OpenAIComprobando
gpt-5.3-codex
openai/gpt-5.3-codex
chat completionsContext: 400KHerramientasJSON
Cargando precio…
Official price$1.75 / $14 / 1 M de tokens
Conversation
Send a message to run gpt-5.3-codex with your account.
Settings
Last response
Usage and charge appear here after a reply.
About this model
gpt-5.3-codex is served through the Hermes AI OpenAI-compatible chat completions endpoint. Send the standard messages array; streaming, function tools, JSON output and reasoning settings pass through unchanged. You pay per token from prepaid credits, at the rates on the Pricing tab.
| Property | Valor |
|---|---|
| Model id | openai/gpt-5.3-codex |
| Vendor | OpenAI |
| Context window | 400,000 tokens |
| Capacidades | Herramientas, JSON, Streaming |
| Endpoint | https://api.hermes-ai.net/api/v1/chat/completions |
| Precio de Hermes AI | Cargando precio… |
Request
POST https://api.hermes-ai.net/api/v1/chat/completions · API key with the chat:create scope, or your signed-in session.
| Campo | Tipo | Notas |
|---|---|---|
| model | string | openai/gpt-5.3-codex |
| messages | array | OpenAI chat messages: system, user, assistant and tool roles; text and image_url parts. |
| stream | boolean | Server-sent events with a final usage chunk. |
| max_tokens | integer | Output cap; also sizes the credit reservation. |
| tools / tool_choice | array / object | Function tools only. |
| response_format | object | JSON mode and JSON schema where the model supports it. |
| temperature, top_p, stop, seed, reasoning_effort | Passed through unchanged. |
Respuesta
| Campo | Notas |
|---|---|
| choices[].message / delta | The OpenAI chat completion shape. |
| usage | prompt_tokens, completion_tokens and prompt_tokens_details.cached_tokens as reported by the model. |
| hermes.charged_micros | Exact charge in USD micro-units (JSON responses). Streams expose it on GET /api/v1/chat/requests. |
| x-hermes-request | Request id header, also used in the billing ledger. |
| error.code | Documented Hermes error code; branch on it, not on the message. |
Token prices
USD per 1M tokens. The tier is chosen by the prompt size of each request.
Cargando precio…
How token billing works
- When a request is accepted, Hermes reserves credits for the estimated prompt plus max_tokens of output (8,192 when you set none). Set max_tokens to keep the reservation small.
- When the response finishes, the reservation is replaced by the exact charge from the usage the model reports: prompt tokens at the input rate, cached prompt tokens at the cache-read rate, completion tokens at the output rate. The remainder is released immediately.
- If the model rejects the request nothing is charged. If a stream is interrupted after tokens were generated, the generated text is estimated and charged.
Ejemplos de código
Any OpenAI SDK works: set the base URL and your Hermes key.
shell
export HERMES_API_KEY=your_api_keycurl https://api.hermes-ai.net/api/v1/chat/completions \
-H "Authorization: Bearer $HERMES_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "openai/gpt-5.3-codex",
"messages": [{"role": "user", "content": "Say hello in one sentence."}],
"max_tokens": 200
}'