gpt-5.3-codex-spark API pricing
OpenAI's GPT series — broad general capability, wide ecosystem, and the native models behind Codex CLI.
openai/gpt-5.3-codex-sparkCall summary
gpt-5.3-codex-spark API endpoints
One API key — call this model through any of the 2 endpoints below.
Production reliability
Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.
99.78%
99.78%
0.00%
0.22%
2,389.81
tokens/s7.45 s
9.47 s
Includes production requests ending in 2xx, 429, or 5xx. Average latency uses successful requests; time to first token and TPS require complete observations from successful streams. Public thresholds are 100 requests and 5 independent principals.
Updated Sep 9, 2026, 8:59 PM UTC
More GPT models on Router One
6 other GPT models on the same gateway and API key — each page lists endpoints, posted price, and context window.
gpt-5.3-codex-spark compared with other models
Spec and price matchups against peer models.
gpt-5.3-codex-spark at a glance
gpt-5.3-codex-spark is a GPT-series text model on Router One. Posted rate: $0.4314 in / $3.4521 out per 1M tokens (was $1.438 in / $11.507 out). Context window: 400K tokens. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path).
gpt-5.3-codex-spark FAQ
How much does gpt-5.3-codex-spark cost on Router One?
gpt-5.3-codex-spark is billed per token on Router One: $0.4314 in / $3.4521 out per 1M tokens, pay as you go.
Which endpoints serve gpt-5.3-codex-spark?
gpt-5.3-codex-spark is served on 2 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path). The same Router One API key works on each.
What is gpt-5.3-codex-spark's context window?
gpt-5.3-codex-spark accepts up to 400K tokens of context per request on Router One.
Start using gpt-5.3-codex-spark
Create an API key and call gpt-5.3-codex-spark at $0.4314 in / $3.4521 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.