gpt-5.6-sol API pricing
OpenAI's GPT series — broad general capability, wide ecosystem, and the native models behind Codex CLI.
openai/gpt-5.6-solCall summary
gpt-5.6-sol API endpoints
One API key — call this model through any of the 2 endpoints below.
gpt-5.6-sol pricing tiers
Pricing is tiered by total input length per request (cache included); once a threshold is crossed, the whole request is billed at that tier. All prices are per 1M tokens.
| Tier | Input | Output |
|---|---|---|
Standard≤ 272K | $5.00$0.50Cache write: $0.50 | $40.00$4.00Cache read: $0.25 |
Long context> 272K | $10.00$1.00Cache write: $6.25 | $45.00$4.50Cache read: $0.50 |
Production reliability
Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.
94.82%
94.82%
1.16%
4.02%
49.12
tokens/s16.57 s
29.82 s
Includes production requests ending in 2xx, 429, or 5xx. Average latency uses successful requests; time to first token and TPS require complete observations from successful streams. Public thresholds are 100 requests and 5 independent principals.
Updated Sep 9, 2026, 8:59 PM UTC
More GPT models on Router One
6 other GPT models on the same gateway and API key — each page lists endpoints, posted price, and context window.
gpt-5.6-sol compared with other models
Spec and price matchups against peer models.
gpt-5.6-sol at a glance
gpt-5.6-sol is a GPT-series text model on Router One. request input ≤ 272,000 tokens: Input $0.50 / 1M tokens, Output $4.00 / 1M tokens, Cache write $0.50 / 1M tokens, Cached input $0.25 / 1M tokens; request input > 272,000 tokens: Input $1.00 / 1M tokens, Output $4.50 / 1M tokens, Cache write $6.25 / 1M tokens, Cached input $0.50 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold. Context window: 1.05M tokens. Accepts text, image; returns text. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path).
gpt-5.6-sol FAQ
How much does gpt-5.6-sol cost on Router One?
request input ≤ 272,000 tokens: Input $0.50 / 1M tokens, Output $4.00 / 1M tokens, Cache write $0.50 / 1M tokens, Cached input $0.25 / 1M tokens; request input > 272,000 tokens: Input $1.00 / 1M tokens, Output $4.50 / 1M tokens, Cache write $6.25 / 1M tokens, Cached input $0.50 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold.
Which endpoints serve gpt-5.6-sol?
gpt-5.6-sol is served on 2 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path). The same Router One API key works on each.
What is gpt-5.6-sol's context window?
gpt-5.6-sol accepts up to 1.05M tokens of context per request on Router One.
Start using gpt-5.6-sol
Create an API key and call gpt-5.6-sol at request input ≤ 272,000 tokens: $0.50 in / $4.00 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.