Claude Sonnet 5 vs Claude Sonnet 4.6
Claude Sonnet 5 and Claude Sonnet 4.6 come to the same total on Router One for a 1M-input + 1M-output mix ($5.40), so price does not decide this pair. Both carry a 1.05M context window. Both answer on /v1/chat/completions and /v1/messages (Claude Code).
Claude Sonnet 5 and Claude Sonnet 4.6 compared on current per-token rates, context window, and capabilities — both callable through one OpenAI-compatible endpoint with per-request cost traces.
Claude Sonnet 5 vs Claude Sonnet 4.6: rates, context window, and capabilities
| Spec | Claude Sonnet 5 | Claude Sonnet 4.6 |
|---|---|---|
| Input / 1M tokens | $0.90 | $0.90 |
| Output / 1M tokens | $4.50 | $4.50 |
| Cached input / 1M tokens | $0.15 | $0.15 |
| Context window | 1.05M | 1.05M |
| Capabilities | Chat, Streaming, Tool calling, Vision | Chat, Streaming, Tool calling, Vision |
| Detail page | Claude Sonnet 5 | Claude Sonnet 4.6 |
What does 1M tokens cost on Claude Sonnet 5 vs Claude Sonnet 4.6?
For a workload of 1M input plus 1M output tokens at current rates, Claude Sonnet 5 and Claude Sonnet 4.6 both come to $5.40. Price does not separate this pair, so choose on context window, capabilities, and the latency your own request traces show; real workloads skew heavily toward input tokens, so compare the input rates by your own ratio.
Switch between Claude Sonnet 5 and Claude Sonnet 4.6 without changing code
Both models are behind the same OpenAI-compatible endpoint, so an A/B test is a one-string change — same key, same code, and every request traced with tokens, cost, and latency in the dashboard:
curl https://api.router.one/v1/chat/completions \
-H "Authorization: Bearer sk-your-router-one-key" \
-H "Content-Type: application/json" \
-d '{"model": "anthropic/claude-sonnet-5", "messages": [{"role": "user", "content": "Hello"}]}'
# Same request, other model — change one string:
# "model": "anthropic/claude-sonnet-4.6"FAQ
Is Claude Sonnet 5 cheaper than Claude Sonnet 4.6?
Input: Claude Sonnet 5 $0.90 vs Claude Sonnet 4.6 $0.90 / 1M tokens; the rates are equal. Output: Claude Sonnet 5 $4.50 vs Claude Sonnet 4.6 $4.50 / 1M tokens; the rates are equal. Total cost depends on the workload's input, output and cache usage. Rates change; the /models page is the live source of truth.
Can I switch between Claude Sonnet 5 and Claude Sonnet 4.6 without changing code?
Yes. Both are served through the same OpenAI-compatible endpoint, so switching is changing the model string in the request — the key, base URL, and request shape stay identical.
Where do these numbers come from?
Specs and prices on this page render from the live Router One catalog — the same data as the /models page — and refresh with it. Pricing methodology is documented on /pricing-methodology. Rates on this page were read from the live catalog on 2026-09-15 (UTC) and refresh within the hour.