claude-sonnet-4.6 API pricing
Anthropic's Claude series — long context, dependable tool calling, and the native models behind Claude Code.
anthropic/claude-sonnet-4.6Call summary
claude-sonnet-4.6 API endpoints
One API key — call this model through any of the 2 endpoints below.
Production reliability
Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.
99.35%
99.35%
0.00%
0.65%
61.8
tokens/s15.82 s
45.73 s
Includes production requests ending in 2xx, 429, or 5xx. Average latency uses successful requests; time to first token and TPS require complete observations from successful streams. Public thresholds are 100 requests and 5 independent principals.
Updated Sep 9, 2026, 8:59 PM UTC
More Claude models on Router One
10 other Claude models on the same gateway and API key — each page lists endpoints, posted price, and context window.
claude-sonnet-4.6 compared with other models
Spec and price matchups against peer models.
claude-sonnet-4.6 at a glance
claude-sonnet-4.6 is a Claude-series text model on Router One. Posted rate: $0.90 in / $4.50 out per 1M tokens (was $3.00 in / $15.00 out). Context window: 1.05M tokens. Accepts text, image; returns text. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path).
claude-sonnet-4.6 FAQ
How much does claude-sonnet-4.6 cost on Router One?
claude-sonnet-4.6 is billed per token on Router One: $0.90 in / $4.50 out per 1M tokens, pay as you go.
Which endpoints serve claude-sonnet-4.6?
claude-sonnet-4.6 is served on 2 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path). The same Router One API key works on each.
What is claude-sonnet-4.6's context window?
claude-sonnet-4.6 accepts up to 1.05M tokens of context per request on Router One.
Start using claude-sonnet-4.6
Create an API key and call claude-sonnet-4.6 at $0.90 in / $4.50 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.