Skip to content
Back to Models
Long context

Claude Haiku 5.5 API pricing

Anthropic's Claude series — long context, dependable tool calling, and the native models behind Claude Code.

Model IDanthropic/claude-haiku-5.5
ChatStreamingTool callingVision

Call summary

Input
$0.10$0.06/ 1M tokens
Cache write $0.075
Cache read $0.006
Output
$0.50$0.30/ 1M tokens

Applies to request input ≤ 100,000 tokens · See pricing tiers

Context Window
1.05M
API Endpoints
2

Claude Haiku 5.5 API endpoints

One API key — call this model through any of the 2 endpoints below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model
POST/v1/messagesAnthropic-native · Native for Claude Code

Claude Haiku 5.5 pricing tiers

Pricing is tiered by total input length per request (cache included); once a threshold is crossed, the whole request is billed at that tier. All prices are per 1M tokens.

Tier
Standard≤ 100K
Input
$0.10$0.06Cache write: $0.075Cache read: $0.006
Output
$0.50$0.30
Tier
Long context> 100K
Input
$0.50$0.30Cache write: $0.375Cache read: $0.03
Output
$2.50$1.50

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Not enough production samples yet

This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.

More Claude models on Router One

28 other Claude models on the same gateway and API key — each page lists endpoints, posted price, and context window.

16 channel versions

Claude Haiku 5.5 at a glance

Claude Haiku 5.5 (model ID anthropic/claude-haiku-5.5) is a Claude-series text model on Router One. Claude Haiku 5.5 is priced in whole-request tiers. Request input ≤ 100,000 tokens: Input $0.06 / 1M tokens, Output $0.30 / 1M tokens, Cache write $0.075 / 1M tokens, Cached input $0.006 / 1M tokens; request input > 100,000 tokens: Input $0.30 / 1M tokens, Output $1.50 / 1M tokens, Cache write $0.375 / 1M tokens, Cached input $0.03 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold. No Router One subscription plan includes Claude Haiku 5.5 today; every call bills the wallet at its posted rate. Context window: 1.05M tokens. Accepts text, image; returns text. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path).

Claude Haiku 5.5 FAQ

How much does Claude Haiku 5.5 cost on Router One?

Claude Haiku 5.5 is priced in whole-request tiers. Request input ≤ 100,000 tokens: Input $0.06 / 1M tokens, Output $0.30 / 1M tokens, Cache write $0.075 / 1M tokens, Cached input $0.006 / 1M tokens; request input > 100,000 tokens: Input $0.30 / 1M tokens, Output $1.50 / 1M tokens, Cache write $0.375 / 1M tokens, Cached input $0.03 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold.

Is Claude Haiku 5.5 included in Router One subscriptions?

No Router One subscription plan includes Claude Haiku 5.5 today; every call bills the wallet at its posted rate.

Which endpoints serve Claude Haiku 5.5?

Claude Haiku 5.5 is served on 2 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path). The same Router One API key works on each.

What is Claude Haiku 5.5's context window?

Claude Haiku 5.5 accepts up to 1.05M tokens of context per request on Router One.

Start using Claude Haiku 5.5

Create an API key and call Claude Haiku 5.5 at $0.06 in / $0.30 out per 1M tokens (request input ≤ 100,000 tokens) on Router One — pay as you go, with per-request cost and latency visibility.