Back to Models
Text model
claude-opus-4.8
Anthropic's Claude series — long context, dependable tool calling, and the native models behind Claude Code.
Model ID
anthropic/claude-opus-4.8ChatStreamingTool callingVision
Call summary
Input
$7.50$2.25/ 1M tokens
Cache write $4.69
Output
$37.50$11.25/ 1M tokens
Cache read $0.38
Context Window
200K
Max output
8K
API Endpoints
2
Available Configurations
3
Production reliability
Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.
Successful response share
99.00%
Success 2xx
99.00%
Rate limit 429
0.00%
Server 5xx
1.00%
TPS
71.43
tokens/sAvg. time to first token
15.42 s
Avg. latency
42.48 s
Includes production requests ending in 2xx, 429, or 5xx. Average latency uses successful requests; time to first token and TPS require complete observations from successful streams. Public thresholds are 100 requests and 5 independent principals.
Updated Jul 23, 2026, 4:34 PM UTC
API Endpoints
One API key — call this model through any of the endpoints below.
Model comparisons
Spec and price matchups against peer models.