Skip to content
Router One
Back to Models
Text model

grok-4.5 API pricing

xAI's Grok series — chat plus image generation.

Model IDgrok-4.5
ChatStreamingTool callingVision

Call summary

Input
$0.80/ 1M tokens
Cache write $0.80
Output
$2.40/ 1M tokens
Cache read $0.20
Context Window
500K
API Endpoints
1
Available Configurations
2

grok-4.5 API endpoints

One API key — this model is served on the endpoint below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Successful response share

99.65%

Success 2xx

99.65%

Rate limit 429

0.00%

Server 5xx

0.35%

TPS

50.32

tokens/s
Avg. time to first token

5.21 s

Avg. latency

29.23 s

Includes production requests ending in 2xx, 429, or 5xx. Average latency uses successful requests; time to first token and TPS require complete observations from successful streams. Public thresholds are 100 requests and 5 independent principals.

Updated Sep 9, 2026, 8:59 PM UTC

More Grok models on Router One

5 other Grok models on the same gateway and API key — each page lists endpoints, posted price, and context window.

grok-4.5 compared with other models

Spec and price matchups against peer models.

grok-4.5 at a glance

grok-4.5 is a Grok-series text model on Router One. Posted rate: $0.80 in / $2.40 out per 1M tokens. Context window: 500K tokens. Endpoint: POST /v1/chat/completions (OpenAI-compatible).

grok-4.5 FAQ

How much does grok-4.5 cost on Router One?

grok-4.5 is billed per token on Router One: $0.80 in / $2.40 out per 1M tokens, pay as you go.

Which endpoints serve grok-4.5?

grok-4.5 is served on POST /v1/chat/completions (OpenAI-compatible). Any Router One API key works there.

What is grok-4.5's context window?

grok-4.5 accepts up to 500K tokens of context per request on Router One.

Start using grok-4.5

Create an API key and call grok-4.5 at $0.80 in / $2.40 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.