Skip to content
Router One
Back to Models
Text model

grok-4.7-build-fast API pricing

xAI's Grok series — chat plus image generation.

Model IDgrok-4.7-build-fast
ChatStreamingTool callingVision

Call summary

Input
$4.40$2.20/ 1M tokens
Cache write $2.20
Output
$13.20$6.60/ 1M tokens
Cache read $0.55
Context Window
500K
API Endpoints
2

grok-4.7-build-fast API endpoints

One API key — call this model through any of the 2 endpoints below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model
POST/v1/responsesResponses API · Native for Codex CLI

grok-4.7-build-fast pricing tiers

Pricing is tiered by total input length per request (cache included); once a threshold is crossed, the whole request is billed at that tier. All prices are per 1M tokens.

Tier
Standard≤ 200K
Input
$4.40$2.20Cache write: $2.20
Output
$13.20$6.60Cache read: $0.55
Tier
Long context> 200K
Input
$8.80$4.40Cache write: $4.40
Output
$26.40$13.20Cache read: $1.10

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Not enough production samples yet

This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.

More Grok models on Router One

7 other Grok models on the same gateway and API key — each page lists endpoints, posted price, and context window.

grok-4.7-build-fast at a glance

grok-4.7-build-fast is a Grok-series text model on Router One. request input ≤ 200,000 tokens: Input $2.20 / 1M tokens, Output $6.60 / 1M tokens, Cache write $2.20 / 1M tokens, Cached input $0.55 / 1M tokens; request input > 200,000 tokens: Input $4.40 / 1M tokens, Output $13.20 / 1M tokens, Cache write $4.40 / 1M tokens, Cached input $1.10 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold. Context window: 500K tokens. Accepts text, image; returns text. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path).

grok-4.7-build-fast FAQ

How much does grok-4.7-build-fast cost on Router One?

request input ≤ 200,000 tokens: Input $2.20 / 1M tokens, Output $6.60 / 1M tokens, Cache write $2.20 / 1M tokens, Cached input $0.55 / 1M tokens; request input > 200,000 tokens: Input $4.40 / 1M tokens, Output $13.20 / 1M tokens, Cache write $4.40 / 1M tokens, Cached input $1.10 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold.

Which endpoints serve grok-4.7-build-fast?

grok-4.7-build-fast is served on 2 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path). The same Router One API key works on each.

What is grok-4.7-build-fast's context window?

grok-4.7-build-fast accepts up to 500K tokens of context per request on Router One.

Start using grok-4.7-build-fast

Create an API key and call grok-4.7-build-fast at request input ≤ 200,000 tokens: $2.20 in / $6.60 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.