Skip to content
Router One
Back to Models
Long context

gpt-6-astra · Azure API pricing

OpenAI's GPT series — broad general capability, wide ecosystem, and the native models behind Codex CLI.

Model IDazure/gpt-6-astra
ChatStreamingTool callingVision

Call summary

Input
$10.00$7.00/ 1M tokens
Cache write $8.75
Output
$50.00$35.00/ 1M tokens
Cache read $0.70
Context Window
1.05M
API Endpoints
2

gpt-6-astra · Azure API endpoints

One API key — call this model through any of the 2 endpoints below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model
POST/v1/responsesResponses API · Native for Codex CLI

gpt-6-astra · Azure pricing tiers

Pricing is tiered by total input length per request (cache included); once a threshold is crossed, the whole request is billed at that tier. All prices are per 1M tokens.

Tier
Standard≤ 272K
Input
$10.00$7.00Cache write: $8.75
Output
$50.00$35.00Cache read: $0.70
Tier
Long context> 272K
Input
$20.00$14.00Cache write: $17.50
Output
$75.00$52.50Cache read: $1.40

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Not enough production samples yet

This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.

More GPT models on Router One

12 other GPT models on the same gateway and API key — each page lists endpoints, posted price, and context window.

gpt-6-astra · Azure at a glance

gpt-6-astra · Azure is a GPT-series text model on Router One. request input ≤ 272,000 tokens: Input $7.00 / 1M tokens, Output $35.00 / 1M tokens, Cache write $8.75 / 1M tokens, Cached input $0.70 / 1M tokens; request input > 272,000 tokens: Input $14.00 / 1M tokens, Output $52.50 / 1M tokens, Cache write $17.50 / 1M tokens, Cached input $1.40 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold. Context window: 1.05M tokens. Accepts text, image; returns text. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path).

gpt-6-astra · Azure FAQ

How much does gpt-6-astra · Azure cost on Router One?

request input ≤ 272,000 tokens: Input $7.00 / 1M tokens, Output $35.00 / 1M tokens, Cache write $8.75 / 1M tokens, Cached input $0.70 / 1M tokens; request input > 272,000 tokens: Input $14.00 / 1M tokens, Output $52.50 / 1M tokens, Cache write $17.50 / 1M tokens, Cached input $1.40 / 1M tokens. The total input tokens in each request select the tier; its rates apply to the whole request, not only tokens above the threshold.

Which endpoints serve gpt-6-astra · Azure?

gpt-6-astra · Azure is served on 2 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/responses (Responses API, the Codex CLI path). The same Router One API key works on each.

What is gpt-6-astra · Azure's context window?

gpt-6-astra · Azure accepts up to 1.05M tokens of context per request on Router One.

Start using gpt-6-astra · Azure

Create an API key and call gpt-6-astra · Azure at request input ≤ 272,000 tokens: $7.00 in / $35.00 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.