Skip to content
Router One
Back to Models
Long context

gemini-3-flash API pricing

Google's Gemini series — very long context windows and strong multimodal input.

Model IDgoogle/gemini-3-flash
ChatStreamingTool callingVision

Call summary

Input
$0.11/ 1M tokens
Cache write $0.11
Output
$0.68/ 1M tokens
Cache read $0.01
Context Window
1.05M
API Endpoints
1
Available Configurations
1

gemini-3-flash API endpoints

One API key — this model is served on the endpoint below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Not enough production samples yet

This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.

More Gemini models on Router One

8 other Gemini models on the same gateway and API key — each page lists endpoints, posted price, and context window.

gemini-3-flash at a glance

gemini-3-flash is a Gemini-series text model on Router One. Posted rate: $0.1125 in / $0.675 out per 1M tokens. Context window: 1.05M tokens. Accepts text, image; returns text. Endpoint: POST /v1/chat/completions (OpenAI-compatible).

gemini-3-flash FAQ

How much does gemini-3-flash cost on Router One?

gemini-3-flash is billed per token on Router One: $0.1125 in / $0.675 out per 1M tokens, pay as you go.

Which endpoints serve gemini-3-flash?

gemini-3-flash is served on POST /v1/chat/completions (OpenAI-compatible). Any Router One API key works there.

What is gemini-3-flash's context window?

gemini-3-flash accepts up to 1.05M tokens of context per request on Router One.

Start using gemini-3-flash

Create an API key and call gemini-3-flash at $0.1125 in / $0.675 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.