Skip to content
Back to Models
Long context

Gemini 3.7 Flash API pricing

Google's Gemini series — very long context windows and strong multimodal input.

Model IDgoogle/gemini-3.7-flash
ChatStreamingTool callingVision

Call summary

Input
$0.75$0.225/ 1M tokens
Cache write $0.225
Cache read $0.225
Output
$3.75$1.125/ 1M tokens
Context Window
1.05M
API Endpoints
1

Gemini 3.7 Flash API endpoints

One API key — this model is served on the endpoint below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Successful response share

98.82%

Success 2xx

98.82%

Rate limit 429

0.16%

Server 5xx

1.02%

TPS

487.51

tokens/s
Avg. time to first token

7.62 s

Avg. latency

10.96 s

Includes production requests ending in 2xx, 429, or 5xx. Average latency uses successful requests; time to first token and TPS require complete observations from successful streams. Public thresholds are 100 requests and 5 independent principals.

Updated Oct 1, 2026, 1:49 AM UTC

More Gemini models on Router One

11 other Gemini models on the same gateway and API key — each page lists endpoints, posted price, and context window.

Gemini 3.7 Flash compared with other models

Spec and price matchups against peer models.

Gemini 3.7 Flash at a glance

Gemini 3.7 Flash (model ID google/gemini-3.7-flash) is a Gemini-series text model on Router One. Posted rate: $0.225 in / $1.125 out per 1M tokens (was $0.75 in / $3.75 out). Router One subscriptions include Gemini 3.7 Flash in the Standard models allowance: Pro 5,000, Max 8,000, Ultra 25,000 requests per 30-day cycle. Usage beyond the allowance bills the wallet at posted rates. Within a plan allowance, a request above 128,000 / 256,000 / 512,000 input tokens counts as 2 / 4 / 8 requests. Context window: 1.05M tokens. Accepts text, image; returns text. Endpoint: POST /v1/chat/completions (OpenAI-compatible).

Gemini 3.7 Flash FAQ

How much does Gemini 3.7 Flash cost on Router One?

Gemini 3.7 Flash is billed per token on Router One: $0.225 in / $1.125 out per 1M tokens, pay as you go.

Is Gemini 3.7 Flash included in Router One subscriptions?

Router One subscriptions include Gemini 3.7 Flash in the Standard models allowance: Pro 5,000, Max 8,000, Ultra 25,000 requests per 30-day cycle. Usage beyond the allowance bills the wallet at posted rates. Within a plan allowance, a request above 128,000 / 256,000 / 512,000 input tokens counts as 2 / 4 / 8 requests.

Which endpoints serve Gemini 3.7 Flash?

Gemini 3.7 Flash is served on POST /v1/chat/completions (OpenAI-compatible). Any Router One API key works there.

What is Gemini 3.7 Flash's context window?

Gemini 3.7 Flash accepts up to 1.05M tokens of context per request on Router One.

Start using Gemini 3.7 Flash

Create an API key and call Gemini 3.7 Flash at $0.225 in / $1.125 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.