Skip to content

Gemini 3.7 Flash vs Gemini 3.6 Flash

Gemini 3.7 Flash is the lower-priced of the two on Router One — about 57% less on a 1M-input + 1M-output mix. Both carry a 1.05M context window. Both answer on /v1/chat/completions.

Gemini 3.7 Flash and Gemini 3.6 Flash compared on current per-token rates, context window, and capabilities — both callable through one OpenAI-compatible endpoint with per-request cost traces.

Gemini 3.7 Flash vs Gemini 3.6 Flash: rates, context window, and capabilities

SpecGemini 3.7 FlashGemini 3.6 Flash
Input / 1M tokens$0.225$0.45
Output / 1M tokens$1.125$2.70
Cached input / 1M tokens$0.225$0.45
Context window1.05M1.05M
CapabilitiesChat, Streaming, Tool calling, VisionChat, Streaming, Tool calling, Vision
Subscription plansStandard models — Pro, Max, UltraStandard models — Pro, Max, Ultra
Detail pageGemini 3.7 FlashGemini 3.6 Flash

What does 1M tokens cost on Gemini 3.7 Flash vs Gemini 3.6 Flash?

For a workload of 1M input plus 1M output tokens at current rates: Gemini 3.7 Flash comes to $1.35, Gemini 3.6 Flash comes to $3.15 — Gemini 3.7 Flash is about 57% cheaper on this mix. Real workloads skew heavily toward input tokens, so weigh the input rate by your own ratio; the cached-input row above is the posted catalog rate for that model.

Switch between Gemini 3.7 Flash and Gemini 3.6 Flash without changing code

Both models are behind the same OpenAI-compatible endpoint, so an A/B test is a one-string change — same key, same code, and every request traced with tokens, cost, and latency in the dashboard:

compare.sh
curl https://api.router.one/v1/chat/completions \
  -H "Authorization: Bearer sk-your-router-one-key" \
  -H "Content-Type: application/json" \
  -d '{"model": "google/gemini-3.7-flash", "messages": [{"role": "user", "content": "Hello"}]}'

# Same request, other model — change one string:
#   "model": "google/gemini-3.6-flash"

FAQ

Is Gemini 3.7 Flash cheaper than Gemini 3.6 Flash?

Input: Gemini 3.7 Flash $0.225 vs Gemini 3.6 Flash $0.45 / 1M tokens; Gemini 3.7 Flash has the lower rate. Output: Gemini 3.7 Flash $1.125 vs Gemini 3.6 Flash $2.70 / 1M tokens; Gemini 3.7 Flash has the lower rate. Total cost depends on the workload's input, output and cache usage. Rates change; the /models page is the live source of truth.

Is Gemini 3.7 Flash or Gemini 3.6 Flash included in a Router One subscription?

Gemini 3.7 Flash and Gemini 3.6 Flash both count against the Standard models allowance on Pro, Max and Ultra: Pro 5,000, Max 8,000 and Ultra 25,000 requests per 30-day cycle, one pool shared by every model in that tier. Requests to Gemini 3.7 Flash or Gemini 3.6 Flash above the model's long-context threshold count as more than one quota request. Requests beyond a plan's allowance bill the wallet per token.

Can I switch between Gemini 3.7 Flash and Gemini 3.6 Flash without changing code?

Yes. Both are served through the same OpenAI-compatible endpoint, so switching is changing the model string in the request — the key, base URL, and request shape stay identical.

Where do these numbers come from?

Specs and prices on this page render from the live Router One catalog — the same data as the /models page — and refresh with it. Pricing methodology is documented on /pricing-methodology. Rates on this page were read from the live catalog on 2026-10-06 (UTC) and refresh within the hour.

More comparisons