Skip to content
Router One
Back to Models
Long context

gemini-3.1-flash-lite API pricing

Google's Gemini series — very long context windows and strong multimodal input.

Model IDgoogle/gemini-3.1-flash-lite

Call summary

Input
$0.25/ 1M tokens
Cache write $0.25
Output
$1.50/ 1M tokens
Cache read $0.25
Context Window
1.05M
API Endpoints
1
Available Configurations
1

gemini-3.1-flash-lite API endpoints

One API key — this model is served on the endpoint below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Not enough production samples yet

This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.

More Gemini models on Router One

10 other Gemini models on the same gateway and API key — each page lists endpoints, posted price, and context window.

gemini-3.1-flash-lite at a glance

gemini-3.1-flash-lite is a Gemini-series text model on Router One. gemini-3.1-flash-lite is the general-availability id; the preview id gemini-3.1-flash-lite-preview is listed separately at its own rate. Posted rate: $0.25 in / $1.50 out per 1M tokens. Context window: 1.05M tokens. Accepts text, image; returns text. Endpoint: POST /v1/chat/completions (OpenAI-compatible).

gemini-3.1-flash-lite FAQ

How much does gemini-3.1-flash-lite cost on Router One?

gemini-3.1-flash-lite is billed per token on Router One: $0.25 in / $1.50 out per 1M tokens, pay as you go.

Which endpoints serve gemini-3.1-flash-lite?

gemini-3.1-flash-lite is served on POST /v1/chat/completions (OpenAI-compatible). Any Router One API key works there.

What is gemini-3.1-flash-lite's context window?

gemini-3.1-flash-lite accepts up to 1.05M tokens of context per request on Router One.

Start using gemini-3.1-flash-lite

Create an API key and call gemini-3.1-flash-lite at $0.25 in / $1.50 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.