gemini-3.1-flash-lite API pricing
Google's Gemini series — very long context windows and strong multimodal input.
google/gemini-3.1-flash-liteCall summary
gemini-3.1-flash-lite API endpoints
One API key — this model is served on the endpoint below.
Production reliability
Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.
Not enough production samples yet
This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.
More Gemini models on Router One
10 other Gemini models on the same gateway and API key — each page lists endpoints, posted price, and context window.
gemini-3.1-flash-lite at a glance
gemini-3.1-flash-lite is a Gemini-series text model on Router One. gemini-3.1-flash-lite is the general-availability id; the preview id gemini-3.1-flash-lite-preview is listed separately at its own rate. Posted rate: $0.25 in / $1.50 out per 1M tokens. Context window: 1.05M tokens. Accepts text, image; returns text. Endpoint: POST /v1/chat/completions (OpenAI-compatible).
gemini-3.1-flash-lite FAQ
How much does gemini-3.1-flash-lite cost on Router One?
gemini-3.1-flash-lite is billed per token on Router One: $0.25 in / $1.50 out per 1M tokens, pay as you go.
Which endpoints serve gemini-3.1-flash-lite?
gemini-3.1-flash-lite is served on POST /v1/chat/completions (OpenAI-compatible). Any Router One API key works there.
What is gemini-3.1-flash-lite's context window?
gemini-3.1-flash-lite accepts up to 1.05M tokens of context per request on Router One.
Start using gemini-3.1-flash-lite
Create an API key and call gemini-3.1-flash-lite at $0.25 in / $1.50 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.