deepseek-v4.1-flash API pricing
DeepSeek's models — cost-efficient reasoning and coding.
deepseek-v4.1-flashCall summary
deepseek-v4.1-flash API endpoints
One API key — call this model through any of the 3 endpoints below.
Production reliability
Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.
Not enough production samples yet
This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.
More DeepSeek models on Router One
1 other DeepSeek model on the same gateway and API key — its page lists endpoints, posted price, and context window.
deepseek-v4.1-flash at a glance
deepseek-v4.1-flash is a DeepSeek-series text model on Router One. Posted rate: $0.18 in / $0.72 out per 1M tokens. Context window: 1M tokens. Accepts text, image; returns text. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path), POST /v1/responses (Responses API, the Codex CLI path).
deepseek-v4.1-flash FAQ
How much does deepseek-v4.1-flash cost on Router One?
deepseek-v4.1-flash is billed per token on Router One: $0.18 in / $0.72 out per 1M tokens, pay as you go.
Which endpoints serve deepseek-v4.1-flash?
deepseek-v4.1-flash is served on 3 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path), POST /v1/responses (Responses API, the Codex CLI path). The same Router One API key works on each.
What is deepseek-v4.1-flash's context window?
deepseek-v4.1-flash accepts up to 1M tokens of context per request on Router One.
Start using deepseek-v4.1-flash
Create an API key and call deepseek-v4.1-flash at $0.18 in / $0.72 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.