Skip to content
Router One
Back to Models
Long context

deepseek-v4.1-flash API pricing

DeepSeek's models — cost-efficient reasoning and coding.

Model IDdeepseek-v4.1-flash
ChatStreamingTool callingVision

Call summary

Input
$0.18/ 1M tokens
Cache write $0.18
Output
$0.72/ 1M tokens
Cache read $0.0036
Context Window
1M
API Endpoints
3
Available Configurations
1

deepseek-v4.1-flash API endpoints

One API key — call this model through any of the 3 endpoints below.

POST/v1/chat/completionsOpenAI-compatible · Works for every model
POST/v1/messagesAnthropic-native · Native for Claude Code
POST/v1/responsesResponses API · Native for Codex CLI

Production reliability

Privacy-safe aggregates from real Router One model calls, so you can assess stability and response latency before integrating.

Not enough production samples yet

This window needs at least 100 eligible requests across 5 independent principals. Metrics appear automatically after that threshold is met.

More DeepSeek models on Router One

1 other DeepSeek model on the same gateway and API key — its page lists endpoints, posted price, and context window.

deepseek-v4.1-flash at a glance

deepseek-v4.1-flash is a DeepSeek-series text model on Router One. Posted rate: $0.18 in / $0.72 out per 1M tokens. Context window: 1M tokens. Accepts text, image; returns text. Endpoints: POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path), POST /v1/responses (Responses API, the Codex CLI path).

deepseek-v4.1-flash FAQ

How much does deepseek-v4.1-flash cost on Router One?

deepseek-v4.1-flash is billed per token on Router One: $0.18 in / $0.72 out per 1M tokens, pay as you go.

Which endpoints serve deepseek-v4.1-flash?

deepseek-v4.1-flash is served on 3 endpoints — POST /v1/chat/completions (OpenAI-compatible), POST /v1/messages (Anthropic-native, the Claude Code path), POST /v1/responses (Responses API, the Codex CLI path). The same Router One API key works on each.

What is deepseek-v4.1-flash's context window?

deepseek-v4.1-flash accepts up to 1M tokens of context per request on Router One.

Start using deepseek-v4.1-flash

Create an API key and call deepseek-v4.1-flash at $0.18 in / $0.72 out per 1M tokens on Router One — pay as you go, with per-request cost and latency visibility.