Deepseek V4 Flash costs $0.140/1M input tokens and $0.280/1M output tokens as of 2026-06-27. Deepseek publishes this model with a 1M tokens context window. o10 lists Deepseek V4 Flash on DeepSeek at routing band C2.
Deepseek V4 Flash pricing — input/output per 1M tokens
AnswersSelf-contained
How much does Deepseek V4 Flash cost per million tokens?
Deepseek V4 Flash costs $0.140/1M input tokens and $0.280/1M output tokens in the gateway catalog snapshot (2026-06-27). Endpoint pricing may differ by fulfillment host.
Sourced from gateway catalog snapshot.
Where can you run Deepseek V4 Flash?
Deepseek V4 Flash is configured on DeepSeek. Each host maps to an upstream model ID listed in the fulfillment routes table. This model is routable through o10 eval-gated routing.
What is Deepseek V4 Flash's context window?
Deepseek V4 Flash supports a 1M tokens context window in the gateway catalog snapshot. Context limits may vary by fulfillment host endpoint.
Deepseek V4 Flash costs $0.140/1M input tokens and $0.280/1M output tokens in the gateway snapshot (2026-06-27). Verify against your provider's published pricing.
What is Deepseek V4 Flash's context window?
Deepseek V4 Flash supports 1M tokens in the gateway catalog snapshot. Host-specific endpoints may differ.
Which hosts serve Deepseek V4 Flash?
Deepseek V4 Flash (deepseek-v4-flash) is configured on: DeepSeek. Upstream model IDs are listed in the fulfillment routes section.
Is Deepseek V4 Flash routable through o10?
Yes. Deepseek V4 Flash is routable at band C2. Send model: "deepseek-v4-flash" or use an o10 smart routing alias.
How do I call Deepseek V4 Flash via o10?
Use o10 slug "deepseek-v4-flash" or gateway ID "deepseek/deepseek-v4-flash" depending on your integration. Gateway-compatible endpoints accept OpenAI-style /v1 requests.
How does o10 route this model?
o10 routes Deepseek V4 Flash only when per-use-case evals clear at your quality floor. Band C2 indicates cost tier. See /kyi for governance and /models/routing for smart aliases.