DeepSeek V4-Flash API Pricing & Cost Calculator
DeepSeek V4-Flash pairs a 1M context window with budget pricing: $0.44 input / $1.32 output per 1M tokens. DeepSeek V3.2 remains cheaper at $0.28/$0.40, but caps out at 163K context — V4-Flash is the pick when you need the long window.
Hosted API for 200+ open models — pay-as-you-go, no subscription.
Affiliate link — we may earn a commission at no extra cost to you.
DeepSeek V4-Flash pricing breakdown
All pricing dimensions including caching and batch discounts.
| Type | Price (per 1M tokens) |
|---|---|
| Input | $0.4400 |
| Output | $1.3200 |
Hosted API for 200+ open models — pay-as-you-go, no subscription.
Affiliate link — we may earn a commission at no extra cost to you.
Last verified 2026-08-24 · DeepSeek official pricing · ⚠️ Spotted a wrong price? Report in 30s →
DeepSeek V4-Flash vs alternatives
Single-call cost (1000 input + 500 output tokens) ranked from cheapest.
| Model | Per call |
|---|---|
DeepSeek V4-Flash DeepSeek · this page | $0.001100 |
| DeepSeek V3.2 DeepSeek | $0.000480 |
| GPT-5.6 Luna OpenAI | $0.000800 |
| Gemini 3.1 Flash Lite Google | $0.001000 |
| GPT-5 mini OpenAI | $0.001250 |
| Mistral Large 3 Mistral | $0.001250 |
When to choose DeepSeek V4-Flash
DeepSeek V4-Flash shines for general-purpose tasks, code generation and review, high-throughput, and low-latency tasks. Token counts are estimated within ~10-20% margin.