Home/DeepSeek/DeepSeek V4-Flash
DeepSeek pricing

DeepSeek V4-Flash API Pricing & Cost Calculator

DeepSeek V4-Flash pairs a 1M context window with budget pricing: $0.44 input / $1.32 output per 1M tokens. DeepSeek V3.2 remains cheaper at $0.28/$0.40, but caps out at 163K context — V4-Flash is the pick when you need the long window.

Input
$0.44
per 1M tokens
Output
$1.32
per 1M tokens
Context window
1000K
tokens
Released
2026-06
Cutoff
≈ Estimated tokenizer·$0.44 in·$1.32 out (per 1M)
Quick start with a use case
Total cost per call$0.001100
Input$0.000440
Output$0.000660
Cost comparison
Standard
$0.001100
With Caching
Not supported
With Batch
Not supported
Run DeepSeek V4-Flash on Novita AI

Hosted API for 200+ open models — pay-as-you-go, no subscription.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Detailed pricing

DeepSeek V4-Flash pricing breakdown

All pricing dimensions including caching and batch discounts.

TypePrice (per 1M tokens)
Input$0.4400
Output$1.3200
Run DeepSeek V4-Flash on Novita AI

Hosted API for 200+ open models — pay-as-you-go, no subscription.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Last verified 2026-08-24 · DeepSeek official pricing · ⚠️ Spotted a wrong price? Report in 30s →

How it compares

DeepSeek V4-Flash vs alternatives

Single-call cost (1000 input + 500 output tokens) ranked from cheapest.

ModelPer call
DeepSeek V4-Flash
DeepSeek · this page
$0.001100
DeepSeek V3.2
DeepSeek
$0.000480
GPT-5.6 Luna
OpenAI
$0.000800
Gemini 3.1 Flash Lite
Google
$0.001000
GPT-5 mini
OpenAI
$0.001250
Mistral Large 3
Mistral
$0.001250
Recommended use

When to choose DeepSeek V4-Flash

DeepSeek V4-Flash shines for general-purpose tasks, code generation and review, high-throughput, and low-latency tasks. Token counts are estimated within ~10-20% margin.

Context window of 1000K tokens handles entire codebases or book-length documents.
Prompt caching available — significant savings for repeated system prompts.
Tool / function calling supported.
FAQ

Frequently asked questions

DeepSeek V4-Flash costs $0.44 per 1M input tokens and $1.32 per 1M output tokens. A typical chat call (1000 input + 500 output tokens) costs approximately $0.0011. Use the calculator above to estimate your specific use case.
DeepSeek V4-Flash supports a 1000K token context window with a max output of 393,216 tokens.
Token counts for DeepSeek V4-Flash are estimated from character ratios (~10-20% margin). Cost calculations use prices verified on 2026-08-24. For final billing accuracy, always verify with DeepSeek's usage dashboard.
DeepSeek V4-Flash does not currently support fine-tuning. Consider another model in the DeepSeek family if fine-tuning is required.