Gemini 3.1 Flash Lite API Pricing & Cost Calculator
Gemini 3.1 Flash Lite runs at $0.25 input / $1.50 output per 1M tokens with a 1049K context window — cheaper than Gemini 3.5 Flash Lite ($0.30/$2.50) but a generation behind it. A 1,000-in / 500-out call costs $0.00100. Google lists a retirement date of 2027-05-07, so treat it as a cost floor rather than a long-term choice.
DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.
Affiliate link — we may earn a commission at no extra cost to you.
Gemini 3.1 Flash Lite pricing breakdown
All pricing dimensions including caching and batch discounts.
| Type | Price (per 1M tokens) |
|---|---|
| Input | $0.2500 |
| Output | $1.5000 |
| Cached input | $0.0250 |
| Batch input | $0.1250 |
| Batch output | $0.7500 |
DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.
Affiliate link — we may earn a commission at no extra cost to you.
Last verified 2026-08-24 · Google official pricing · ⚠️ Spotted a wrong price? Report in 30s →
Gemini 3.1 Flash Lite vs alternatives
Single-call cost (1000 input + 500 output tokens) ranked from cheapest.
| Model | Per call |
|---|---|
Gemini 3.1 Flash Lite Google · this page | $0.001000 |
| DeepSeek V3.2 DeepSeek | $0.000480 |
| GPT-5.6 Luna OpenAI | $0.000800 |
| DeepSeek V4-Flash DeepSeek | $0.001100 |
| GPT-5 mini OpenAI | $0.001250 |
| Mistral Large 3 Mistral | $0.001250 |
When to choose Gemini 3.1 Flash Lite
Gemini 3.1 Flash Lite shines for high-throughput, low-latency tasks, general-purpose tasks, long documents and large codebases, and image and document understanding. Token counts are estimated within ~10-20% margin.