Gemini 3.5 Flash Lite Cost Calculator
Gemini 3.5 Flash Lite is Google's low-cost multimodal tier: 1M context, native audio + vision, at $0.30/$2.50 per 1M tokens — 5x cheaper than Gemini 3.5 Flash. Cached input at $0.03/1M.
DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.
Affiliate link — we may earn a commission at no extra cost to you.
Gemini 3.5 Flash Lite pricing breakdown
All pricing dimensions including caching and batch discounts.
| Type | Price (per 1M tokens) |
|---|---|
| Input | $0.3000 |
| Output | $2.5000 |
| Cached input | $0.0300 |
| Batch input | $0.1500 |
| Batch output | $1.2500 |
DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.
Affiliate link — we may earn a commission at no extra cost to you.
Last verified 2026-08-01 · Google official pricing · ⚠️ Spotted a wrong price? Report in 30s →
Gemini 3.5 Flash Lite vs alternatives
Single-call cost (1000 input + 500 output tokens) ranked from cheapest.
| Model | Per call |
|---|---|
Gemini 3.5 Flash Lite Google · this page | $0.001550 |
| DeepSeek V4-Flash DeepSeek | $0.000280 |
| DeepSeek V3.2 DeepSeek | $0.000480 |
| GPT-5.6 Luna OpenAI | $0.000800 |
| GPT-5 mini OpenAI | $0.001250 |
| Mistral Large 3 Mistral | $0.001250 |
When to choose Gemini 3.5 Flash Lite
Gemini 3.5 Flash Lite shines for general-purpose tasks, image and document understanding, high-throughput, low-latency tasks, and audio understanding and generation. Token counts are estimated within ~10-20% margin.