Gemini 3.5 Flash API Pricing & Cost Calculator
Gemini 3.5 Flash is Google's newest Flash generation at $1.50/$9 per 1M tokens — 3x the price of Gemini 3 Flash but with substantially stronger reasoning. 1M context, native audio + vision.
Pricing below is still current. Worth factoring the retirement date into anything you are building on it long-term.
Still available from Google: Gemini 3.1 Pro, Gemini 3.6 Flash, Gemini 3.7 Flash
DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.
Affiliate link — we may earn a commission at no extra cost to you.
Gemini 3.5 Flash pricing breakdown
All pricing dimensions including caching and batch discounts.
| Type | Price (per 1M tokens) |
|---|---|
| Input | $1.5000 |
| Output | $9.0000 |
| Cached input | $0.1500 |
| Batch input | $0.7500 |
| Batch output | $4.5000 |
DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.
Affiliate link — we may earn a commission at no extra cost to you.
Last verified 2026-08-01 · Google official pricing · ⚠️ Spotted a wrong price? Report in 30s →
Gemini 3.5 Flash vs alternatives
Single-call cost (1000 input + 500 output tokens) ranked from cheapest.
| Model | Per call |
|---|---|
Gemini 3.5 Flash Google · this page | $0.006000 |
| DeepSeek V3.2 DeepSeek | $0.000480 |
| GPT-5.6 Luna OpenAI | $0.000800 |
| Gemini 3.1 Flash Lite Google | $0.001000 |
| DeepSeek V4-Flash DeepSeek | $0.001100 |
| GPT-5 mini OpenAI | $0.001250 |
When to choose Gemini 3.5 Flash
Gemini 3.5 Flash shines for general-purpose tasks, image and document understanding, high-throughput, low-latency tasks, and audio understanding and generation. Token counts are estimated within ~10-20% margin.