Gemini 3.5 Flash-Lite pricing
Google · list prices as of 2026-09-29
| Input | $0.30 per 1M tokens |
|---|---|
| Output | $2.50 per 1M tokens |
| Cached input (read) | $0.03 per 1M tokens |
| Batch / async | 50% off |
| Context window | 1.049M tokens |
| Longest answer | 66K tokens |
What a call costs
| Typical call | Input tokens | Output tokens | Per call | Per month at 1,000 a day |
|---|---|---|---|---|
| Chat reply | 1,000 | 300 | $0.00105 | $31.50 |
| Document summary | 6,000 | 400 | $0.0028 | $84.00 |
| Code change | 3,000 | 1,200 | $0.0039 | $117.00 |
No caching or batch discount. Your own prompt will differ: count it exactly below.
Price your own prompt on Gemini 3.5 Flash-LiteSimilarly priced models
- Gemini 2.5 Flash (Google): about the same
- Llama 3.3 70B (Meta): 1.1× more
- GPT-3.5 Turbo (OpenAI): 1.1× cheaper
- Mistral Large 3 (Mistral): 1.1× cheaper
- DeepSeek R1 (DeepSeek): 1.1× more
Questions
How much does Gemini 3.5 Flash-Lite cost?
Gemini 3.5 Flash-Lite costs $0.30 per million input tokens and $2.50 per million output tokens, as of 2026-09-29.
How much is one Gemini 3.5 Flash-Lite call?
A call with 1,000 input tokens and a 300-token answer costs about $0.00105, or $31.50 a month at 1,000 calls a day.
How big is the Gemini 3.5 Flash-Lite context window?
1,048,576 tokens.