Gemini 3.5 Flash-Lite
$0.30 per million input tokens, $2.50 per million output.
| Input | $0.30 / MTok |
|---|---|
| Output | $2.50 / MTok |
| Cache read | $0.030 / MTok |
| Cache write (5 min) | billed at the base input rate |
| Context window | not published here |
| Tokenizer family | Gemini |
What that costs per month
Monthly spend for a prompt of a given size, with a 500-token response, no caching. Rows are prompt size; columns are requests per day.
| Prompt size | 1,000 / day | 10,000 / day | 100,000 / day |
|---|---|---|---|
| 1,000 tokens | $47.15 | $471.46 | $4,715 |
| 5,000 tokens | $83.65 | $836.46 | $8,365 |
| 20,000 tokens | $220.52 | $2,205 | $22,052 |
Prompt caching on Gemini 3.5 Flash-Lite
A 10,000-token static prefix at 10,000 requests a day costs $1,338 a month uncached. With caching at an 80% hit rate it costs $681.33 — a saving of $657.00, or 49%. Caching pays for itself after 0 reads of the same prefix.
How cache economics workCheaper on input
Lower input price. Whether they are cheaper for your task depends on token counts and on whether they hold up on your evals.
- GPT-5 mini$0.25 / $2OpenAI
- Gemini 3.1 Flash-Lite$0.25 / $1.50Google
- GPT-5.6 Luna$0.20 / $1.20OpenAI
- GPT-5.4 nano$0.20 / $1.25OpenAI
- GPT-5 nano$0.050 / $0.40OpenAI
Price your actual prompt
These figures assume a prompt size. Paste your real one and the analyser will count it with Gemini 3.5 Flash-Lite’s tokenizer family, find what is wasted, and compare against every other model at your volume.
Open the analyserPrices last verified 2026-08-10. Confirm against Google’s own pricing page before committing to anything.