Gemini 2.5 Pro
$1.25 per million input tokens, $10 per million output. Superseded — kept here so you can price existing code.
Input price doubles above a 200k-token prompt.
| Input | $1.25 / MTok |
|---|---|
| Output | $10 / MTok |
| Cache read | $0.13 / MTok |
| Cache write (5 min) | billed at the base input rate |
| Above 200,000 tokens | $2.50 in / $15 out |
| Context window | 1,048,576 tokens |
| Tokenizer family | Gemini |
What that costs per month
Monthly spend for a prompt of a given size, with a 500-token response, no caching. Rows are prompt size; columns are requests per day.
| Prompt size | 1,000 / day | 10,000 / day | 100,000 / day |
|---|---|---|---|
| 1,000 tokens | $190.10 | $1,901 | $19,010 |
| 5,000 tokens | $342.19 | $3,422 | $34,219 |
| 20,000 tokens | $912.50 | $9,125 | $91,250 |
Prompt caching on Gemini 2.5 Pro
A 10,000-token static prefix at 10,000 requests a day costs $5,513 a month uncached. With caching at an 80% hit rate it costs $2,776 — a saving of $2,738, or 50%. Caching pays for itself after 0 reads of the same prefix.
How cache economics workCheaper on input
Lower input price. Whether they are cheaper for your task depends on token counts and on whether they hold up on your evals.
- Claude Haiku 4.5$1 / $5Anthropic
- GPT-5.4 mini$0.75 / $4.50OpenAI
- Gemini 3 Flash Preview$0.50 / $3Google
- Gemini 3.5 Flash-Lite$0.30 / $2.50Google
- GPT-5 mini$0.25 / $2OpenAI
Price your actual prompt
These figures assume a prompt size. Paste your real one and the analyser will count it with Gemini 2.5 Pro’s tokenizer family, find what is wasted, and compare against every other model at your volume.
Open the analyserPrices last verified 2026-08-10. Confirm against Google’s own pricing page before committing to anything.