Skip to content
S/T

Claude Sonnet 4.5

$3 per million input tokens, $15 per million output.

Input$3 / MTok
Output$15 / MTok
Cache read$0.30 / MTok
Cache write (5 min)$3.75 / MTok
Cache write (1 hour)$6 / MTok
Batch API$1.50 in / $7.50 out
Context window200,000 tokens
Tokenizer familyClaude (Sonnet 4.6 and earlier)
Tool-use overhead496 tokens added when any tool is defined

What that costs per month

Monthly spend for a prompt of a given size, with a 500-token response, no caching. Rows are prompt size; columns are requests per day.

Prompt size1,000 / day10,000 / day100,000 / day
1,000 tokens$319.37$3,194$31,938
5,000 tokens$684.38$6,844$68,438
20,000 tokens$2,053$20,531$205,313

Prompt caching on Claude Sonnet 4.5

A 10,000-token static prefix at 10,000 requests a day costs $11,863 a month uncached. With caching at an 80% hit rate it costs $5,749 — a saving of $6,114, or 52%. Caching pays for itself after 1 read of the same prefix.

How cache economics work

Cheaper on input

Lower input price. Whether they are cheaper for your task depends on token counts and on whether they hold up on your evals.

Price your actual prompt

These figures assume a prompt size. Paste your real one and the analyser will count it with Claude Sonnet 4.5’s tokenizer family, find what is wasted, and compare against every other model at your volume.

Open the analyser

Prices last verified 2026-08-10. Confirm against Anthropic’s own pricing page before committing to anything.