Skip to content
S/T

GPT-4o

$2.50 per million input tokens, $10 per million output. Superseded — kept here so you can price existing code.

Input$2.50 / MTok
Output$10 / MTok
Cache read$1.25 / MTok
Cache write (5 min)billed at the base input rate
Context window128,000 tokens
Tokenizer familyo200k_base (OpenAI)

What that costs per month

Monthly spend for a prompt of a given size, with a 500-token response, no caching. Rows are prompt size; columns are requests per day.

Prompt size1,000 / day10,000 / day100,000 / day
1,000 tokens$228.13$2,281$22,813
5,000 tokens$532.29$5,323$53,229
20,000 tokens$1,673$16,729$167,292

Prompt caching on GPT-4o

A 10,000-token static prefix at 10,000 requests a day costs $9,505 a month uncached. With caching at an 80% hit rate it costs $6,464 — a saving of $3,042, or 32%. Caching pays for itself after 0 reads of the same prefix.

How cache economics work

Cheaper on input

Lower input price. Whether they are cheaper for your task depends on token counts and on whether they hold up on your evals.

Price your actual prompt

These figures assume a prompt size. Paste your real one and the analyser will count it with GPT-4o’s tokenizer family, find what is wasted, and compare against every other model at your volume.

Open the analyser

Prices last verified 2026-08-10. Confirm against OpenAI’s own pricing page before committing to anything.