AiCostCompare

GPT-4o (2024-05-13)

by OpenAI · openai/gpt-4o-2024-05-13 · #295 cheapest of 325 paid models

Prices updated Jul 22, 2026, 1:42 AM UTC · refreshed hourly

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5Gemini 3.6 Flash

Pricing

Input

$5.00

per 1M tokens

Output

$15.00

per 1M tokens

Blended (3:1)

$7.50

3 input : 1 output

Cache read

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for GPT-4o (2024-05-13) (no cache-read rate published — hits billed as normal input).

Cache hit rateEffective blended $/1MExample request*vs no cache
0%$7.50$0.525
50%$7.50$0.525−0%
90%$7.50$0.525−0%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates for this provider’s batch API. Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0.007$7.00
Document summary8,0001,000$0.055$55.00
Codebase question (RAG)30,0002,000$0.18$180.00
Long-context analysis150,0005,000$0.825$825.00

Performance

Intelligence Index

8.6

AA composite quality

Coding Index

24.2

Math Index

Agentic Index

Output speed

107 tok/s

median

Time to first token

0.51s

median

Time to first answer

0.51s

after reasoning tokens

Artificial Analysis benchmarks

BenchmarkScore
MMLU-Pro74.0%
GPQA Diamond52.6%
Humanity's Last Exam2.8%
LiveCodeBench33.4%
SciCode30.9%
MATH-50079.1%
AIME11.0%

Specs

Context window

128K

tokens

Max output

4K

tokens

Input modalities

text, image, file

Tokenizer

GPT