AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5o3
Claude Haiku 4.5

Anthropic

o3

OpenAI

Pricing
Input $/1M$1.00$2.00
Output $/1M$5.00$8.00
Blended $/1M (3:1)$2.00$3.50
Cache read $/1M$0.10$0.50
Cache write $/1M$1.25
Quality indices
Intelligence Index29.630.4
Coding Index43.9
Math Index88.3
Agentic Index16.4
Speed & latency
Output speed (tok/s)162
Time to first token4.84s
Time to first answer token4.84s
Benchmarks (Artificial Analysis)
MMLU-Pro85.3%
GPQA Diamond82.7%
Humanity's Last Exam20.0%
LiveCodeBench80.8%
SciCode41.0%
MATH-50099.2%
AIME90.3%
AIME 202588.3%
IFBench71.4%
AA-LCR (Long Context Reasoning)69.3%
Terminal-Bench Hard37.1%
τ²-Bench (Telecom)80.7%
Design Arena (Elo)
Website1,1491,064
UI components1,1421,060
Data viz1,1601,200
SVG1,072
3D1,131
Game dev1,1551,093
Code categories1,1481,053
ASCII art1,182
Specs
Context window200K200K
Max output tokens64K100K
Input modalitiestext, image, fileimage, text, file
TokenizerClaudeGPT