AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5GPT-5.1 (batch)
Claude Haiku 4.5

Anthropic

GPT-5.1 (batch)

OpenAI

Pricing
Input $/1M$1.00$0.625
Output $/1M$5.00$5.00
Blended $/1M (3:1)$2.00$1.72
Cache read $/1M$0.10$0.063
Cache write $/1M$1.25
Quality indices
Intelligence Index16.924.7
Coding Index43.949.4
Math Index94
Agentic Index8
Speed & latency
Output speed (tok/s)0
Time to first token0s
Time to first answer token0s
Benchmarks (Artificial Analysis)
MMLU-Pro87.0%
GPQA Diamond87.3%
Humanity's Last Exam28.5%
LiveCodeBench86.8%
AIME 202594.0%
IFBench72.9%
AA-LCR (Long Context Reasoning)80.0%
Terminal-Bench Hard45.5%
Terminal-Bench 2.152.4%
τ²-Bench (Telecom)81.9%
τ³-Bench Banking15.9%
Design Arena (Elo)
Website1,1351,199
UI components1,1151,182
Data viz1,1391,219
SVG1,0491,173
3D1,1011,089
Game dev1,1201,202
Code categories1,1301,186
ASCII art1,1591,137
Specs
Context window200K400K
Max output tokens64K128K
Input modalitiestext, image, fileimage, text, file
TokenizerClaudeGPT