AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5GPT-5.1
Claude Haiku 4.5

Anthropic

GPT-5.1

OpenAI

Pricing
Input $/1M$1.00$1.25
Output $/1M$5.00$10.00
Blended $/1M (3:1)$2.00$3.44
Cache read $/1M$0.10$0.125
Cache write $/1M$1.25
Quality indices
Intelligence Index29.636.9
Coding Index43.949.4
Math Index94
Agentic Index16.421
Speed & latency
Output speed (tok/s)110
Time to first token41.75s
Time to first answer token41.75s
Benchmarks (Artificial Analysis)
MMLU-Pro87.0%
GPQA Diamond87.3%
Humanity's Last Exam26.5%
LiveCodeBench86.8%
SciCode43.3%
AIME 202594.0%
IFBench72.9%
AA-LCR (Long Context Reasoning)75.0%
Terminal-Bench Hard45.5%
Terminal-Bench 2.152.4%
τ²-Bench (Telecom)81.9%
τ³-Bench Banking14.0%
Design Arena (Elo)
Website1,1491,215
Web apps1,075
Mobile apps1,122
UI components1,1421,209
Data viz1,1601,241
SVG1,0721,196
3D1,1311,120
Game dev1,1551,237
Code categories1,1481,204
ASCII art1,1821,160
Specs
Context window200K400K
Max output tokens64K128K
Input modalitiestext, image, fileimage, text, file
TokenizerClaudeGPT