AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude 3 HaikuGPT-4.1
Claude 3 Haiku

Anthropic

GPT-4.1

OpenAI

Pricing
Input $/1M$0.25$2.00
Output $/1M$1.25$8.00
Blended $/1M (3:1)$0.50$3.50
Cache read $/1M$0.03$0.50
Cache write $/1M$0.30
Quality indices
Intelligence Index3.919.4
Coding Index
Math Index34.7
Agentic Index
Speed & latency
Output speed (tok/s)0151
Time to first token0s0.56s
Time to first answer token0s0.56s
Benchmarks (Artificial Analysis)
MMLU-Pro80.6%
GPQA Diamond37.4%66.6%
Humanity's Last Exam3.9%4.6%
LiveCodeBench15.4%45.7%
SciCode18.6%38.1%
MATH-50039.4%91.3%
AIME1.0%43.7%
AIME 202534.7%
IFBench36.1%43.0%
AA-LCR (Long Context Reasoning)21.0%61.0%
Terminal-Bench Hard0.8%13.6%
τ²-Bench (Telecom)21.1%47.1%
Design Arena (Elo)
Website1,066
UI components1,043
Data viz1,140
3D908
Game dev1,138
Code categories1,059
Specs
Context window200K1.05M
Max output tokens4K33K
Input modalitiestext, imageimage, text, file
TokenizerClaudeGPT