AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5Grok 4.6
Claude Haiku 4.5

Anthropic

Grok 4.6

xAI

Pricing
Input $/1M$1.00$2.00
Output $/1M$5.00$6.00
Blended $/1M (3:1)$2.00$3.00
Cache read $/1M$0.10$0.50
Cache write $/1M$1.25
Quality indices
Intelligence Index16.944.3
Coding Index43.976.8
Math Index
Agentic Index853
Speed & latency
Output speed (tok/s)58
Time to first token34.50s
Time to first answer token34.50s
Benchmarks (Artificial Analysis)
GPQA Diamond94.9%
Humanity's Last Exam42.9%
SciCode56.5%
AA-LCR (Long Context Reasoning)80.3%
Terminal-Bench 2.188.4%
τ³-Bench Banking50.7%
Design Arena (Elo)
Website1,1351,304
Web apps1,264
Full-stack1,274
Mobile apps1,267
Android native1,301
UI components1,1151,305
Data viz1,1391,303
SVG1,0491,263
3D1,1011,305
Game dev1,1201,321
Agentic game dev1,210
HTML slides1,253
Python→PPTX slides1,242
Code categories1,1301,309
ASCII art1,1591,297
Specs
Context window200K500K
Max output tokens64K450K
Input modalitiestext, image, filetext, image, file
TokenizerClaudeGrok