AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5Grok 4.3 (batch)
Claude Haiku 4.5

Anthropic

Grok 4.3 (batch)

xAI

Pricing
Input $/1M$1.00$1.00
Output $/1M$5.00$2.00
Blended $/1M (3:1)$2.00$1.25
Cache read $/1M$0.10$0.16
Cache write $/1M$1.25
Quality indices
Intelligence Index16.924.9
Coding Index43.942.2
Math Index
Agentic Index815.5
Speed & latency
Output speed (tok/s)0
Time to first token0s
Time to first answer token0s
Benchmarks (Artificial Analysis)
GPQA Diamond90.1%
Humanity's Last Exam37.2%
SciCode48.3%
IFBench81.3%
AA-LCR (Long Context Reasoning)73.0%
Terminal-Bench Hard37.9%
Terminal-Bench 2.139.7%
τ²-Bench (Telecom)97.7%
τ³-Bench Banking12.4%
Design Arena (Elo)
Website1,1351,207
Web apps1,144
Full-stack1,017
Mobile apps1,092
Android native972
UI components1,1151,207
Data viz1,1391,197
SVG1,0491,109
3D1,1011,154
Game dev1,1201,198
Agentic game dev1,007
Godot game dev1,035
HTML slides1,011
PPTX slides1,071
Python→PPTX slides1,072
Agentic slides1,072
Agentic HTML slides1,066
Agentic slides (HTML)1,066
Agentic slides (Python PPTX)1,068
Code categories1,1301,202
ASCII art1,1591,157
Specs
Context window200K1M
Max output tokens64K900K
Input modalitiestext, image, filetext, image, file
TokenizerClaudeGrok