AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5Qwen3.7 Max
Claude Haiku 4.5

Anthropic

Qwen3.7 Max

Alibaba (Qwen)

Pricing
Input $/1M$1.00$1.48
Output $/1M$5.00$4.42
Blended $/1M (3:1)$2.00$2.21
Cache read $/1M$0.10$0.295
Cache write $/1M$1.25$1.84
Quality indices
Intelligence Index29.646
Coding Index43.966
Math Index
Agentic Index16.430.6
Speed & latency
Output speed (tok/s)207
Time to first token1.55s
Time to first answer token13.16s
Benchmarks (Artificial Analysis)
GPQA Diamond92.3%
Humanity's Last Exam38.1%
SciCode48.8%
IFBench80.5%
AA-LCR (Long Context Reasoning)69.0%
Terminal-Bench Hard50.8%
Terminal-Bench 2.174.5%
τ²-Bench (Telecom)94.7%
τ³-Bench Banking10.9%
Design Arena (Elo)
Website1,1491,291
Web apps1,253
Full-stack1,225
Mobile apps1,208
Android native1,182
UI components1,1421,313
Data viz1,1601,304
SVG1,0721,267
3D1,1311,310
Game dev1,1551,315
Agentic game dev1,180
Godot game dev1,248
HTML slides1,191
Python→PPTX slides1,229
Code categories1,1481,299
ASCII art1,1821,261
Specs
Context window200K1M
Max output tokens64K66K
Input modalitiestext, image, filetext
TokenizerClaudeQwen