AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5Qwen3.5 397B A17B
Claude Haiku 4.5

Anthropic

Qwen3.5 397B A17B

Alibaba (Qwen)

Pricing
Input $/1M$1.00$0.55
Output $/1M$5.00$3.50
Blended $/1M (3:1)$2.00$1.29
Cache read $/1M$0.10$0.225
Cache write $/1M$1.25
Quality indices
Intelligence Index16.918.4
Coding Index43.948.2
Math Index
Agentic Index88.3
Speed & latency
Output speed (tok/s)87
Time to first token1.59s
Time to first answer token38.12s
Benchmarks (Artificial Analysis)
GPQA Diamond89.3%
Humanity's Last Exam29.0%
SciCode44.8%
IFBench78.8%
AA-LCR (Long Context Reasoning)77.3%
Terminal-Bench Hard40.9%
Terminal-Bench 2.151.3%
τ²-Bench (Telecom)95.6%
τ³-Bench Banking13.4%
Design Arena (Elo)
Website1,1351,203
UI components1,1151,179
Data viz1,1391,190
SVG1,0491,163
3D1,1011,189
Game dev1,1201,163
Code categories1,1301,195
ASCII art1,159
Specs
Context window200K262K
Max output tokens64K236K
Input modalitiestext, image, filetext, image, video
TokenizerClaudeQwen3