AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5Qwen3.8 2.4T A95B
Claude Haiku 4.5

Anthropic

Qwen3.8 2.4T A95B

Alibaba (Qwen)

Pricing
Input $/1M$1.00$2.00
Output $/1M$5.00$6.00
Blended $/1M (3:1)$2.00$3.00
Cache read $/1M$0.10$0.25
Cache write $/1M$1.25
Quality indices
Intelligence Index16.939.9
Coding Index43.971.9
Math Index
Agentic Index850.1
Speed & latency
Output speed (tok/s)39
Time to first token1.88s
Time to first answer token53.08s
Benchmarks (Artificial Analysis)
GPQA Diamond93.5%
Humanity's Last Exam42.4%
SciCode54.1%
AA-LCR (Long Context Reasoning)80.3%
Terminal-Bench 2.182.0%
τ³-Bench Banking49.1%
Design Arena (Elo)
Website1,135
UI components1,115
Data viz1,139
SVG1,049
3D1,101
Game dev1,120
Code categories1,130
ASCII art1,159
Specs
Context window200K1.05M
Max output tokens64K131K
Input modalitiestext, image, filetext
TokenizerClaudeQwen