AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5R1 Distill Llama 70B
Claude Haiku 4.5

Anthropic

R1 Distill Llama 70B

DeepSeek

Pricing
Input $/1M$1.00$0.80
Output $/1M$5.00$0.80
Blended $/1M (3:1)$2.00$0.80
Cache read $/1M$0.10
Cache write $/1M$1.25
Quality indices
Intelligence Index29.69.9
Coding Index43.9
Math Index53.7
Agentic Index16.4
Speed & latency
Output speed (tok/s)32
Time to first token0.61s
Time to first answer token62.16s
Benchmarks (Artificial Analysis)
MMLU-Pro79.5%
GPQA Diamond40.2%
Humanity's Last Exam6.1%
LiveCodeBench26.6%
SciCode31.3%
MATH-50093.5%
AIME67.0%
AIME 202553.7%
IFBench27.6%
AA-LCR (Long Context Reasoning)11.0%
Terminal-Bench Hard1.5%
τ²-Bench (Telecom)21.9%
Design Arena (Elo)
Website1,149
UI components1,142
Data viz1,160
SVG1,072
3D1,131
Game dev1,155
Code categories1,148
ASCII art1,182
Specs
Context window200K8K
Max output tokens64K8K
Input modalitiestext, image, filetext
TokenizerClaudeLlama3