AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5Gemini 3.1 Pro Preview
Claude Haiku 4.5

Anthropic

Gemini 3.1 Pro Preview

Google

Pricing
Input $/1M$1.00$2.00
Output $/1M$5.00$12.00
Blended $/1M (3:1)$2.00$4.50
Cache read $/1M$0.10$0.20
Cache write $/1M$1.25$0.375
Quality indices
Intelligence Index29.646.5
Coding Index43.968.8
Math Index
Agentic Index16.421.4
Speed & latency
Output speed (tok/s)135
Time to first token19.70s
Time to first answer token19.70s
Benchmarks (Artificial Analysis)
GPQA Diamond94.1%
Humanity's Last Exam44.7%
SciCode58.9%
IFBench77.1%
AA-LCR (Long Context Reasoning)72.7%
Terminal-Bench Hard53.8%
Terminal-Bench 2.173.8%
τ²-Bench (Telecom)95.6%
τ³-Bench Banking16.5%
Design Arena (Elo)
Website1,1491,278
Web apps1,182
Full-stack1,127
Mobile apps1,167
Android native1,038
UI components1,1421,309
Data viz1,1601,262
SVG1,0721,334
3D1,1311,291
Game dev1,1551,258
Agentic game dev1,128
Godot game dev1,159
HTML slides1,202
PPTX slides1,110
Python→PPTX slides1,109
Agentic slides1,112
Agentic HTML slides1,226
Agentic slides (HTML)1,219
Agentic slides (Python PPTX)1,107
Code categories1,1481,274
ASCII art1,1821,312
Specs
Context window200K1.05M
Max output tokens64K66K
Input modalitiestext, image, fileaudio, file, image, text, video
TokenizerClaudeGemini