AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Sonnet 5GPT-5.5 (batch)
Claude Sonnet 5

Anthropic

GPT-5.5 (batch)

OpenAI

Pricing
Input $/1M$2.00$2.50
Output $/1M$10.00$15.00
Blended $/1M (3:1)$4.00$5.63
Cache read $/1M$0.20$0.25
Cache write $/1M$2.50
Quality indices
Intelligence Index38.238.4
Coding Index71.574.9
Math Index
Agentic Index43.636.4
Speed & latency
Output speed (tok/s)820
Time to first token118.14s0s
Time to first answer token118.14s0s
Benchmarks (Artificial Analysis)
GPQA Diamond91.1%93.5%
Humanity's Last Exam41.3%45.8%
SciCode54.3%55.8%
IFBench75.9%
AA-LCR (Long Context Reasoning)82.0%84.3%
Terminal-Bench Hard60.6%
Terminal-Bench 2.180.5%84.3%
τ²-Bench (Telecom)93.9%
τ³-Bench Banking37.3%39.0%
Design Arena (Elo)
Website1,2881,269
Web apps1,2601,132
Full-stack1,2601,099
Mobile apps1,2491,171
Android native1,2401,175
UI components1,2991,273
Data viz1,2601,273
SVG1,2121,260
3D1,2891,226
Game dev1,3131,320
Agentic game dev1,2271,179
Godot game dev1,2381,183
HTML slides1,2201,070
PPTX slides1,157
Python→PPTX slides1,2161,152
Agentic slides1,150
Agentic HTML slides1,084
Agentic slides (HTML)1,077
Agentic slides (Python PPTX)1,155
Code categories1,2931,274
ASCII art1,2261,276
Specs
Context window1M1.05M
Max output tokens128K128K
Input modalitiestext, image, filefile, image, text
TokenizerClaudeGPT