AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Opus 4.6GLM 5.3 (batch)
Claude Opus 4.6

Anthropic

GLM 5.3 (batch)

Z.ai

Pricing
Input $/1M$5.00$0.70
Output $/1M$25.00$2.20
Blended $/1M (3:1)$10.00$1.07
Cache read $/1M$0.50$0.13
Cache write $/1M$6.25
Quality indices
Intelligence Index26.444.8
Coding Index74.8
Math Index
Agentic Index53.1
Speed & latency
Output speed (tok/s)0
Time to first token0s
Time to first answer token0s
Benchmarks (Artificial Analysis)
GPQA Diamond84.0%
Humanity's Last Exam19.1%
IFBench44.6%
AA-LCR (Long Context Reasoning)67.0%
Terminal-Bench Hard48.5%
τ²-Bench (Telecom)84.8%
Design Arena (Elo)
Website1,3031,314
Web apps1,219
Full-stack1,229
Mobile apps1,2381,223
Android native1,216
UI components1,3001,348
Data viz1,2961,274
SVG1,2511,320
3D1,3031,388
Game dev1,3011,375
Agentic game dev1,215
HTML slides1,189
Python→PPTX slides1,258
Code categories1,3041,329
ASCII art1,2731,245
Specs
Context window1M1.05M
Max output tokens128K944K
Input modalitiestext, image, filetext
TokenizerClaudeOther