AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Opus 4.6Kimi K2.5
Claude Opus 4.6

Anthropic

Kimi K2.5

Moonshot AI

Pricing
Input $/1M$5.00$0.57
Output $/1M$25.00$2.85
Blended $/1M (3:1)$10.00$1.14
Cache read $/1M$0.50$0.095
Cache write $/1M$6.25
Quality indices
Intelligence Index37.835.4
Coding Index46.8
Math Index
Agentic Index21.7
Speed & latency
Output speed (tok/s)54
Time to first token1.89s
Time to first answer token1.89s
Benchmarks (Artificial Analysis)
GPQA Diamond84.0%
Humanity's Last Exam18.6%
SciCode45.7%
IFBench44.6%
AA-LCR (Long Context Reasoning)58.3%
Terminal-Bench Hard48.5%
τ²-Bench (Telecom)84.8%
Design Arena (Elo)
Website1,3231,277
Web apps1,2611,181
Full-stack1,2791,173
Mobile apps1,2801,181
Android native1,1951,112
UI components1,3361,278
Data viz1,3221,262
SVG1,2741,193
3D1,3351,264
Game dev1,3371,264
Godot game dev1,211
Code categories1,3271,270
ASCII art1,2981,210
Specs
Context window1M262K
Max output tokens128K262K
Input modalitiestext, image, filetext, image
TokenizerClaudeOther