AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Opus 4.6Kimi K2.6
Claude Opus 4.6

Anthropic

Kimi K2.6

Moonshot AI

Pricing
Input $/1M$5.00$0.684
Output $/1M$25.00$3.42
Blended $/1M (3:1)$10.00$1.37
Cache read $/1M$0.50$0.144
Cache write $/1M$6.25
Quality indices
Intelligence Index37.844.2
Coding Index61.8
Math Index
Agentic Index30.3
Speed & latency
Output speed (tok/s)54
Time to first token1.89s
Time to first answer token1.89s
Benchmarks (Artificial Analysis)
GPQA Diamond84.0%
Humanity's Last Exam18.6%
SciCode45.7%
IFBench44.6%
AA-LCR (Long Context Reasoning)58.3%
Terminal-Bench Hard48.5%
τ²-Bench (Telecom)84.8%
Design Arena (Elo)
Website1,3231,304
Web apps1,2611,268
Full-stack1,2791,216
Mobile apps1,2801,248
Android native1,1951,294
UI components1,3361,305
Data viz1,3221,295
SVG1,2741,225
3D1,3351,333
Game dev1,3371,307
Agentic game dev1,157
Godot game dev1,168
HTML slides1,229
PPTX slides1,181
Python→PPTX slides1,180
Agentic slides1,187
Agentic HTML slides1,248
Agentic slides (HTML)1,252
Agentic slides (Python PPTX)1,186
Code categories1,3271,311
ASCII art1,2981,192
Specs
Context window1M262K
Max output tokens128K262K
Input modalitiestext, image, filetext, image
TokenizerClaudeOther