AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Opus 4.6Kimi K3
Claude Opus 4.6

Anthropic

Kimi K3

Moonshot AI

Pricing
Input $/1M$5.00$1.70
Output $/1M$25.00$8.50
Blended $/1M (3:1)$10.00$3.40
Cache read $/1M$0.50$0.17
Cache write $/1M$6.25
Quality indices
Intelligence Index26.443.6
Coding Index76.2
Math Index
Agentic Index50
Speed & latency
Output speed (tok/s)0
Time to first token0s
Time to first answer token0s
Benchmarks (Artificial Analysis)
GPQA Diamond84.0%
Humanity's Last Exam19.1%
IFBench44.6%
AA-LCR (Long Context Reasoning)67.0%
Terminal-Bench Hard48.5%
τ²-Bench (Telecom)84.8%
Design Arena (Elo)
Website1,3031,352
Web apps1,2191,309
Full-stack1,2291,333
Mobile apps1,2381,279
Android native1,2161,259
UI components1,3001,371
Data viz1,2961,358
SVG1,2511,330
3D1,3031,425
Game dev1,3011,407
Agentic game dev1,2151,252
Godot game dev1,199
HTML slides1,259
Python→PPTX slides1,273
Code categories1,3041,386
ASCII art1,273
Specs
Context window1M1.05M
Max output tokens128K944K
Input modalitiestext, image, filetext, image, video
TokenizerClaudeOther