AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Opus 4.6MiniMax M3
Claude Opus 4.6

Anthropic

MiniMax M3

MiniMax

Pricing
Input $/1M$5.00$0.30
Output $/1M$25.00$1.20
Blended $/1M (3:1)$10.00$0.525
Cache read $/1M$0.50$0.06
Cache write $/1M$6.25
Quality indices
Intelligence Index37.844.4
Coding Index58.6
Math Index
Agentic Index35.4
Speed & latency
Output speed (tok/s)5497
Time to first token1.89s1.29s
Time to first answer token1.89s21.90s
Benchmarks (Artificial Analysis)
GPQA Diamond84.0%92.9%
Humanity's Last Exam18.6%37.1%
SciCode45.7%45.4%
IFBench44.6%82.9%
AA-LCR (Long Context Reasoning)58.3%74.0%
Terminal-Bench Hard48.5%42.4%
Terminal-Bench 2.165.2%
τ²-Bench (Telecom)84.8%88.9%
τ³-Bench Banking13.0%
Design Arena (Elo)
Website1,3231,287
Web apps1,2611,263
Full-stack1,2791,250
Mobile apps1,2801,256
Android native1,1951,099
UI components1,3361,280
Data viz1,3221,277
SVG1,2741,227
3D1,3351,280
Game dev1,3371,280
Code categories1,3271,287
ASCII art1,2981,198
Specs
Context window1M1.05M
Max output tokens128K512K
Input modalitiestext, image, filetext, image, video
TokenizerClaudeOther