AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Claude Haiku 4.5MiniMax M3 (batch)
Claude Haiku 4.5

Anthropic

MiniMax M3 (batch)

MiniMax

Pricing
Input $/1M$1.00$0.30
Output $/1M$5.00$1.20
Blended $/1M (3:1)$2.00$0.525
Cache read $/1M$0.10$0.06
Cache write $/1M$1.25
Quality indices
Intelligence Index16.929.2
Coding Index43.958.6
Math Index
Agentic Index829.5
Speed & latency
Output speed (tok/s)107
Time to first token1.01s
Time to first answer token19.78s
Benchmarks (Artificial Analysis)
GPQA Diamond92.9%
Humanity's Last Exam39.0%
SciCode47.1%
IFBench82.9%
AA-LCR (Long Context Reasoning)83.0%
Terminal-Bench Hard42.4%
Terminal-Bench 2.165.2%
τ²-Bench (Telecom)88.9%
τ³-Bench Banking15.3%
Design Arena (Elo)
Website1,1351,272
Web apps1,215
Full-stack1,204
Mobile apps1,195
Android native1,167
UI components1,1151,261
Data viz1,1391,250
SVG1,0491,191
3D1,1011,238
Game dev1,1201,239
Agentic game dev1,158
HTML slides1,183
Python→PPTX slides1,207
Code categories1,1301,265
ASCII art1,1591,183
Specs
Context window200K524K
Max output tokens64K472K
Input modalitiestext, image, filetext, image, video
TokenizerClaudeOther