AiCostCompare

MiniMax M2.5

by MiniMax · minimax/minimax-m2.5 · #89 cheapest of 325 paid models

Prices updated Jul 22, 2026, 1:31 AM UTC · refreshed hourly

MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol

Pricing

Input

$0.15

per 1M tokens

Output

$0.90

per 1M tokens

Blended (3:1)

$0.337

3 input : 1 output

Cache read

$0.05

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for MiniMax M2.5 with prompt caching.

Cache hit rateEffective blended $/1MExample request*vs no cache
0%$0.337$0.0162
50%$0.30$0.0112−31%
90%$0.27$0.0072−56%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0.000345$0.345
Document summary8,0001,000$0.0021$2.10
Codebase question (RAG)30,0002,000$0.0063$6.30
Long-context analysis150,0005,000$0.027$27.00

Performance

Intelligence Index

33.7

AA composite quality

Coding Index

Math Index

Agentic Index

Output speed

87 tok/s

median

Time to first token

1.02s

median

Time to first answer

24.07s

after reasoning tokens

Artificial Analysis benchmarks

BenchmarkScore
GPQA Diamond84.8%
Humanity's Last Exam19.1%
SciCode42.6%
IFBench71.6%
AA-LCR (Long Context Reasoning)66.0%
Terminal-Bench Hard34.8%
τ²-Bench (Telecom)95.3%

Design Arena Elo

CategoryElo
Website1,249
UI components1,215
Data viz1,208
SVG1,197
3D1,228
Game dev1,235
Code categories1,240

Specs

Context window

205K

tokens

Max output

197K

tokens

Input modalities

text

Tokenizer

Other