AiCostCompare

MiniMax M2.1

by MiniMax · minimax/minimax-m2.1 · #127 cheapest of 325 paid models

Prices updated Jul 22, 2026, 1:31 AM UTC · refreshed hourly

MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol

Pricing

Input

$0.30

per 1M tokens

Output

$1.20

per 1M tokens

Blended (3:1)

$0.525

3 input : 1 output

Cache read

$0.03

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for MiniMax M2.1 with prompt caching.

Cache hit rateEffective blended $/1MExample request*vs no cache
0%$0.525$0.0318
50%$0.424$0.0183−42%
90%$0.343$0.0075−76%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0.00051$0.51
Document summary8,0001,000$0.0036$3.60
Codebase question (RAG)30,0002,000$0.0114$11.40
Long-context analysis150,0005,000$0.051$51.00

Performance

Intelligence Index

31.4

AA composite quality

Coding Index

Math Index

82.7

Agentic Index

Output speed

89 tok/s

median

Time to first token

1s

median

Time to first answer

23.51s

after reasoning tokens

Artificial Analysis benchmarks

BenchmarkScore
MMLU-Pro87.5%
GPQA Diamond83.0%
Humanity's Last Exam22.2%
LiveCodeBench81.0%
SciCode40.7%
AIME 202582.7%
IFBench69.9%
AA-LCR (Long Context Reasoning)59.0%
Terminal-Bench Hard28.8%
τ²-Bench (Telecom)85.4%

Design Arena Elo

CategoryElo
Website1,229
UI components1,265
Data viz1,243
SVG1,179
3D1,224
Game dev1,193
Code categories1,224

Specs

Context window

205K

tokens

Max output

131K

tokens

Input modalities

text

Tokenizer

Other