Mistral Small 3.2 24B
by Mistral · mistralai/mistral-small-3.2-24b-instruct · #61 cheapest of 422 paid models
Prices updated Sep 19, 2026, 6:22 PM UTC · refreshed hourly
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...
Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol
Pricing
Input
$0.094
per 1M tokens
Output
$0.25
per 1M tokens
Blended (3:1)
$0.133
3 input : 1 output
Cache read
—
per 1M cached tokens
Cache write
—
per 1M tokens
Cache & batch economics
Effective prices for Mistral Small 3.2 24B (no cache-read rate published — hits billed as normal input).
| Cache hit rate | Effective blended $/1M | Example request* | vs no cache |
|---|---|---|---|
| 0% | $0.133 | $0.009813 | — |
| 50% | $0.133 | $0.009813 | −0% |
| 90% | $0.133 | $0.009813 | −0% |
*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.
What a request costs
| Workload | Input tokens | Output tokens | Cost / request | Cost / 1K requests |
|---|---|---|---|---|
| Short chat message | 500 | 300 | $0.000122 | $0.1219 |
| Document summary | 8,000 | 1,000 | $0.001 | $1.00 |
| Codebase question (RAG) | 30,000 | 2,000 | $0.003312 | $3.31 |
| Long-context analysis | 150,000 | 5,000 | $0.0153 | $15.31 |
Performance
Intelligence Index
—
AA composite quality
Coding Index
—
Math Index
—
Agentic Index
—
Output speed
—
median
Time to first token
—
median
Time to first answer
—
after reasoning tokens
Design Arena Elo
| Category | Elo |
|---|---|
| Website | 908 |
| UI components | 922 |
| Data viz | 942 |
| Game dev | 912 |
| Code categories | 922 |
Specs
Context window
256K
tokens
Max output
16K
tokens
Input modalities
image, text
Tokenizer
Mistral