Mistral Small 3.2 24B
by Mistral · mistralai/mistral-small-3.2-24b-instruct · #52 cheapest of 325 paid models
Prices updated Jul 22, 2026, 2:19 AM UTC · refreshed hourly
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...
Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol
Pricing
Input
$0.10
per 1M tokens
Output
$0.30
per 1M tokens
Blended (3:1)
$0.15
3 input : 1 output
Cache read
$0.01
per 1M cached tokens
Cache write
—
per 1M tokens
Cache & batch economics
Effective prices for Mistral Small 3.2 24B with prompt caching.
| Cache hit rate | Effective blended $/1M | Example request* | vs no cache |
|---|---|---|---|
| 0% | $0.15 | $0.0105 | — |
| 50% | $0.116 | $0.006 | −43% |
| 90% | $0.089 | $0.0024 | −77% |
*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.
What a request costs
| Workload | Input tokens | Output tokens | Cost / request | Cost / 1K requests |
|---|---|---|---|---|
| Short chat message | 500 | 300 | $0.00014 | $0.14 |
| Document summary | 8,000 | 1,000 | $0.0011 | $1.10 |
| Codebase question (RAG) | 30,000 | 2,000 | $0.0036 | $3.60 |
| Long-context analysis | 150,000 | 5,000 | $0.0165 | $16.50 |
Performance
Intelligence Index
—
AA composite quality
Coding Index
—
Math Index
—
Agentic Index
—
Output speed
—
median
Time to first token
—
median
Time to first answer
—
after reasoning tokens
Design Arena Elo
| Category | Elo |
|---|---|
| Website | 923 |
| UI components | 950 |
| Data viz | 965 |
| Game dev | 947 |
| Code categories | 940 |
Specs
Context window
256K
tokens
Max output
—
tokens
Input modalities
image, text
Tokenizer
Mistral