o1-pro
by OpenAI · openai/o1-pro · #325 cheapest of 325 paid models
Prices updated Jul 22, 2026, 1:41 AM UTC · refreshed hourly
The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5Gemini 3.6 Flash
Pricing
Input
$150.00
per 1M tokens
Output
$600.00
per 1M tokens
Blended (3:1)
$262.50
3 input : 1 output
Cache read
—
per 1M cached tokens
Cache write
—
per 1M tokens
Cache & batch economics
Effective prices for o1-pro (no cache-read rate published — hits billed as normal input).
| Cache hit rate | Effective blended $/1M | Example request* | vs no cache |
|---|---|---|---|
| 0% | $262.50 | $15.90 | — |
| 50% | $262.50 | $15.90 | −0% |
| 90% | $262.50 | $15.90 | −0% |
*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates for this provider’s batch API. Confirm on the vendor’s pricing page.
What a request costs
| Workload | Input tokens | Output tokens | Cost / request | Cost / 1K requests |
|---|---|---|---|---|
| Short chat message | 500 | 300 | $0.255 | $255.00 |
| Document summary | 8,000 | 1,000 | $1.80 | $1,800 |
| Codebase question (RAG) | 30,000 | 2,000 | $5.70 | $5,700 |
| Long-context analysis | 150,000 | 5,000 | $25.50 | $25,500 |
Performance
Intelligence Index
18.9
AA composite quality
Coding Index
—
Math Index
—
Agentic Index
—
Output speed
0 tok/s
median
Time to first token
0s
median
Time to first answer
0s
after reasoning tokens
Specs
Context window
200K
tokens
Max output
100K
tokens
Input modalities
text, image, file
Tokenizer
GPT