GPT-5.4 Pro
by OpenAI · openai/gpt-5.4-pro · #324 cheapest of 325 paid models
Prices updated Jul 22, 2026, 1:39 AM UTC · refreshed hourly
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5Gemini 3.6 Flash
Pricing
Input
$30.00
per 1M tokens
Output
$180.00
per 1M tokens
Blended (3:1)
$67.50
3 input : 1 output
Cache read
—
per 1M cached tokens
Cache write
—
per 1M tokens
Cache & batch economics
Effective prices for GPT-5.4 Pro (no cache-read rate published — hits billed as normal input).
| Cache hit rate | Effective blended $/1M | Example request* | vs no cache |
|---|---|---|---|
| 0% | $67.50 | $3.24 | — |
| 50% | $67.50 | $3.24 | −0% |
| 90% | $67.50 | $3.24 | −0% |
*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates for this provider’s batch API. Confirm on the vendor’s pricing page.
What a request costs
| Workload | Input tokens | Output tokens | Cost / request | Cost / 1K requests |
|---|---|---|---|---|
| Short chat message | 500 | 300 | $0.069 | $69.00 |
| Document summary | 8,000 | 1,000 | $0.42 | $420.00 |
| Codebase question (RAG) | 30,000 | 2,000 | $1.26 | $1,260 |
| Long-context analysis | 150,000 | 5,000 | $5.40 | $5,400 |
Performance
Intelligence Index
—
AA composite quality
Coding Index
—
Math Index
—
Agentic Index
—
Output speed
0 tok/s
median
Time to first token
0s
median
Time to first answer
0s
after reasoning tokens
Specs
Context window
1.05M
tokens
Max output
128K
tokens
Input modalities
text, image, file
Tokenizer
GPT