GPT-5.4 mini charges $0.75 per 1M input tokens and $4.50 per 1M output tokens (cached input: $0.075/M). Real bills depend on your input/output split — the tables below price GPT-5.4 mini at 1K–1M tokens across common workload shapes. Verified Aug 16, 2026
Cost by token volume
| Volume | Input-heavy (90/10) | Balanced (70/30) | Output-heavy (30/70) | All input | All output |
|---|---|---|---|---|---|
| 1,000 tokens | $0.001125 | $0.001875 | $0.003375 | $0.00075 | $0.0045 |
| 10,000 tokens | $0.0113 | $0.0188 | $0.0338 | $0.0075 | $0.045 |
| 100,000 tokens | $0.1125 | $0.1875 | $0.3375 | $0.075 | $0.45 |
| 1 million tokens | $1.125 | $1.875 | $3.375 | $0.75 | $4.50 |
Uncached. Volume links open interactive pages with adjustable splits.
Prompt caching impact
| Cached share of input | Chat request (4K in / 800 out) | Monthly @ 10K req/day | Savings / month |
|---|---|---|---|
| 0% | $0.0066 | $1,980.00 | — |
| 50% | $0.00525 | $1,575.00 | −$405.00 |
| 75% | $0.004575 | $1,372.50 | −$607.50 |
Monthly projections
| Scale | Requests / day | Chat workload / month | Agentic workload / month* |
|---|---|---|---|
| 100 req/day | 3,000 | $15.75 | $57.60 |
| 1K req/day | 30,000 | $157.50 | $576.00 |
| 10K req/day | 300,000 | $1,575.00 | $5,760.00 |
| 100K req/day | 3,000,000 | $15,750.00 | $57,600.00 |
Chat: 4K in / 800 out, 50% cached. *Agentic: 40K in / 1.5K out, 65% cached. Model your exact mix in the monthly cost calculator.
FAQ
1M purely input tokens cost $0.75; 1M purely output tokens cost $4.50. A typical mixed workload lands in between — see the volume table for exact splits.
A chat-style request (4,000 input + 800 output) costs about $0.00525 with 50% cached input, or $0.0066 without caching.
At 1,000 requests/day of a chat workload (4K in / 800 out, 50% cached): $157.50/month. At 10,000 requests/day: $1,575.00/month. Model your exact mix with the monthly cost calculator.