Llama 4 Maverick (400B MoE) charges $0.45 per 1M input tokens and $1.25 per 1M output tokens. Real bills depend on your input/output split — the tables below price Llama 4 Maverick (400B MoE) at 1K–1M tokens across common workload shapes. Verified Aug 19, 2026
Cost by token volume
| Volume | Input-heavy (90/10) | Balanced (70/30) | Output-heavy (30/70) | All input | All output |
|---|---|---|---|---|---|
| 1,000 tokens | $0.00053 | $0.00069 | $0.00101 | $0.00045 | $0.00125 |
| 10,000 tokens | $0.0053 | $0.0069 | $0.0101 | $0.0045 | $0.0125 |
| 100,000 tokens | $0.053 | $0.069 | $0.101 | $0.045 | $0.125 |
| 1 million tokens | $0.53 | $0.69 | $1.01 | $0.45 | $1.25 |
Uncached. Volume links open interactive pages with adjustable splits.
Monthly projections
| Scale | Requests / day | Chat workload / month | Agentic workload / month* |
|---|---|---|---|
| 100 req/day | 3,000 | $8.40 | $59.625 |
| 1K req/day | 30,000 | $84.00 | $596.25 |
| 10K req/day | 300,000 | $840.00 | $5,962.50 |
| 100K req/day | 3,000,000 | $8,400.00 | $59,625.00 |
Chat: 4K in / 800 out, 50% cached. *Agentic: 40K in / 1.5K out, 65% cached. Model your exact mix in the monthly cost calculator.
FAQ
1M purely input tokens cost $0.45; 1M purely output tokens cost $1.25. A typical mixed workload lands in between — see the volume table for exact splits.
A chat-style request (4,000 input + 800 output) costs about $0.0028 with 50% cached input, or $0.0028 without caching.
At 1,000 requests/day of a chat workload (4K in / 800 out, 50% cached): $84.00/month. At 10,000 requests/day: $840.00/month. Model your exact mix with the monthly cost calculator.