gpt-4.1 · OpenAI
The long-context veteran with a 1.05M-token window, still listed for existing production pipelines.
Last checked Aug 16, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Confirm the current rate at the official provider source.
Input
$2.00
per 1M tokens
Output
$8.00
per 1M tokens
Cached input
$0.50
per 1M tokens
Context window
1M
max output 32.8K
GPT-4.1 cost calculator
What GPT-4.1 costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.006125 | $183.75 |
| RAG / search-augmented answers | 8,000 | 500 | $0.014 | $420.00 |
| AI coding assistant | 12,000 | 2,000 | $0.0292 | $876.00 |
| Document summarization | 25,000 | 600 | $0.0511 | $1,531.50 |
| Agentic workflow | 40,000 | 1,500 | $0.053 | $1,590.00 |
| Content generation | 800 | 1,200 | $0.0108 | $325.20 |
| Data extraction & tagging | 2,000 | 250 | $0.0054 | $162.00 |
| Translation | 5,000 | 5,500 | $0.0529 | $1,586.25 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| GPT-4.1 | $2.00 | $8.00 | 1M | $0.0114 |
| GPT-5.4compare | $2.50 | $15.00 | 1.1M | $0.0175 |
| Gemini 3.6 Flash | $1.50 | $7.50 | 1M | $0.0093 |
| GPT-5.2 | $1.75 | $14.00 | 400K | $0.0151 |
*4,000 in + 800 out tokens, 50% cached input where available.
Related calculations
FAQ
GPT-4.1 costs $2.00 per 1M input tokens and $8.00 per 1M output tokens, with cached input at $0.50 per 1M tokens. Verified Aug 16, 2026 against https://developers.openai.com/api/docs/pricing.
GPT-4.1 supports a 1,047,576-token context window with up to 32,768 output tokens per request, tokenized with o200k_base.
At $2.00/M input and $8.00/M output, GPT-4.1 sits below GPT-5.4 ($2.50/M in, $15.00/M out) — see the comparison table for full-workload differences.