gpt-5.6-luna · OpenAI
The cheap 5.6 tier at $0.20/$1.20 — classification, routing and high-volume features with frontier-family latency.
Last checked Aug 16, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Note: Long-context requests bill at 2× ($0.40/$1.80). Confirm the current rate at the official provider source.
Input
$0.20
per 1M tokens
Output
$1.20
per 1M tokens
Cached input
$0.02
per 1M tokens
Context window
1.1M
max output 128K
GPT-5.6 Luna cost calculator
What GPT-5.6 Luna costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.000679 | $20.37 |
| RAG / search-augmented answers | 8,000 | 500 | $0.00148 | $44.40 |
| AI coding assistant | 12,000 | 2,000 | $0.003504 | $105.12 |
| Document summarization | 25,000 | 600 | $0.00527 | $158.10 |
| Agentic workflow | 40,000 | 1,500 | $0.00512 | $153.60 |
| Content generation | 800 | 1,200 | $0.001557 | $46.704 |
| Data extraction & tagging | 2,000 | 250 | $0.000628 | $18.84 |
| Translation | 5,000 | 5,500 | $0.007465 | $223.95 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| GPT-5.6 Luna | $0.20 | $1.20 | 1.1M | $0.0014 |
| Gemini 3.5 Flash-Litecompare | $0.30 | $2.50 | 1M | $0.00266 |
| Claude Haiku 4.5compare | $1.00 | $5.00 | 200K | $0.0062 |
| Mistral Small 4 | $0.15 | $0.60 | 131.1K | $0.00081 |
*4,000 in + 800 out tokens, 50% cached input where available.
Related calculations
FAQ
GPT-5.6 Luna costs $0.20 per 1M input tokens and $1.20 per 1M output tokens, with cached input at $0.02 per 1M tokens. Verified Aug 16, 2026 against https://developers.openai.com/api/docs/pricing.
GPT-5.6 Luna supports a 1,050,000-token context window with up to 128,000 output tokens per request, tokenized with o200k_base.
At $0.20/M input and $1.20/M output, GPT-5.6 Luna sits below Gemini 3.5 Flash-Lite ($0.30/M in, $2.50/M out) — see the comparison table for full-workload differences.