gemini-2.5-flash · Google
The 2025 volume workhorse at $0.30/$2.50, still listed with a 1M-token window.
Last checked Aug 16, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Confirm the current rate at the official provider source.
Input
$0.30
per 1M tokens
Output
$2.50
per 1M tokens
Cached input
$0.03
per 1M tokens
Context window
1M
max output 65.5K
Gemini 2.5 Flash cost calculator
What Gemini 2.5 Flash costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.001263 | $37.905 |
| RAG / search-augmented answers | 8,000 | 500 | $0.00257 | $77.10 |
| AI coding assistant | 12,000 | 2,000 | $0.006656 | $199.68 |
| Document summarization | 25,000 | 600 | $0.008325 | $249.75 |
| Agentic workflow | 40,000 | 1,500 | $0.00873 | $261.90 |
| Content generation | 800 | 1,200 | $0.003175 | $95.256 |
| Data extraction & tagging | 2,000 | 250 | $0.001117 | $33.51 |
| Translation | 5,000 | 5,500 | $0.015 | $451.425 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| Gemini 2.5 Flash | $0.30 | $2.50 | 1M | $0.00266 |
| Gemini 3.5 Flash-Lite | $0.30 | $2.50 | 1M | $0.00266 |
| GPT-5.4 mini | $0.75 | $4.50 | 1.1M | $0.00525 |
| Claude Haiku 4.5 | $1.00 | $5.00 | 200K | $0.0062 |
*4,000 in + 800 out tokens, 50% cached input where available.
Related calculations
FAQ
Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, with cached input at $0.03 per 1M tokens. Verified Aug 16, 2026 against https://ai.google.dev/gemini-api/docs/pricing.
Gemini 2.5 Flash supports a 1,048,576-token context window with up to 65,536 output tokens per request, tokenized with Gemini SentencePiece (~262K vocab).
At $0.30/M input and $2.50/M output, Gemini 2.5 Flash sits below Gemini 3.5 Flash-Lite ($0.30/M in, $2.50/M out) — see the comparison table for full-workload differences.