gemini-3.5-flash · Google
Google multimodal Flash route for reasoning, coding, and high-volume agent workloads.
Last checked Aug 26, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Note: Live OpenRouter route snapshot checked 26 Aug 2026. Google AI API tier, promotion, and availability may differ. Confirm the current rate at the official provider source.
Quick answer
At the published base rate, Gemini 3.5 Flash (OpenRouter) costs $1.50 per 1M input tokens and $9.00 per 1M output tokens. It supports a 1M context window and is currently marked ga.
Method & trust
This rate card uses the provider's listed base input, output, and cached-input prices. Workload tables below apply those rates to explicit token counts, cache assumptions, and request volumes; verify provider-specific tiers before committing budget.
Input
$1.50
per 1M tokens
Output
$9.00
per 1M tokens
Cached input
$0.15
per 1M tokens
Context window
1M
max output 65.5K
Gemini 3.5 Flash (OpenRouter) cost calculator
What Gemini 3.5 Flash (OpenRouter) costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.005093 | $152.775 |
| RAG / search-augmented answers | 8,000 | 500 | $0.0111 | $333.00 |
| AI coding assistant | 12,000 | 2,000 | $0.0263 | $788.40 |
| Document summarization | 25,000 | 600 | $0.0395 | $1,185.75 |
| Agentic workflow | 40,000 | 1,500 | $0.0384 | $1,152.00 |
| Content generation | 800 | 1,200 | $0.0117 | $350.28 |
| Data extraction & tagging | 2,000 | 250 | $0.00471 | $141.30 |
| Translation | 5,000 | 5,500 | $0.056 | $1,679.625 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Benchmark scores
Third-party composite indices (0-100) via OpenRouter's model catalog, verified Aug 28, 2026. Coding per $$9.00 of output: 7.8.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| Gemini 3.5 Flash (OpenRouter) | $1.50 | $9.00 | 1M | $0.0105 |
| GPT-5.6 Solcompare | $5.00 | $30.00 | 1.1M | $0.035 |
| GPT-5.6 Terracompare | $2.00 | $12.00 | 1.1M | $0.014 |
| GPT-5.6 Lunacompare | $0.20 | $1.20 | 1.1M | $0.0014 |
| GPT-5.6 Cybercompare | $12.50 | $75.00 | 1.1M | $0.0875 |
| GPT-5.5 Standardcompare | $5.00 | $30.00 | 512K | $0.035 |
| GPT-5.5 Procompare | $30.00 | $180.00 | 512K | $0.21 |
| GPT-5.4 Workhorsecompare | $2.50 | $15.00 | 256K | $0.0175 |
| GPT-5.4 minicompare | $0.75 | $4.50 | 256K | $0.00525 |
*4,000 in + 800 out tokens, 50% cached input where available.
Related calculations
FAQ
Gemini 3.5 Flash (OpenRouter) costs $1.50 per 1M input tokens and $9.00 per 1M output tokens, with cached input at $0.15 per 1M tokens. Verified Aug 26, 2026 against https://openrouter.ai/api/v1/models.
Gemini 3.5 Flash (OpenRouter) supports a 1,048,576-token context window with up to 65,536 output tokens per request, tokenized with sentencepiece.
At $1.50/M input and $9.00/M output, Gemini 3.5 Flash (OpenRouter) sits below GPT-5.6 Sol ($5.00/M in, $30.00/M out) — see the comparison table for full-workload differences.