Gemini 3.x and 2.5 families: Pro with tiered long-context pricing, Flash for volume, Flash-Lite for the cheapest usable tier — plus explicit context-caching rates.
7 models tracked · cheapest input: Gemini 2.5 Flash-Lite at $0.10/M · official pricing page
gemini-3.1-pro
Input /M
$2.00
Output /M
$12.00
Context
1M
Google's flagship with tiered long-context pricing: $2/$12 up to 200K prompt tokens, doubling above.
gemini-3.6-flash
Input /M
$1.50
Output /M
$7.50
Context
1M
The newest Flash at $1.50/$7.50 — flagship-adjacent quality at half of 3.1 Pro pricing.
gemini-3.5-flash
Input /M
$1.50
Output /M
$9.00
Context
1M
Proven 3.5 Flash generation at $1.50/$9.00 — still listed alongside its 3.6 successor.
gemini-3.5-flash-lite
Input /M
$0.30
Output /M
$2.50
Context
1M
The cheap 3.5 tier at $0.30/$2.50 — classification and summarization at scale.
gemini-2.5-flash
Input /M
$0.30
Output /M
$2.50
Context
1M
The 2025 volume workhorse at $0.30/$2.50, still listed with a 1M-token window.
gemini-2.5-flash-lite
Input /M
$0.10
Output /M
$0.40
Context
1M
The absolute cheapest Gemini tier at $0.10/$0.40 — still a going concern for sub-cent pipelines.
gemini-2.5-pro
Input /M
$1.25
Output /M
$10.00
Context
1M
The 2025 flagship at $1.25/$10 (≤200K), superseded by 3.1 Pro but still listed and widely deployed.
Other providers
All Google prices verified Aug 16, 2026. Prices change frequently — each model page links to the authoritative source.