50% promotional credit
50% billing credit on Gemini 3.6 / 3.7 / 3.8 Flash
Compare Gemini API input, output and cached-token pricing across Flash, Pro and preview models, then run the same workload through TokenCOGS calculators.
Start with the models developers are most likely to compare, then use the full table for the rest of the Gemini lineup.
USD per 1M text tokens for Google Gemini models. Compare standard input, output and cached-token rates before modeling your workload.
| Model | Input | Output | Cached input | Context | |
|---|---|---|---|---|---|
| Gemini 3.8 Flash | $0.75 | $3.75 | $0.075 | 1.048576M tokens | Details → |
| Gemini 3.7 FlashCurrent / previous generation | $0.75 | $3.75 | $0.075 | 1M tokens | Details → |
| Gemini 3.6 Flash | $0.75 | $3.75 | $0.075 | 1M tokens | Calculate → |
| Gemini 3.5 Flash | $1.5 | $9 | $0.15 | Not listed in registry | Calculate → |
| Gemini 3.5 Flash-Lite | $0.3 | $2.5 | $0.03 | Not listed in registry | Calculate → |
| Gemini 3.1 Pro Preview Preview | $2 | $12 | $0.2 | 1M tokens | Details → |
| Gemini 3.1 Flash-Lite | $0.25 | $1.5 | $0.025 | Not listed in registry | Calculate → |
| Gemini 2.5 Pro | $1.25 | $10 | $0.125 | 1M tokens | Details → |
| Gemini 2.5 Flash | $0.3 | $2.5 | $0.03 | 1M tokens | Calculate → |
| Gemini 2.5 Flash-Lite | $0.1 | $0.4 | $0.01 | Not listed in registry | Calculate → |
50% billing credit on Gemini 3.6 / 3.7 / 3.8 Flash
Google for Startups AI cloud credits
Use this table to eliminate models that do not fit your unit economics, then benchmark the remaining candidates on your own task for output quality, latency, rate limits and operational reliability.
Gemini 2.5 Flash-Lite currently has the lowest tracked input-token price in this Google set. That does not automatically make it the lowest-cost model for every workload because output ratios, caching, long-context rules and tool usage can change the result.