Up to $350K in Cloud credits
Google for Startups AI cloud credits
Current introductory standard and Batch API pricing, published January 2027 rates, token limits and workload calculators for Gemini 3.8 Flash.
Introductory pricing through Dec 31, 2026. Pricing page fallback verified September 3, 2026; client checks TokenCOGS live registry on load.
| Input / 1M | $0.75 |
| Output / 1M | $3.75 |
| Cached input / 1M | $0.075 |
| Input context limit | 1,048,576 tokens |
| Max output | 65,536 tokens |
| Released | Sep 2, 2026 |
| Batch input / 1M | $0.375 |
| Batch output / 1M | $1.875 |
| Batch cached input / 1M | $0.0375 |
| Equivalent discount | 50% |
| Through Dec 31, 2026 | From Jan 1, 2027 | |
|---|---|---|
| Input | $0.75 | $1.50 |
| Output | $3.75 | $7.50 |
| Cached input | $0.075 | $0.15 |
| Batch input | $0.375 | $0.75 |
| Batch output | $1.875 | $3.75 |
| Batch cached | $0.0375 | $0.075 |
TokenCOGS uses the existing date-aware priceSchedule resolver, so calculators use the introductory rate before Jan 1, 2027 and the published next rate on or after that date.
Example: 2,000 input tokens + 500 output tokens per request.
| Volume | Estimated API cost | Approx. / request |
|---|---|---|
| 1,000 requests | $3.38 | $0.0034 |
| 100,000 requests | $337.50 | $0.0034 |
| 1,000,000 requests | $3,375.00 | $0.0034 |
Gemini 3.8 Flash is the current Flash model in the TokenCOGS registry. Gemini 3.7 Flash remains available for existing workloads and comparisons.
Price alone is not a quality, latency or reliability ranking. Benchmark the exact task you plan to run.
Compare all Google models →Google for Startups AI cloud credits
Through December 31, 2026, the current TokenCOGS schedule is $0.75 input, $3.75 output and $0.075 cached input per 1M tokens.
Google’s published next pricing takes effect January 1, 2027: $1.50 input, $7.50 output and $0.15 cached input per 1M tokens. Batch pricing changes to $0.75 input, $3.75 output and $0.075 cached input.