Gemini 3.8 Flash API Pricing
Gemini 3.8 Flash by Google costs $1.50 per million input tokens and $7.50 per million output tokens.
Price verified from Google's pricing page.
Rates and limits
- Input / 1M
- $1.50
- Output / 1M
- $7.50
- Cached input / 1M
- $0.15
- Batch discount
- 50% off
- Context window
- 1.05M tokens
- Max output
- 66K tokens
- Released
- September 2026
Pricing note: Standard (post-promotion) rates shown. Promotional rates of $0.75 in / $0.075 cached / $3.75 out apply through December 31, 2026; the listed rates start January 1, 2027. Cache storage $1.00 per 1M tokens per hour ($0.50 during the promotion).
What Gemini 3.8 Flash costs per month
Token volume split 3:1 between input and output, with no prompt caching.
| Tokens per month | Input / output | Standard | Batch API |
|---|---|---|---|
| 1M | 750K / 250K | $3.00 | $1.50 |
| 10M | 7.5M / 2.5M | $30.00 | $15.00 |
| 100M | 75M / 25M | $300.00 | $150.00 |
Where to run Gemini 3.8 Flash
Prices above are Google's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.
- Google AI Studio (Gemini API)Direct API
- OpenRouterCloud / platform
Gemini 3.8 Flash pricing FAQ
How much does Gemini 3.8 Flash cost per million tokens?
Gemini 3.8 Flash costs $1.50 per million input tokens and $7.50 per million output tokens, as listed on Google's pricing page on September 30, 2026.
How much does Gemini 3.8 Flash cost for 10 million tokens a month?
About $30.00 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $15.00 through the batch API.
What is the context window of Gemini 3.8 Flash?
Gemini 3.8 Flash accepts up to 1,048,576 tokens of context and returns up to 65,536 output tokens per request.
Does Gemini 3.8 Flash support prompt caching?
Yes. Cached input tokens cost $0.15 per million, 90% less than the standard input rate.
Is there a batch discount for Gemini 3.8 Flash?
Yes. Requests sent through the batch API cost 50% less: $0.75 per million input tokens and $3.75 per million output tokens.
Where can I use the Gemini 3.8 Flash API?
Gemini 3.8 Flash is available through Google AI Studio (Gemini API) and OpenRouter.
Compare with other models
More from Google
Prices shown as input / output per million tokens. All model prices · Side-by-side comparison