Gemini 3.8 Flash API Pricing

Gemini 3.8 Flash by Google costs $1.50 per million input tokens and $7.50 per million output tokens.

Price verified from Google's pricing page.

Rates and limits

Input / 1M
$1.50
Output / 1M
$7.50
Cached input / 1M
$0.15
Batch discount
50% off
Context window
1.05M tokens
Max output
66K tokens
Released
September 2026

Pricing note: Standard (post-promotion) rates shown. Promotional rates of $0.75 in / $0.075 cached / $3.75 out apply through December 31, 2026; the listed rates start January 1, 2027. Cache storage $1.00 per 1M tokens per hour ($0.50 during the promotion).

What Gemini 3.8 Flash costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandardBatch API
1M750K / 250K$3.00$1.50
10M7.5M / 2.5M$30.00$15.00
100M75M / 25M$300.00$150.00
Estimate your own workload with Gemini 3.8 Flash

Where to run Gemini 3.8 Flash

Prices above are Google's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

Gemini 3.8 Flash pricing FAQ

How much does Gemini 3.8 Flash cost per million tokens?

Gemini 3.8 Flash costs $1.50 per million input tokens and $7.50 per million output tokens, as listed on Google's pricing page on September 30, 2026.

How much does Gemini 3.8 Flash cost for 10 million tokens a month?

About $30.00 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $15.00 through the batch API.

What is the context window of Gemini 3.8 Flash?

Gemini 3.8 Flash accepts up to 1,048,576 tokens of context and returns up to 65,536 output tokens per request.

Does Gemini 3.8 Flash support prompt caching?

Yes. Cached input tokens cost $0.15 per million, 90% less than the standard input rate.

Is there a batch discount for Gemini 3.8 Flash?

Yes. Requests sent through the batch API cost 50% less: $0.75 per million input tokens and $3.75 per million output tokens.

Where can I use the Gemini 3.8 Flash API?

Gemini 3.8 Flash is available through Google AI Studio (Gemini API) and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison