Gemini 2.5 Flash API Pricing

Gemini 2.5 Flash by Google costs $0.30 per million input tokens and $2.50 per million output tokens.

Price verified from Google's pricing page.

Rates and limits

Input / 1M
$0.30
Output / 1M
$2.50
Cached input / 1M
$0.03
Batch discount
50% off
Context window
1M tokens
Max output
64K tokens
Knowledge cutoff
January 2025
Released
April 2025

Pricing note: Audio input is $1.00 (cached $0.10). Cache storage $1.00 per 1M tokens per hour.

What Gemini 2.5 Flash costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandardBatch API
1M750K / 250K$0.85$0.43
10M7.5M / 2.5M$8.50$4.25
100M75M / 25M$85.00$42.50
Estimate your own workload with Gemini 2.5 Flash

Where to run Gemini 2.5 Flash

Prices above are Google's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

Benchmarks reported by Google

MMLU
80.9%
HumanEval
87.8%
GPQA
67.5%
AIME
72%
SWE-bench Verified
48.9%
MMMU
76.7%

Gemini 2.5 Flash pricing FAQ

How much does Gemini 2.5 Flash cost per million tokens?

Gemini 2.5 Flash costs $0.30 per million input tokens and $2.50 per million output tokens, as listed on Google's pricing page on September 30, 2026.

How much does Gemini 2.5 Flash cost for 10 million tokens a month?

About $8.50 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $4.25 through the batch API.

What is the context window of Gemini 2.5 Flash?

Gemini 2.5 Flash accepts up to 1,000,000 tokens of context and returns up to 64,000 output tokens per request.

Does Gemini 2.5 Flash support prompt caching?

Yes. Cached input tokens cost $0.03 per million, 90% less than the standard input rate.

Is there a batch discount for Gemini 2.5 Flash?

Yes. Requests sent through the batch API cost 50% less: $0.15 per million input tokens and $1.25 per million output tokens.

Where can I use the Gemini 2.5 Flash API?

Gemini 2.5 Flash is available through Google AI Studio (Gemini API), Google Cloud Vertex AI, and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison