Gemini 2.5 Flash Lite API Pricing

Gemini 2.5 Flash Lite by Google costs $0.10 per million input tokens and $0.40 per million output tokens.

Price verified from Google's pricing page.

Rates and limits

Input / 1M
$0.10
Output / 1M
$0.40
Cached input / 1M
$0.01
Batch discount
50% off
Context window
1.05M tokens
Released
September 2025

Pricing note: Audio input is $0.30 (cached $0.03). Cache storage $1.00 per 1M tokens per hour.

What Gemini 2.5 Flash Lite costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandardBatch API
1M750K / 250K$0.18$0.09
10M7.5M / 2.5M$1.75$0.88
100M75M / 25M$17.50$8.75
Estimate your own workload with Gemini 2.5 Flash Lite

Where to run Gemini 2.5 Flash Lite

Prices above are Google's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

Gemini 2.5 Flash Lite pricing FAQ

How much does Gemini 2.5 Flash Lite cost per million tokens?

Gemini 2.5 Flash Lite costs $0.10 per million input tokens and $0.40 per million output tokens, as listed on Google's pricing page on September 30, 2026.

How much does Gemini 2.5 Flash Lite cost for 10 million tokens a month?

About $1.75 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $0.88 through the batch API.

What is the context window of Gemini 2.5 Flash Lite?

Gemini 2.5 Flash Lite accepts up to 1,048,576 tokens of context.

Does Gemini 2.5 Flash Lite support prompt caching?

Yes. Cached input tokens cost $0.01 per million, 90% less than the standard input rate.

Is there a batch discount for Gemini 2.5 Flash Lite?

Yes. Requests sent through the batch API cost 50% less: $0.05 per million input tokens and $0.20 per million output tokens.

Where can I use the Gemini 2.5 Flash Lite API?

Gemini 2.5 Flash Lite is available through Google AI Studio (Gemini API), Google Cloud Vertex AI, and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison