Gemini 2.5 Flash Lite API Pricing
Gemini 2.5 Flash Lite by Google costs $0.10 per million input tokens and $0.40 per million output tokens.
Price verified from Google's pricing page.
Rates and limits
- Input / 1M
- $0.10
- Output / 1M
- $0.40
- Cached input / 1M
- $0.01
- Batch discount
- 50% off
- Context window
- 1.05M tokens
- Released
- September 2025
Pricing note: Audio input is $0.30 (cached $0.03). Cache storage $1.00 per 1M tokens per hour.
What Gemini 2.5 Flash Lite costs per month
Token volume split 3:1 between input and output, with no prompt caching.
| Tokens per month | Input / output | Standard | Batch API |
|---|---|---|---|
| 1M | 750K / 250K | $0.18 | $0.09 |
| 10M | 7.5M / 2.5M | $1.75 | $0.88 |
| 100M | 75M / 25M | $17.50 | $8.75 |
Where to run Gemini 2.5 Flash Lite
Prices above are Google's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.
- Google AI Studio (Gemini API)Direct API
- Google Cloud Vertex AICloud / platform
- OpenRouterCloud / platform
Gemini 2.5 Flash Lite pricing FAQ
How much does Gemini 2.5 Flash Lite cost per million tokens?
Gemini 2.5 Flash Lite costs $0.10 per million input tokens and $0.40 per million output tokens, as listed on Google's pricing page on September 30, 2026.
How much does Gemini 2.5 Flash Lite cost for 10 million tokens a month?
About $1.75 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $0.88 through the batch API.
What is the context window of Gemini 2.5 Flash Lite?
Gemini 2.5 Flash Lite accepts up to 1,048,576 tokens of context.
Does Gemini 2.5 Flash Lite support prompt caching?
Yes. Cached input tokens cost $0.01 per million, 90% less than the standard input rate.
Is there a batch discount for Gemini 2.5 Flash Lite?
Yes. Requests sent through the batch API cost 50% less: $0.05 per million input tokens and $0.20 per million output tokens.
Where can I use the Gemini 2.5 Flash Lite API?
Gemini 2.5 Flash Lite is available through Google AI Studio (Gemini API), Google Cloud Vertex AI, and OpenRouter.
Compare with other models
More from Google
Prices shown as input / output per million tokens. All model prices · Side-by-side comparison