Gemini 2.5 Pro API Pricing

Gemini 2.5 Pro by Google costs $1.25 per million input tokens and $10.00 per million output tokens.

Price verified from Google's pricing page.

Rates and limits

Input / 1M
$1.25
Output / 1M
$10.00
Cached input / 1M
$0.125
Batch discount
50% off
Context window
2M tokens
Max output
64K tokens
Knowledge cutoff
January 2025
Released
March 2025

Pricing note: Prompts over 200K tokens: $2.50 in / $0.25 cached / $15 out. Cache storage $4.50 per 1M tokens per hour.

What Gemini 2.5 Pro costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandardBatch API
1M750K / 250K$3.44$1.72
10M7.5M / 2.5M$34.38$17.19
100M75M / 25M$343.75$171.88
Estimate your own workload with Gemini 2.5 Pro

Where to run Gemini 2.5 Pro

Prices above are Google's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

Benchmarks reported by Google

MMLU
86.3%
HumanEval
91.5%
GPQA
84%
AIME
88%
SWE-bench Verified
63.8%
MMMU
79.6%

Gemini 2.5 Pro pricing FAQ

How much does Gemini 2.5 Pro cost per million tokens?

Gemini 2.5 Pro costs $1.25 per million input tokens and $10.00 per million output tokens, as listed on Google's pricing page on September 30, 2026.

How much does Gemini 2.5 Pro cost for 10 million tokens a month?

About $34.38 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $17.19 through the batch API.

What is the context window of Gemini 2.5 Pro?

Gemini 2.5 Pro accepts up to 2,000,000 tokens of context and returns up to 64,000 output tokens per request.

Does Gemini 2.5 Pro support prompt caching?

Yes. Cached input tokens cost $0.125 per million, 90% less than the standard input rate.

Is there a batch discount for Gemini 2.5 Pro?

Yes. Requests sent through the batch API cost 50% less: $0.625 per million input tokens and $5.00 per million output tokens.

Where can I use the Gemini 2.5 Pro API?

Gemini 2.5 Pro is available through Google AI Studio (Gemini API), Google Cloud Vertex AI, and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison