GPT-4o mini API Pricing

GPT-4o mini by OpenAI costs $0.15 per million input tokens and $0.60 per million output tokens.

Price verified from OpenAI's pricing page.

Rates and limits

Input / 1M
$0.15
Output / 1M
$0.60
Cached input / 1M
$0.075
Batch discount
50% off
Context window
128K tokens
Max output
16K tokens
Knowledge cutoff
October 2023
Released
July 2024

What GPT-4o mini costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandardBatch API
1M750K / 250K$0.26$0.13
10M7.5M / 2.5M$2.63$1.31
100M75M / 25M$26.25$13.13
Estimate your own workload with GPT-4o mini

Where to run GPT-4o mini

Prices above are OpenAI's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

Benchmarks reported by OpenAI

MMLU
82%
HumanEval
87.2%
GPQA
40.2%
MMMU
59.4%

GPT-4o mini pricing FAQ

How much does GPT-4o mini cost per million tokens?

GPT-4o mini costs $0.15 per million input tokens and $0.60 per million output tokens, as listed on OpenAI's pricing page on September 30, 2026.

How much does GPT-4o mini cost for 10 million tokens a month?

About $2.63 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $1.31 through the batch API.

What is the context window of GPT-4o mini?

GPT-4o mini accepts up to 128,000 tokens of context and returns up to 16,384 output tokens per request.

Does GPT-4o mini support prompt caching?

Yes. Cached input tokens cost $0.075 per million, 50% less than the standard input rate.

Is there a batch discount for GPT-4o mini?

Yes. Requests sent through the batch API cost 50% less: $0.075 per million input tokens and $0.30 per million output tokens.

Where can I use the GPT-4o mini API?

GPT-4o mini is available through OpenAI API, Azure OpenAI Service, Azure AI Foundry, and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison