GPT-4.1 Mini API Pricing

GPT-4.1 Mini by OpenAI costs $0.40 per million input tokens and $1.60 per million output tokens.

Price verified from OpenAI's pricing page.

Rates and limits

Input / 1M
$0.40
Output / 1M
$1.60
Cached input / 1M
$0.10
Batch discount
50% off
Context window
1.05M tokens
Max output
16K tokens
Knowledge cutoff
October 2024
Released
April 2025

What GPT-4.1 Mini costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandardBatch API
1M750K / 250K$0.70$0.35
10M7.5M / 2.5M$7.00$3.50
100M75M / 25M$70.00$35.00
Estimate your own workload with GPT-4.1 Mini

Where to run GPT-4.1 Mini

Prices above are OpenAI's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

Benchmarks reported by OpenAI

MMLU
83%
HumanEval
88%
GPQA
45%
MMMU
65%

GPT-4.1 Mini pricing FAQ

How much does GPT-4.1 Mini cost per million tokens?

GPT-4.1 Mini costs $0.40 per million input tokens and $1.60 per million output tokens, as listed on OpenAI's pricing page on September 30, 2026.

How much does GPT-4.1 Mini cost for 10 million tokens a month?

About $7.00 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching, or $3.50 through the batch API.

What is the context window of GPT-4.1 Mini?

GPT-4.1 Mini accepts up to 1,047,576 tokens of context and returns up to 16,384 output tokens per request.

Does GPT-4.1 Mini support prompt caching?

Yes. Cached input tokens cost $0.10 per million, 75% less than the standard input rate.

Is there a batch discount for GPT-4.1 Mini?

Yes. Requests sent through the batch API cost 50% less: $0.20 per million input tokens and $0.80 per million output tokens.

Where can I use the GPT-4.1 Mini API?

GPT-4.1 Mini is available through OpenAI API, Azure OpenAI Service, and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison