GLM 5.3 Flash API Pricing

GLM 5.3 Flash by Zhipu costs $0.15 per million input tokens and $0.50 per million output tokens.

Price verified from Zhipu's pricing page.

Rates and limits

Input / 1M
$0.15
Output / 1M
$0.50
Cached input / 1M
$0.03
Batch discount
None
Context window
1M tokens
Max output
128K tokens
Released
August 2026

Pricing note: Z.ai international USD list price. Cached-input storage is free for a limited time. A faster GLM-5.3-FlashX is $0.37 in / $0.075 cached / $1.25 out.

What GLM 5.3 Flash costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandard
1M750K / 250K$0.24
10M7.5M / 2.5M$2.38
100M75M / 25M$23.75
Estimate your own workload with GLM 5.3 Flash

Where to run GLM 5.3 Flash

Prices above are Zhipu's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

GLM 5.3 Flash pricing FAQ

How much does GLM 5.3 Flash cost per million tokens?

GLM 5.3 Flash costs $0.15 per million input tokens and $0.50 per million output tokens, as listed on Zhipu's pricing page on September 30, 2026.

How much does GLM 5.3 Flash cost for 10 million tokens a month?

About $2.38 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching.

What is the context window of GLM 5.3 Flash?

GLM 5.3 Flash accepts up to 1,000,000 tokens of context and returns up to 128,000 output tokens per request.

Does GLM 5.3 Flash support prompt caching?

Yes. Cached input tokens cost $0.03 per million, 80% less than the standard input rate.

Is there a batch discount for GLM 5.3 Flash?

No batch API price is listed for GLM 5.3 Flash.

Where can I use the GLM 5.3 Flash API?

GLM 5.3 Flash is available through Zhipu Open Platform and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison