GLM 5.3 Flash API Pricing
GLM 5.3 Flash by Zhipu costs $0.15 per million input tokens and $0.50 per million output tokens.
Price verified from Zhipu's pricing page.
Rates and limits
- Input / 1M
- $0.15
- Output / 1M
- $0.50
- Cached input / 1M
- $0.03
- Batch discount
- None
- Context window
- 1M tokens
- Max output
- 128K tokens
- Released
- August 2026
Pricing note: Z.ai international USD list price. Cached-input storage is free for a limited time. A faster GLM-5.3-FlashX is $0.37 in / $0.075 cached / $1.25 out.
What GLM 5.3 Flash costs per month
Token volume split 3:1 between input and output, with no prompt caching.
| Tokens per month | Input / output | Standard |
|---|---|---|
| 1M | 750K / 250K | $0.24 |
| 10M | 7.5M / 2.5M | $2.38 |
| 100M | 75M / 25M | $23.75 |
Where to run GLM 5.3 Flash
Prices above are Zhipu's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.
- Zhipu Open PlatformDirect API
- OpenRouterCloud / platform
GLM 5.3 Flash pricing FAQ
How much does GLM 5.3 Flash cost per million tokens?
GLM 5.3 Flash costs $0.15 per million input tokens and $0.50 per million output tokens, as listed on Zhipu's pricing page on September 30, 2026.
How much does GLM 5.3 Flash cost for 10 million tokens a month?
About $2.38 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching.
What is the context window of GLM 5.3 Flash?
GLM 5.3 Flash accepts up to 1,000,000 tokens of context and returns up to 128,000 output tokens per request.
Does GLM 5.3 Flash support prompt caching?
Yes. Cached input tokens cost $0.03 per million, 80% less than the standard input rate.
Is there a batch discount for GLM 5.3 Flash?
No batch API price is listed for GLM 5.3 Flash.
Where can I use the GLM 5.3 Flash API?
GLM 5.3 Flash is available through Zhipu Open Platform and OpenRouter.
Compare with other models
More from Zhipu
Prices shown as input / output per million tokens. All model prices · Side-by-side comparison