GLM 4.7 FlashX API Pricing
GLM 4.7 FlashX by Zhipu costs $0.07 per million input tokens and $0.40 per million output tokens.
Price verified from Zhipu's pricing page.
Rates and limits
- Input / 1M
- $0.07
- Output / 1M
- $0.40
- Cached input / 1M
- $0.01
- Batch discount
- None
- Context window
- 200K tokens
- Max output
- 128K tokens
- Released
- January 2026
Pricing note: Paid, higher-speed tier of GLM-4.7-Flash on Z.ai. Cached-input storage is listed as "Limited-time Free".
What GLM 4.7 FlashX costs per month
Token volume split 3:1 between input and output, with no prompt caching.
| Tokens per month | Input / output | Standard |
|---|---|---|
| 1M | 750K / 250K | $0.15 |
| 10M | 7.5M / 2.5M | $1.53 |
| 100M | 75M / 25M | $15.25 |
Where to run GLM 4.7 FlashX
Prices above are Zhipu's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.
- Zhipu Open PlatformDirect API
GLM 4.7 FlashX pricing FAQ
How much does GLM 4.7 FlashX cost per million tokens?
GLM 4.7 FlashX costs $0.07 per million input tokens and $0.40 per million output tokens, as listed on Zhipu's pricing page on September 30, 2026.
How much does GLM 4.7 FlashX cost for 10 million tokens a month?
About $1.53 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching.
What is the context window of GLM 4.7 FlashX?
GLM 4.7 FlashX accepts up to 200,000 tokens of context and returns up to 128,000 output tokens per request.
Does GLM 4.7 FlashX support prompt caching?
Yes. Cached input tokens cost $0.01 per million, 86% less than the standard input rate.
Is there a batch discount for GLM 4.7 FlashX?
No batch API price is listed for GLM 4.7 FlashX.
Where can I use the GLM 4.7 FlashX API?
GLM 4.7 FlashX is available through Zhipu Open Platform.
Compare with other models
More from Zhipu
Prices shown as input / output per million tokens. All model prices · Side-by-side comparison