GLM 4.7 Flash API Pricing

GLM 4.7 Flash by Zhipu costs $0.00 per million input tokens and $0.00 per million output tokens.

Price verified from Zhipu's pricing page.

Rates and limits

Input / 1M
$0.00
Output / 1M
$0.00
Cached input / 1M
$0.00
Batch discount
None
Context window
203K tokens
Released
January 2026

Pricing note: Free tier: Z.ai lists GLM-4.7-Flash input, cached input and output all as "Free". The pricing page states no rate limits or usage terms for it. The paid, faster tier is GLM-4.7-FlashX (listed separately).

What GLM 4.7 Flash costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandard
1M750K / 250K$0.00
10M7.5M / 2.5M$0.00
100M75M / 25M$0.00
Estimate your own workload with GLM 4.7 Flash

Where to run GLM 4.7 Flash

Prices above are Zhipu's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

GLM 4.7 Flash pricing FAQ

How much does GLM 4.7 Flash cost per million tokens?

GLM 4.7 Flash costs $0.00 per million input tokens and $0.00 per million output tokens, as listed on Zhipu's pricing page on September 30, 2026.

How much does GLM 4.7 Flash cost for 10 million tokens a month?

About $0.00 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching.

What is the context window of GLM 4.7 Flash?

GLM 4.7 Flash accepts up to 202,752 tokens of context.

Does GLM 4.7 Flash support prompt caching?

Yes. Cached input tokens cost $0.00 per million, NaN% less than the standard input rate.

Is there a batch discount for GLM 4.7 Flash?

No batch API price is listed for GLM 4.7 Flash.

Where can I use the GLM 4.7 Flash API?

GLM 4.7 Flash is available through Zhipu Open Platform and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison