GLM 4.7 FlashX API Pricing

GLM 4.7 FlashX by Zhipu costs $0.07 per million input tokens and $0.40 per million output tokens.

Price verified from Zhipu's pricing page.

Rates and limits

Input / 1M
$0.07
Output / 1M
$0.40
Cached input / 1M
$0.01
Batch discount
None
Context window
200K tokens
Max output
128K tokens
Released
January 2026

Pricing note: Paid, higher-speed tier of GLM-4.7-Flash on Z.ai. Cached-input storage is listed as "Limited-time Free".

What GLM 4.7 FlashX costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandard
1M750K / 250K$0.15
10M7.5M / 2.5M$1.53
100M75M / 25M$15.25
Estimate your own workload with GLM 4.7 FlashX

Where to run GLM 4.7 FlashX

Prices above are Zhipu's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

GLM 4.7 FlashX pricing FAQ

How much does GLM 4.7 FlashX cost per million tokens?

GLM 4.7 FlashX costs $0.07 per million input tokens and $0.40 per million output tokens, as listed on Zhipu's pricing page on September 30, 2026.

How much does GLM 4.7 FlashX cost for 10 million tokens a month?

About $1.53 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching.

What is the context window of GLM 4.7 FlashX?

GLM 4.7 FlashX accepts up to 200,000 tokens of context and returns up to 128,000 output tokens per request.

Does GLM 4.7 FlashX support prompt caching?

Yes. Cached input tokens cost $0.01 per million, 86% less than the standard input rate.

Is there a batch discount for GLM 4.7 FlashX?

No batch API price is listed for GLM 4.7 FlashX.

Where can I use the GLM 4.7 FlashX API?

GLM 4.7 FlashX is available through Zhipu Open Platform.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison