GLM-4-Plus API Pricing

GLM-4-Plus by Zhipu costs $0.745 per million input tokens and $0.745 per million output tokens.

Price verified from Zhipu's pricing page.

Rates and limits

Input / 1M
$0.745
Output / 1M
$0.745
Cached input / 1M
$0.373
Batch discount
None
Context window
128K tokens
Max output
8K tokens
Knowledge cutoff
June 2024
Released
June 2024

Pricing note: Sold only on Zhipu BigModel (China), priced in CNY: ¥5 input / ¥5 output / ¥2.5 cache hit per 1M tokens. Converted at 6.7110 CNY per USD (US Federal Reserve H.10, rate for 2026-09-25, released 2026-09-28). Not listed on the international Z.ai platform.

What GLM-4-Plus costs per month

Token volume split 3:1 between input and output, with no prompt caching.

Tokens per monthInput / outputStandard
1M750K / 250K$0.75
10M7.5M / 2.5M$7.45
100M75M / 25M$74.50
Estimate your own workload with GLM-4-Plus

Where to run GLM-4-Plus

Prices above are Zhipu's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.

Benchmarks reported by Zhipu

MMLU
81.2%
HumanEval
78.4%

GLM-4-Plus pricing FAQ

How much does GLM-4-Plus cost per million tokens?

GLM-4-Plus costs $0.745 per million input tokens and $0.745 per million output tokens, as listed on Zhipu's pricing page on September 30, 2026.

How much does GLM-4-Plus cost for 10 million tokens a month?

About $7.45 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching.

What is the context window of GLM-4-Plus?

GLM-4-Plus accepts up to 128,000 tokens of context and returns up to 8,192 output tokens per request.

Does GLM-4-Plus support prompt caching?

Yes. Cached input tokens cost $0.373 per million, 50% less than the standard input rate.

Is there a batch discount for GLM-4-Plus?

No batch API price is listed for GLM-4-Plus.

Where can I use the GLM-4-Plus API?

GLM-4-Plus is available through Zhipu Open Platform and OpenRouter.

Prices shown as input / output per million tokens. All model prices · Side-by-side comparison