GLM-4.7 — API Pricing & Benchmarks

GLM-4.7 is Zhipu AI (Z.ai)'s mid reasoning model. The API costs $0.60 per 1M input tokens and $2.20 per 1M output tokens ($1.00 blended at a 3:1 ratio), with cached input at $0.11. On the independently measured Artificial Analysis Intelligence Index it scores 34.5, ranking #47 of 68 models tracked on this site, with a median output speed of 113 tokens/sec.

Explore provider routes and additional benchmark sources for GLM-4.7

ProviderZhipu AI (Z.ai)
Input price / 1M tokens$0.60
Output price / 1M tokens$2.20
Cached input / 1M tokens$0.11
Blended price (3:1)$1.00
Context window
Max output tokens128K
Knowledge cutoff
Licenseopen weights
Input modalitiesT
Output speed (tok/s)113
Released
Artificial Analysis Intelligence Index34.5
LMArena Elo1436
GPQA Diamond85.9%
Humanity's Last Exam27.4%

Where GLM-4.7 fits

On a blended 3:1 basis GLM-4.7 costs $1.00 per 1M tokens, cheaper than 64% of the 80 priced models in this index. It sits in the mid-tier bracket where most production chat, RAG and summarization workloads run: meaningfully cheaper than flagships while keeping most of their capability. The weights are open, so it can be self-hosted for data-residency or cost reasons instead of consumed via API. Cached input is priced at $0.11 (82% below standard input), which rewards prompt structures with long stable prefixes — see our caching guide.

Price history

PeriodInput $/1MOutput $/1MCached $/1M
2026-07-18 → current$0.60$2.20$0.11

No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).

Alternatives at a similar price

About Zhipu AI (Z.ai)

Zhipu AI (Z.ai) publishes the GLM series as open weights. GLM-5.3 benchmarks within a few points of the proprietary frontier at roughly a third of the price, and GLM-5.3-Flash keeps most of that score at small-model pricing.

More Zhipu AI (Z.ai) models

← Full comparison table (60+ models, live filters)