LLM Insights中文
Zhipu logo

Zhipu API Pricing & Model Comparison

We track 2 Zhipu models, with input pricing from $0.40 to $0.49 per million tokens. The line-up's blended rate runs 72% below the catalog average. Below: price, context, speed and value compared across the range.

Models tracked2
Lowest input price$0.40 / 1M
Top quality score93
Largest context1000K

For cost, Zhipu's cheapest option is GLM-4.7 ($0.40 input / $1.75 output); the quality ceiling is GLM-5.2 at a quality score of 93. On a 50,000-turn-per-month support workload that is $84 versus $88 — the spread inside a single provider is often wider than the spread between providers.

On value (quality score ÷ output price) the pick of the Zhipu range is GLM-5.2. The largest context window is 1000K, offered only by GLM-5.2, and the fastest output is GLM-4.7 at 130 tok/s. Modalities across the line-up: Text, Vision.

Models and pricing

Sorted by blended rate, cheapest first (3:1 input-to-output mix). Follow a model name for full specs and monthly cost estimates.

ModelInput /1MOutput /1MBlended /1MQualityContextSpeedValueUse cases
GLM-4.7$0.40$1.75$0.7487205K130 tok/s49.7
ChatbotCodingAffordable
GLM-5.2$0.49$1.54$0.75931000K100 tok/s60.4
CodingReasoningLong Context

Frequently asked questions

Which Zhipu model is cheapest?

GLM-4.7, at $0.40 input and $1.75 output per million tokens, with a quality score of 87.

What is Zhipu's flagship model?

By quality score it is GLM-5.2 (93), priced at $0.49 input / $1.54 output with a 1000K context window.

Does Zhipu pricing change?

Yes. llmprice.app runs a daily collection job and refreshes these figures, but there can be a lag after a provider re-prices — confirm against Zhipu's official pricing page before committing.

Browse all providers

Prices and specs are collected daily by llmprice.app. Confirm against the official pricing page before production use.