Published 2026-08-16 · 赵予安

Direct answer

Zhipu GLM 5.1 CN is listed at 0.120 / 0.479 per million tokens. Context window and CacheSafety for this exact id stay unpublished. This guide is written for product and platform teams comparing model quality, cost, routing policy, and production rollout risk.

What shipped

GLM 5.1 CN is now live in the NextModel gateway catalog. The model id is z-ai/glm-5.1-cn, and the public slug is z-ai--glm-5.1-cn. The provider is Z Ai. The model has streaming support. The catalog assigns it to the quality use case. Its latency band is 1000 to 3000 ms.

Pricing

Input price is $0.120 per 1M tokens. Output price is $0.479 per 1M tokens. The context window is unpublished in the catalog snapshot. Related catalog entries include GLM 5 CN at $0.084 per 1M tokens in and $0.373 per 1M tokens out. GLM 5.2 CN is listed at $0.171 per 1M tokens in and $0.596 per 1M tokens out.

Who this is for

This is for teams that need a quality-focused model at a price point between GLM 5 CN and GLM 5.2 CN. The latency band fits interactive workloads. Streaming output is available for real-time use.

What is still unverified

The release date is unpublished in this catalog snapshot. Independent third-party quality scores are not in these notes. The context window is unpublished when the catalog value is zero. No first-party CacheSafety Bench row exists for this exact model. Safe Hit, Bad Hit, and trap rates are not reported.