發布於 2026-08-16 · 赵予安
直接回答
智谱 GLM 5.1 CN 標價為每百萬 token 0.120 / 0.479。上下文窗口和該精確 ID 的 CacheSafety 仍未公布。 這篇指南面向正在比較模型品質、成本、路由策略和正式環境上線風險的台灣產品與平台團隊。
What shipped
GLM 5.1 CN is now live in the NextModel gateway catalog. The model id is z-ai/glm-5.1-cn, and the public slug is z-ai--glm-5.1-cn. The provider is Z Ai. The model has streaming support. The catalog assigns it to the quality use case. Its latency band is 1000 to 3000 ms.
Pricing
Input price is $0.120 per 1M tokens. Output price is $0.479 per 1M tokens. The context window is unpublished in the catalog snapshot. Related catalog entries include GLM 5 CN at $0.084 per 1M tokens in and $0.373 per 1M tokens out. GLM 5.2 CN is listed at $0.171 per 1M tokens in and $0.596 per 1M tokens out.
Who this is for
This is for teams that need a quality-focused model at a price point between GLM 5 CN and GLM 5.2 CN. The latency band fits interactive workloads. Streaming output is available for real-time use.
What is still unverified
The release date is unpublished in this catalog snapshot. Independent third-party quality scores are not in these notes. The context window is unpublished when the catalog value is zero. No first-party CacheSafety Bench row exists for this exact model. Safe Hit, Bad Hit, and trap rates are not reported.