Currently listed through these providers:
Model details
GLM-5.1
GLM-5.1 extends the GLM line of large language models from Z.ai, and its presence in the NVIDIA NGC catalog under the nim/zai-org namespace confirms it as a Z.ai-published release intended for enterprise and developer distribution through standard model hubs. A March 2026 NVIDIA developer-forum request explicitly asked for GLM-5.1 to replace GLM-5 on NIM, signaling active community interest in deploying this newer revision alongside its predecessor on optimized serving infrastructure. Together, the NGC listing and forum thread establish GLM-5.1 as a recognized successor model within the broader GLM family rather than an obscure or unreleased variant. For practitioners, the practical fit is that of a text-only, open-weights general-purpose LLM suited to workloads that previously used GLM-5 and now want the refreshed 5.1 revision. The combination of open-weight availability and inclusion in NVIDIA's model catalog positions it for self-hosted inference, experimentation, and integration into agentic or tool-assisted pipelines where the underlying model weights can be inspected, fine-tuned, or deployed on-premises. As a cataloged NGC entry, GLM-5.1 is positioned as a drop-in upgrade candidate for teams already standardizing their GLM-based stacks on NIM-class serving platforms.
From a capability and lineage perspective, GLM-5.1 inherits the design intent of the GLM series as a versatile reasoning and generation model, and its appearance on NGC alongside other flagship open models reflects Z.ai's pattern of advancing the family with each numbered release. The forum demand to swap GLM-5 for GLM-5.1 on NIM suggests measurable improvements that users expect to carry over into existing agentic and reasoning pipelines without redesigning integrations. As an open-weights release distributed through NVIDIA's ecosystem, it offers a practical path for teams seeking a modern, Z.ai-developed LLM that can be served on NIM-compatible infrastructure while retaining the flexibility of self-hosted deployment.
Quick Info
Powered by- Provider
- EBCloud
- Model key
- GLM-5.1
- Release date
- Apr 7, 2026
- Last updated
- Apr 7, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.8571
- Output token cost
- $3.4286
Limits
- Output tokens
- 131,072 tokens
- Context window
- 200,000 tokens
Transparent token rates
Compare GLM-5.1 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM-5.1
No articles yet. Fetch the latest news to show it here.