Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenCode Zen logo

Model details

GLM-5

GLM-5 marks a significant architectural leap from its predecessor, scaling up to 744 billion total parameters with 40 billion active parameters per token, up from the 355 billion total and 32 billion active seen in GLM-4.5. This growth is paired with expanded pre-training data—28.5 trillion tokens compared to 23 trillion—giving the model richer world knowledge to draw from. To keep deployment practical despite the larger scale, GLM-5 integrates DeepSeek Sparse Attention, which preserves long-context reasoning capabilities while reducing operational overhead. The model is purpose-built for complex systems engineering and multi-stage agentic workflows, with design emphasis on the kind of deep, extended problem-solving that powers production-grade coding agents and long-horizon task execution.

The team developed an internal infrastructure called "slime"—an asynchronous reinforcement learning framework—to address the challenge of scaling RL training for large language models. This approach substantially improves training throughput and enables more granular post-training iterations, bridging the gap between a model's base competence and excellence in real-world tasks. The result is a model that achieves best-in-class performance among all open-source models on reasoning, coding, and agentic benchmarks, narrowing the gap with frontier closed models. Released under the MIT license, GLM-5 is designed to be deployed in production environments, with partner platforms offering optimized inference pipelines, one-click deployment, autoscaling, and tools like LoRA adapters and quantization to balance performance and cost for real workloads.

OpenCode Zenglm-5glm

Quick Info

Powered by
Provider
OpenCode Zen
Model key
glm-5
Release date
Feb 11, 2026
Last updated
Feb 11, 2026
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$3.20

Limits

Output tokens
131,072 tokens
Context window
204,800 tokens

Transparent token rates

Compare GLM-5 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-5

OpenCode Zen

CoverageBenchmark

The SemiAnalysis InferenceX technical page treats GLM-5 as Zhipu AI's (operating internationally as Z.ai, Hugging Face org zai-org) flagship open-weights large language model, with the Hugging Face weight repository created on 2026-02-11 and public release dated February 11, 2026. Architecture-wise GLM-5 scales from GL The same page also documents GLM-5.1, the follow-up point release on the same architecture, positioned as Z.ai's "next-generation flagship model for agentic engineering" with significantly stronger coding capabilities than its predecessor and SOTA performance on SWE-Bench Pro plus wide-margin leads over GLM-5 on NL2Rep

Videos about GLM-5

More models around GLM-5