Currently listed through these providers:
Model details
GLM-5
GLM-5 is positioned for complex systems engineering and long-running agentic tasks, extending the earlier family through substantially larger sparse architecture and broader pre-training. The official launch materials describe 744 billion total parameters with 40 billion active, alongside increased pre-training data and DeepSeek Sparse Attention intended to reduce deployment cost while retaining long-context capability. This makes it a practical fit for coding, systems work, and tool-driven workflows that require sustained reasoning rather than short, isolated responses.
The model’s later lineage shows continued gains from post-training: the August 2026 GLM-5.3 release used the same base model as GLM-5.2 and reported improvements in complex coding and long-horizon tasks, including a 50% improvement on Z.ai’s in-house Code Bench and state-of-the-art results on CyberGym. Those findings are useful indicators of the family’s direction, but they are vendor-reported claims rather than independent benchmark validation.
Quick Info
Powered by- Provider
- Deep Infra
- Model key
- zai-org/GLM-5
- Release date
- Feb 12, 2026
- Last updated
- Feb 12, 2026
- Knowledge cutoff
- 2025-12
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $2.08
Limits
- Output tokens
- 16,384 tokens
- Context window
- 202,752 tokens
Transparent token rates
Compare GLM-5 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM-5
No articles yet. Fetch the latest news to show it here.