Amazon Bedrock
Z.ai launched GLM-5 on 2026-02-12, scaling from GLM-4.5's 355B parameters (32B active, 23T tokens) to 744B parameters (40B active, 28.5T pre-training tokens) and integrating DeepSeek Sparse Attention to cut long-context deployment costs. The new slime asynchronous RL infrastructure enables more fine-grained post-training iterations at scale. GLM-5 targets complex systems engineering and long-horizon agentic tasks, with Z.ai reporting best-in-class open-source performance on reasoning, coding, and agentic benchmarks and narrowing the gap to Claude Opus 4.5 on CC-Bench-V2 and Vending Bench 2. Weights were released under MIT on Hugging Face and ModelScope, with API access via api.z.ai and BigModel.cn and Claude Code/OpenClaw compatibility.