GLM 5.2 is a large open-weights text model from Z.ai designed to deliver frontier-tier reasoning and coding capabilities at a fraction of the cost of proprietary systems. Independent coverage describes it as a 753-billion-parameter Mixture-of-Experts architecture with roughly 40 billion active parameters, and the full weights were released under an MIT license shortly after a closed coding-plan preview. The model extends usable context to about one million tokens, making it well suited for long-document analysis, repository-scale code review, and multi-step agent workflows that combine reasoning with tool use and structured output.
On community benchmarks, GLM 5.2 is widely cited as one of the most capable open-weight models of its generation, with Artificial Analysis ranking it at the top of its open-weights leaderboard and a Hacker News security-bug-hunting test calling it a strong but not best-in-class performer. It is offered through multiple providers, including an FP8/NVFP4 quantized build served on Lilac's inference platform that fits a 524,288-token serving context. Practically, GLM 5.2 fits teams that need open-weight deployment for agentic and coding pipelines, sustained long-context reasoning, and cost-efficient access to frontier-level generation without locking into a closed vendor.