Currently listed through these providers:
Model details
GLM-5.2
GLM-5.2 is positioned as Z.ai's flagship model for long-horizon work, marking a substantial capability leap over its predecessor GLM-5.1. Its central design priority is making extended context genuinely useful rather than merely available: the model is engineered to sustain quality across long, messy coding-agent trajectories where earlier models often degraded. This focus shapes both its headline capability, a solid one-the cataloged API limit, and its improved internal mechanics aimed at keeping that context reliable under real workload pressure.
Architecturally, GLM-5.2 introduces IndexShare, an approach that reuses a single indexer across every four sparse attention layers and reduces per-token compute at one-the cataloged API limit context lengths, paired with an enhanced multi-token prediction layer that extends speculative-decoding acceptance. The model also offers configurable thinking effort levels, letting developers tune the trade-off between coding performance and latency for different stages of an agent pipeline. Weights are publicly released under an MIT license with no regional restrictions, hosted on HuggingFace under the zai-org organization alongside code on GitHub, making GLM-5.2 a strong fit for teams building long-running coding agents, repository-scale refactoring workflows, and other tasks that benefit from sustained, high-quality reasoning across very large contexts.
Quick Info
Powered by- Provider
- Model Oracle AI
- Model key
- glm-5.2
- Release date
- Jun 13, 2026
- Last updated
- Jun 13, 2026
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 131,072 tokens
- Context window
- 1,000,000 tokens