Currently listed through these providers:
Model details
GLM-5.2
GLM-5.2 was introduced as Z.ai's flagship model aimed squarely at long-horizon tasks, marking a substantial leap over its predecessor GLM-5.1 and, for the first time in the lineup, pairing that capability with a solid the cataloged API limit context. The release emphasizes that long context only matters when the model can sustain quality across messy, extended coding-agent trajectories, not merely accept more tokens. The model is released under an MIT license as open weights with no regional access limits, with artifacts hosted on Hugging Face under the zai-org organization and the broader GLM-5 repository on GitHub.
Architecturally, GLM-5.2 introduces an efficiency-oriented component called IndexShare, which reuses the same indexer across every four sparse attention layers and is credited with reducing per-token compute at the cataloged API limit context, alongside an improved multi-token prediction layer that boosts speculative decoding acceptance length. These changes underpin both the long-context gains and a stronger coding capability exposed through configurable thinking effort levels that let users trade latency against performance. GLM-5.2 also served as the base model for the later GLM-5.3 release, whose entire improvement came from scaled post-training on long-horizon environments, indicating that GLM-5.2 is best understood as a foundation designed for sustained agentic and coding workloads.
Quick Info
Powered by- Provider
- Alibaba (China)
- Model key
- glm-5.2
- Release date
- Jun 13, 2026
- Last updated
- Jun 13, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.10
- Output token cost
- $3.851
Limits
- Output tokens
- 128,000 tokens
- Context window
- 1,000,000 tokens