Currently listed through these providers:
Model details
GLM-5.2
GLM-5.2 is Z.ai's flagship successor to GLM-5.1, engineered specifically to sustain quality across lengthy, messy coding-agent trajectories rather than merely accepting more tokens. Its headline capability is a solid one-the cataloged API limit that the team describes as the first time long-horizon behavior is delivered reliably at that scale. The model is positioned for agentic and long-horizon coding workflows where maintaining coherence over an entire project or session matters more than peak single-prompt performance, and Z.ai ships it under a permissive MIT license with no regional access restrictions.
Beyond raw context length, GLM-5.2 introduces two notable technical advances. First, it adds configurable thinking-effort levels for coding, letting developers trade latency for capability depending on the task's difficulty. Second, the model adopts a new architectural pattern called IndexShare, which the team reports reuses a single indexer across every four sparse-attention layers and roughly triples per-token FLOP efficiency at the cataloged API limit lengths, paired with an improved multi-token prediction layer that lengthens speculative-decoding acceptance by up to twenty percent. Together these make the model a strong practical fit for sustained agentic coding, large repository analysis, and other workflows that need both depth and endurance.
Quick Info
Powered by- Provider
- UnoRouter
- Model key
- glm-5.2
- Release date
- Jun 13, 2026
- Last updated
- Jun 13, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.6001
- Output token cost
- $5.0288
Limits
- Output tokens
- 131,072 tokens
- Context window
- 1,000,000 tokens