Sulat.com
AI models
UnoRouter logo

Model details

GLM-5.2

GLM-5.2 is Z.ai's flagship successor to GLM-5.1, engineered specifically to sustain quality across lengthy, messy coding-agent trajectories rather than merely accepting more tokens. Its headline capability is a solid one-the cataloged API limit that the team describes as the first time long-horizon behavior is delivered reliably at that scale. The model is positioned for agentic and long-horizon coding workflows where maintaining coherence over an entire project or session matters more than peak single-prompt performance, and Z.ai ships it under a permissive MIT license with no regional access restrictions.

Beyond raw context length, GLM-5.2 introduces two notable technical advances. First, it adds configurable thinking-effort levels for coding, letting developers trade latency for capability depending on the task's difficulty. Second, the model adopts a new architectural pattern called IndexShare, which the team reports reuses a single indexer across every four sparse-attention layers and roughly triples per-token FLOP efficiency at the cataloged API limit lengths, paired with an improved multi-token prediction layer that lengthens speculative-decoding acceptance by up to twenty percent. Together these make the model a strong practical fit for sustained agentic coding, large repository analysis, and other workflows that need both depth and endurance.

UnoRouterglm-5.2glm

Quick Info

Powered by
Provider
UnoRouter
Model key
glm-5.2
Release date
Jun 13, 2026
Last updated
Jun 13, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.6001
Output token cost
$5.0288

Limits

Output tokens
131,072 tokens
Context window
1,000,000 tokens

Latest news about GLM-5.2

Videos about GLM-5.2

More models around GLM-5.2