Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
routing.run logo

Model details

GLM-5.2

GLM-5.2 is presented as a flagship release focused on long-horizon tasks, meaning extended multi-step coding and agent workflows rather than single-turn answers. The defining capability is a solid one-million-token context designed to stay coherent across very long, messy agent trajectories, framing the model as a foundation for engineering-heavy assistants that must keep state and reasoning intact across many turns and tool calls.

The model introduces architectural changes aimed at making that long context practical rather than just nominally large. It uses an IndexShare sparse attention design that reuses a single indexer across every four sparse attention layers, reportedly cutting per-token compute by about 2.9× at 1M tokens, and upgrades the MTP speculative-decoding layer to lift acceptance length by up to 20%. Coding capability is described as notably stronger than the prior GLM-5.1, with configurable thinking-effort levels to trade latency for depth. The release is distributed as a fully open weights model under an MIT license with no regional access limits, suiting teams that want on-premise or self-hosted long-context coding agents without proprietary restrictions.

routing.runglm-5.2glm

Quick Info

Powered by
Provider
routing.run
Model key
glm-5.2
Release date
Jun 13, 2026
Last updated
Jun 13, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.80
Output token cost
$2.40

Limits

Output tokens
32,000 tokens
Context window
200,000 tokens

Transparent token rates

Compare GLM-5.2 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-5.2

routing.run

Coverage

Simon Willison's independent write-up confirms GLM-5.2 is a 753B-parameter, 1.51TB Mixture-of-Experts model with 40 active parameters, released by Chinese AI lab Z.ai to coding-plan subscribers on June 13, 2026, with full open weights under an MIT license on June 16, 2026. It expands the context window from GLM-5.1's 2 On benchmarks, Artificial Analysis ranks GLM-5.2 at 51 on the Intelligence Index v4.1—leading open weights ahead of MiniMax-M3 (44), DeepSeek V4 Pro max (44), and Kimi K2.6 (43)—and it sits 2nd on the Code Arena WebDev leaderboard behind Claude Fable 5. Willison flags token-hunger as a practical cost concern: GLM-5.2 u

Videos about GLM-5.2

More models around GLM-5.2