ZenMux
Z.ai announced GLM-5.2 on June 16, 2026 as its flagship model built for long-horizon tasks, introducing a solid 1-million-token context window designed to remain reliable across extended coding-agent trajectories. The release emphasizes that a long context must be engineering-usable, not merely wide, and the model ship Architecturally, GLM-5.2 introduces IndexShare, which reuses the same indexer across every four sparse attention layers, reducing per-token FLOPs by 2.9× at the 1M context length, along with an upgraded Multi-Token Prediction (MTP) layer that increases speculative-decoding acceptance length by up to 20%. The model is r