Impossibl
Z.ai announced GLM-4.6 on 2025-09-30 as the latest version of its flagship model, expanding the context window from 128K to 200K tokens to handle more complex agentic tasks. The model delivers higher scores on code benchmarks and improved real-world performance in Claude Code, Cline, Roo Code, and Kilo Code, including more visually polished front-end generation. Z.ai reports clear gains over GLM-4.5 across eight public benchmarks covering agents, reasoning, and coding. On the extended CC-Bench evaluation, human evaluators ran GLM-4.6 inside isolated Docker containers on multi-turn tasks spanning front-end development, tool building, data analysis, testing, and algorithms, with the model reaching near parity with Claude Sonnet 4 at a 48.6% win rate. Z.ai also notes GLM-4.6 finishes tasks with about 15% fewer tokens than GLM-4.5, and the full evaluation details plus trajectory data are publicly available on Hugging Face for community research and reproduction.