Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
EmpirioLabs AI logo

Model details

GLM 5.1

GLM-5.1 is positioned as a next-generation flagship model for agentic engineering, with a design that explicitly targets long-horizon work rather than single-turn answers. According to Z.ai's release notes and research blog, the model is built to operate independently for up to eight hours in a single run, sustaining a complete loop from planning and execution through iterative refinement to final delivery. This emphasis on endurance reflects a broader shift away from models that exhaust familiar techniques early and plateau, toward systems that can keep reasoning, experimenting, and revising their strategy across extended sessions. Training combines multi-turn supervised fine-tuning, reinforcement learning, and a process-quality evaluation framework, all aimed at improving stability, consistency, and tool use over drawn-out engineering tasks.

On evaluation, GLM-5.1 is reported to achieve state-of-the-art performance on SWE-Bench Pro, and to lead its predecessor GLM-5 by a wide margin on NL2Repo repository generation and Terminal-Bench 2.0 real-world terminal tasks, with quoted figures such as 58.4 on SWE-Bench Pro versus 55.1 for GLM-5. The blog frames the model as reaching comprehensive capability alignment with Claude Opus 4.6 across engineering intelligence, autonomous planning, sustained execution, bug fixing, and strategy iteration. Practically, GLM-5.1 fits teams that need an open-weights coding assistant capable of multi-hour debugging, refactoring, and repository-level synthesis, where the value comes from sustained judgment and revisitation rather than a single strong answer.

EmpirioLabs AIglm-5-1glm

Quick Info

Powered by
Provider
EmpirioLabs AI
Model key
glm-5-1
Release date
Apr 7, 2026
Last updated
Jun 12, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.825
Output token cost
$3.301

Limits

Output tokens
128,000 tokens
Context window
202,000 tokens

Transparent token rates

Compare GLM 5.1 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM 5.1

EmpirioLabs AI

CoverageBenchmark

Morph LLM's technical reference page characterizes GLM-5.1 as Z.ai's 744B MoE coding model released April 7, 2026 under the MIT license with a 200K context window and up to 128K-131K max output. It launched with a state-of-the-art 58.4 SWE-bench Pro score, ahead of GPT-5.4 (57.7) and Claude Opus 4.6 (57.3), and posted The page details GLM-5.1's real architecture, including DSA sparse attention and asynchronous agent RL infrastructure, and notes it was trained for 8-hour autonomous agentic runs across hundreds of tool-call rounds. It serves via transformers, vLLM, SGLang, KTransformers, and xLLM, and drops into Claude Code or Cline.

EmpirioLabs AI

CoverageBenchmark

InferenceX's analysis covers the GLM-5 / GLM-5.1 lineage from Zhipu AI (operating internationally as Z.ai), noting the GLM-5 weight repository was created on Hugging Face on 2026-02-11 and the GLM-5.1 repository on 2026-04-03, with public release dated April 7, 2026. Compared with GLM-4.5, GLM-5 scales from 355B parame GLM-5.1 is described as the follow-up point release on the same architecture, reaching state-of-the-art performance on SWE-Bench Pro and leading GLM-5 on NL2Repo and Terminal-Bench 2.0, with a distinguishing claim of sustaining optimization over hundreds of rounds and thousands of tool calls. Both models are served thr

EmpirioLabs AI

CoverageRelease Notes

Z.ai's official developer documentation release notes explicitly list GLM-5.1, dated 2026-04-07, among the GLM family updates. The first-party entry states GLM-5.1 is designed for long-horizon tasks and can work independently for up to 8 hours in a single run, covering planning, execution, iterative refinement, and fin As the primary first-party source, this release-notes page anchors the model's capabilities and timing directly to Z.ai (the model creator). The entry is embedded in a multi-release index that also documents subsequent versions such as GLM-5.2 and GLM-5.3, but the GLM-5.1 block itself is explicitly version-named and da

Videos about GLM 5.1

More models around GLM 5.1