Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
DevPass (LLM Gateway) logo

Model details

GLM-5.1

GLM-5.1 is a 754-billion parameter flagship model built specifically for agentic engineering—tasks that demand sustained reasoning and autonomous execution over hours rather than minutes. Its architecture is engineered to stay effective across long horizons, unlike predecessors that plateau after initial gains. The model handles ambiguous problems with adaptive judgment, breaks complex issues into experiments, reads results, and revises strategy through repeated iteration. This design makes it particularly strong at automating multi-hour engineering workflows, from writing and rewriting CUDA kernels to executing complex software development cycles end-to-end.

While explicit pre-training details remain limited in public sources, GLM-5.1 represents a significant leap over GLM-5, achieving state-of-the-art performance on SWE-Bench Pro with a score of 58.4—outpacing GPT-5.4, Opus 4.6, and Gemini 3.1 Pro on complex software engineering benchmarks. The model has been released as open-weight through Zhipu's GitHub and HuggingFace repositories, enabling community access and further development. Practical applications include automated repo generation, real-world terminal task execution, and cybersecurity problem-solving, with the model sustaining productive performance across extended sessions in ways previous GLM generations could not.

DevPass (LLM Gateway)glm-5.1glm

Quick Info

Powered by
Provider
DevPass (LLM Gateway)
Model key
glm-5.1
Release date
Apr 7, 2026
Last updated
Apr 7, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.931
Output token cost
$2.93

Limits

Output tokens
131,072 tokens
Context window
204,800 tokens

Transparent token rates

Compare GLM-5.1 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-5.1

DevPass (LLM Gateway)

CoverageRelease Notes

The official Z.ai developer documentation release notes confirm GLM-5.1 was released on April 7, 2026, designed specifically for long-horizon tasks. According to the official release note, GLM-5.1 can work independently for up to 8 hours in a single run, enabling a full loop from planning and execution through iterativ Z.ai's official release notes state that GLM-5.1 achieves comprehensive capability alignment with Claude Opus 4.6. The model was built with multi-turn SFT (supervised fine-tuning), RL (reinforcement learning), and process-based training methods. The same documentation page also documents subsequent GLM family releases

Videos about GLM-5.1

More models around GLM-5.1