Sulat.com
AI models
Wafer logo

Model details

GLM-5.1

GLM-5.1 serves as Z.AI's flagship foundation model, engineered for tasks that demand sustained, autonomous execution over extended periods. According to official documentation, the model can operate continuously on a single task for up to eight hours, moving through planning, execution, and iterative optimization to deliver production-grade results. This long-horizon capability, combined with its open-weight availability, positions GLM-5.1 as a practical foundation for building autonomous agents and coding assistants that need to maintain coherence across complex, multi-step engineering workflows rather than single-turn interactions.

In terms of qualitative performance positioning, Z.AI documents GLM-5.1 as broadly aligned with leading proprietary models in both general capability and coding, while demonstrating stronger sustained execution on complex engineering optimization and real-world development tasks. The model supports multiple thinking modes for different reasoning scenarios, alongside streaming output, function calling, context caching, and structured output formats like JSON, making it well-suited for integration with external toolsets and production pipelines. Third-party analysis places GLM-5.1 in the competitive tier of Chinese frontier open-weight models, where it serves as a stepping stone in the lineage leading to its successor, offering developers a proven, open foundation for agentic software engineering before migrating to newer releases.

WaferGLM-5.1glm

Quick Info

Powered by
Provider
Wafer
Model key
GLM-5.1
Release date
Apr 7, 2026
Last updated
Jun 1, 2026
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$3.20

Limits

Output tokens
131,072 tokens
Context window
202,752 tokens

Transparent token rates

Compare GLM-5.1 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-5.1

Wafer

Coverage

GLM-5.1 by Wafer, released on Apr 7, 2026. Compare capabilities, pricing, limits, providers, and latest news on Sulat.

Wafer

Coverage

Mehul Gupta's Medium walkthrough positions GLM-5.2 (released in June 2026, just months after GLM-5.1) as Zhipu's doubled-down bet on agentic software engineering, with a 1 million token context window, new reasoning modes tuned for coding, and explicit emphasis on tool usage, multi-step reasoning, repository analysis, For developers coming from the Wafer-hosted GLM-5.1, the piece traces the same release lineage — GLM-5 → GLM-5.1 → GLM-5.2 — and explains how the 1M-token context window and coding-focused reasoning modes inherited from and expanded upon in GLM-5.1 underpin GLM-5.2's open-weight coding performance, offering a migration

Wafer

CoverageAnalysis

For developers evaluating Wafer's GLM-5.1 against other hosted open-weight options, Artificial Analysis provides useful quantitative baseline figures for the same model architecture: GLM-5.1 sits at an Intelligence Index of 40, uses roughly 26k output tokens per Intelligence Index task, and costs about $0.25 per task o The same Artificial Analysis article confirms that GLM-5.2 (the immediate successor) is priced at $1.4 per 1M input, $4.4 per 1M output, and $0.26 per 1M cache-hit tokens on Z.ai's first-party API, and is MIT-licensed. While this is GLM-5.2 data rather than GLM-5.1-specific, it gives Wafer customers a reference point f

Wafer

Coverage

Z.ai's GLM series sits within the Tier 4 Chinese frontier models grouping in the May 9, 2026 LLM Encyclopedia, alongside ERNIE, Kimi, Qwen, Hunyuan and others, reflecting GLM-5.1's positioning as an open-weight coding-focused model family competing with proprietary leaders on benchmarks and developer tools. Although the explainx.ai blog is primarily framed around the GLM-5.2 release and the so-called Fable 5 export-ban episode, it provides concrete ecosystem-level signal relevant to GLM-5.1's current standing: GLM-5.2 remains text-only while topping open coding benchmarks, Kimi K2.7 Code is now GA in GitHub Copilot, and C

Wafer

Coverage

The Stochastic Sandbox LLM Encyclopedia (May 9, 2026) catalog of 60+ models from 22 use cases places Zhipu's GLM-5 and GLM-5.1 (alongside ChatGLM) under Tier 4 — Chinese Frontier Models, alongside Baidu ERNIE, Moonshot Kimi, Baichuan, 01.AI Yi, Hunyuan, InternLM and ByteDance Seed, as pricing and benchmark tables are r For developers evaluating the Wafer-hosted GLM-5.1 endpoint, this encyclopedia serves as an independent third-party reference confirming GLM-5.1's place in the open-weight Chinese frontier tier; however, the supplied excerpt does not contain GLM-5.1-specific benchmark deltas, pricing, or Wafer routing details, so it fu

Wafer

CoverageBenchmark

Z.ai releases GLM-5.1, an open-source model scoring 94.6% of Claude Opus 4.6 in coding benchmarks. Full breakdown of benchmarks, pricing, architecture, and what's still unverified.

Wafer

Coverage

The Chinese company said its new open-source model can continue to improve over hundreds of iterations, as AI vendors race to build tools that can handle longer software tasks.

Videos about GLM-5.1

More models around GLM-5.1