Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Wafer logo

Model details

GLM-5.2

GLM-5.2 sits within the glm family and is positioned as an open-weights text model aimed at developer and agent-coding workloads. Wafer surfaces it on its serverless inference platform alongside peers such as glm5.2-fast, GLM-5.1, and Kimi-K2.6, offering pay-as-you-go access through integrations with coding-agent clients including Linzumi, Claude Code, Conductor, Codex, OpenClaw, Hermes Agent, Cline, Roo Code, Kilo Code, OpenHands, and LibreChat. The platform markets itself on making open models substantially faster than generic inference providers, framing GLM-5.2 as a model suited to interactive, tool-driven development environments rather than as a general chat assistant.

Independent reviewers have begun examining GLM-5.2 in depth, with a June 2026 Medium analysis from Devansh exploring its behavior beyond launch headlines and a late-June 2026 episode of Lenny's Newsletter's How I AI series dedicated to a hands-on GLM 5.2 review. These pieces indicate active practitioner interest in how the model performs inside real agent loops and coding workflows, complementing Wafer's positioning of GLM-5.2 as a practical backbone for autonomous coding agents. For builders choosing among open alternatives on Wafer, GLM-5.2 represents the higher-capability option in the lineup while glm5.2-fast targets latency-sensitive paths.

WaferGLM-5.2glm

Quick Info

Powered by
Provider
Wafer
Model key
GLM-5.2
Release date
Jun 13, 2026
Last updated
Jun 22, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.20
Output token cost
$4.10

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare GLM-5.2 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-5.2

Wafer

CoverageBenchmark

OpenRouter lists Z.ai’s GLM 5.2 as a large-scale reasoning model released on June 16, 2026, supporting text input and output with a 1-million-token context window. The page says it is designed for long-horizon agent workflows, project-level software engineering, and complex multistep automation, with high and xhigh rea The page provides a point-in-time comparison of more than a dozen hosting providers and routing options for GLM 5.2. The displayed provider table includes input, output, and cache-read prices per million tokens alongside latency, throughput, uptime, and quantization; listed rates range from $0.343/$1.078 for input/outp

Videos about GLM-5.2

More models around GLM-5.2