Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

GLM 5.2

GLM-5.2 is positioned as Z.ai's flagship foundation model, explicitly designed for long-horizon tasks rather than short single-turn exchanges. According to Z.ai's developer documentation, the model is described as built "for the era of long-horizon tasks" and as having been "tested to handle project-scale engineering context," with claims of more stable long-task execution and higher success rates in development scenarios. This framing suggests a deliberate emphasis on sustained, multi-step workflows such as requirements gathering through to deployable products, where coherence across a large working memory matters more than one-shot answer quality.

Practically, the model pairs a very wide context window with output capacity suited to producing sizable artifacts in a single run, and the documentation highlights several quality-of-life capabilities: a configurable Thinking Mode for different reasoning depths, real-time streaming output, function calling, context caching, structured output for system integration, and flexible MCP tool/data integration. The combination of long context with these orchestration features points to a fit for agentic coding assistants, multi-file refactors, and research workflows where the model must track many constraints simultaneously. Independent scrutiny is also underway, as a NIST CAISI assessment page dated July 17, 2026, documents an evaluation completed on July 8, 2026, indicating the model is receiving formal third-party review alongside vendor-led claims.

Vercel AI Gatewayzai/glm-5.2glm

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
zai/glm-5.2
Release date
Jun 13, 2026
Last updated
Jun 13, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.80
Output token cost
$2.55

Limits

Output tokens
128,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare GLM 5.2 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM 5.2

Vercel AI Gateway

Official sourceAnnouncement

Z.ai (Zhipu AI) announced GLM-5.2 as its latest flagship model engineered for long-horizon tasks, published on its official blog dated June 16, 2026. The model delivers a solid 1-million-token context window that the vendor says holds up under sustained coding-agent workloads, and introduces a new architecture componen On three long-horizon coding benchmarks reported by Z.ai, GLM-5.2 trails Claude Opus 4.8 by only 1% on FrontierSWE while edging out GPT-5.5 by 1% and Opus 4.7 by 11%; on PostTrainBench, where each agent is given an H100 GPU for model post-training, GLM-5.2 ranks second only to Opus 4.8, outperforming both Opus 4.7 and

Vercel AI Gateway

CoverageAnalysis

A Baidu Cloud technical analysis dated September 7, 2026 examines GLM-5.2's architecture and performance, confirming the model maintains the GLM-5 series' 744 billion total parameters with 40 billion activated parameters per token. The article attributes the model's long-context efficiency to a sparse attention mechani The Baidu analysis reports benchmark results including a 32% improvement in long-sequence reasoning scores versus GLM-5.1 in KingBench 3 testing and 91% code completion accuracy in cross-file reference scenarios, positioning GLM-5.2 as outperforming multimodal models limited to 256K context windows in complex engineeri

Vercel AI Gateway

CoverageRelease Notes

Featherless published a Day Zero launch breakdown on 2026-06-18 confirming that GLM-5.2 ships as a Mixture-of-Experts model roughly 753B parameters in size with about 39B parameters activated per token, released by Z.ai under an MIT license and exposed through an OpenAI-compatible API. The piece attributes the version- The same post lists concrete benchmark deltas attributed to GLM-5.2 versus predecessor GLM-5.1: Terminal-Bench 2.1 rose from 63.5 to 81.0, SWE-bench Pro from 58.4 to 62.1, FrontierSWE from 30.5 to 74.4, and SWE-Marathon from 1.0 to 13.0, with reasoning gains on AIME 2026 (95.3 to 99.2) and GPQA-Diamond (86.2 to 91.2).

Vercel AI Gateway

CoverageBenchmark

A third-party review on aiforanything.io covers Zhipu AI's GLM-5.2, reporting a release date of June 13, 2026, and listing confirmed specs: a Mixture-of-Experts architecture with 744 billion total parameters and 40 billion active parameters per token, a 1,000,000-token input context window, a maximum output of 131,072 The review positions GLM-5.2 as the strongest open model for coding and agentic work, citing pricing roughly 10× cheaper than comparable frontier access (with a GLM Coding Plan starting at roughly $10/month), but flags that Zhipu shipped it without published benchmarks in the reviewer's view, meaning performance claims

Vercel AI Gateway

Official sourceRelease Notes

Z.ai's official developer release-notes index explicitly lists GLM-5.2 as a 2026-06-16 entry, distinct from later GLM-5.3 (2026-08-18) and GLM-5.3-Flash (2026-08-26) entries, helping fix the exact version under discussion. The GLM-5.2 entry states the model supports 1M lossless context, significantly improving long-hor By enumerating GLM-5.1, GLM-5.2, GLM-5.3, and GLM-5.3-Flash as separate line items with their own dates, the release-notes page functions as a version discriminator that rules out inadvertently splicing sibling-family evidence onto GLM-5.2. The page is first-party Z.ai material, so creator attribution to Z.ai (Zhipu) i

Videos about GLM 5.2

More models around GLM 5.2