Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Venice AI logo

Model details

GLM 5 Turbo

GLM-5 Turbo is built on the GLM-5 architecture with a 744 billion parameter foundation and 40 billion active parameters during inference, leveraging DeepSeek Sparse Attention to keep deployment costs manageable without sacrificing capability. The model targets agent-driven workflows and scenarios like OpenClaw, where sustained task execution and reliable tool use matter more than raw benchmark scores. Its design philosophy centers on handling long execution chains with stability, making it suited for applications that require persistent, multi-step reasoning rather than single-shot responses.

The model is deeply optimized for real-world agent scenarios, with improvements in complex instruction decomposition, scheduled and persistent execution, and overall stability across extended tasks. These optimizations aim to make long-chain task execution more practical at scale. GLM-5 Turbo positions itself as a foundation model for developers building autonomous agents, workflow automation, and systems that need to maintain context and reliability across prolonged interactions rather than just handling individual prompts effectively.

Venice AIz-ai-glm-5-turboglm

Quick Info

Powered by
Provider
Venice AI
Model key
z-ai-glm-5-turbo
Release date
Mar 15, 2026
Last updated
Jun 11, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.20
Output token cost
$4.00

Limits

Output tokens
32,768 tokens
Context window
200,000 tokens

Transparent token rates

Compare GLM 5 Turbo pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM 5 Turbo

Venice AI

CoverageBenchmark

BenchGecko's model page lists GLM 5 Turbo as a proprietary text model from Z.ai released March 2026 with a 203K-token context window, priced at $1.20 per million input and $4.00 per million output tokens, noting it was tested on three benchmarks with no fully populated score. It reports Artificial Analysis composite sc The page situates GLM 5 Turbo within the z-ai GLM 5 family alongside base GLM 5 (Feb 2026), highlighting a $0.60/M input price increase and a 59K larger context versus its sibling. It links to z-ai pricing, developer documentation, a research/technical report, an API playground, and community channels, while flagging t

Venice AI

CoverageRelease Notes

The Opper AI release tracker explicitly lists GLM-5-Turbo among Z.ai's dated model releases, confirming its launch date as March 15, 2026. The entry shows a 203K context window and an intelligence index score of 27, corroborating the VentureBeat coverage of the model as a Turbo-branded variant within Z.ai's GLM family. The tracker entry lists pricing for GLM-5-Turbo at $1.20 per million input tokens and $4.00 per million output tokens — figures that diverge from VentureBeat's OpenRouter-cited $0.96/$3.20 pricing, a discrepancy likely reflecting different aggregator snapshots or routing markups. Despite the pricing difference, both so

Videos about GLM 5 Turbo

More models around GLM 5 Turbo