Currently listed through these providers:
Model details
GLM-5-Turbo
GLM-5-Turbo represents a deliberate pivot from general-purpose language modeling toward agent-native execution. It is a proprietary derivative of the open-source GLM-5 family, but unlike its base model, it was optimized at the training level specifically for tool calling, instruction following, and sustained multi-step task chains. The model is built to power what Z.ai calls "claws" — autonomous proxy agents that handle persistent, real-world automation — and it targets the emerging OpenClaw ecosystem where models must operate reliably in proxy scenarios rather than just chat interfaces. Its design philosophy prioritizes execution stability and throughput over conversational polish, favoring hard, repeatable performance in agentic programming benchmarks over flashy "thinking" demonstrations.
The lineage from GLM-5 to GLM-5-Turbo traces a path of specialization rather than simple scaling. GLM-5 itself achieved open-source state-of-the-art performance on agentic benchmarks like SWE-bench Verified and Terminal Bench 2.0, reaching parity with Claude Opus 4.5 on core programming tasks. GLM-5-Turbo carries forward that benchmark strength while adding training-level integration of agent primitives — meaning the model was shaped for autonomous behavior from its foundation rather than patched in afterward. It integrates natively with developer tooling like Claude Code, Cline, Cursor, and MCP-compatible clients, making it practical for coding workflows where autonomous agents need to plan, execute, and iterate across long sessions. The result is a model that trades some general-purpose versatility for deeper fit in automated development and proxy-driven use cases.
Quick Info
Powered by- Provider
- Z.AI
- Model key
- glm-5-turbo
- Release date
- Mar 16, 2026
- Last updated
- Mar 16, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.20
- Output token cost
- $4.00
Limits
- Output tokens
- 131,072 tokens
- Context window
- 200,000 tokens
Transparent token rates
Compare glm pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM-5-Turbo
No articles yet. Fetch the latest news to show it here.