Currently listed through these providers:
Model details
GLM (latest)
GLM (latest) on Privatemode AI is the alias for GLM-5.3, an open-weights large language model intended for general-purpose text generation, reasoning, and tool-based workflows. According to Privatemode's v1.55.0 changelog, this entry replaced the earlier GLM-5.2 preview, which is now deprecated and automatically routes to GLM-5.3 pending future removal. The model can be addressed either as glm-5.3 for an explicit version pin or as glm-latest for an always-current pointer, giving integrators a straightforward migration path from the previous 5.2 line.
For practical use, GLM-5.3 behaves as a reasoning-capable model with one notable operational constraint: the reasoning_effort=none option is not supported, and requests using that value are silently remapped to reasoning_effort=max, so reasoning behavior cannot be switched off through that parameter. Z.ai documentation summarized by a third-party reviewer also describes GLM-5.3 as supporting long context windows, streaming output, tool calls, and structured output across OpenAI-compatible endpoints, which makes it a reasonable fit for agent-style applications, multi-step problem solving, and structured data extraction tasks where continuous reasoning is acceptable.
Quick Info
Powered by- Provider
- Privatemode AI
- Model key
- glm-latest
- Release date
- Aug 14, 2026
- Last updated
- Aug 14, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.791
- Output token cost
- $8.9436
Limits
- Output tokens
- 131,072 tokens
- Context window
- 256,000 tokens
Transparent token rates
Compare GLM (latest) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM (latest)
No articles yet. Fetch the latest news to show it here.