Currently listed through these providers:
Model details
GLM 5.2 Short Flex
GLM 5.2 Short Flex extends the GLM 5.2 family with a profile that combines the shorter context tier of the GLM 5.2 Short line with the lower Flex pricing tier, positioning it as a budget-oriented variant for assistant-style workloads that still benefit from reasoning and tool use. On the Neuralwatt catalog, the GLM 5.2 family spans the full-context base, Fast, and Flex variants alongside Short and Short Fast siblings, and Short Flex sits at the intersection of those axes, keeping the more compact 199,984-token window while adopting the Flex rate card. Its siblings in the same directory are listed with reasoning, tool calling, and temperature control flags turned on, suggesting Short Flex is intended for the same agentic, instruction-following use cases as the rest of the GLM 5.2 line rather than as a specialized model.
Because the supplied provider excerpt does not contain a dedicated "GLM 5.2 Short Flex" row, the catalog's specific input and output figures, cache read price, and capability flags for this exact model key cannot be directly evidenced from the sources provided. What is supported is the broader pattern: Flex variants in the GLM 5.2 family are priced at the lower tier, Short variants share the 199,984-token context and output window, and the no-reasoning Fast variants are paired with reasoning-enabled siblings. The practical takeaway is that GLM 5.2 Short Flex is best understood as a cost-optimized, reasoning-capable endpoint for shorter-prompt workloads where the full million-token GLM 5.2 context is unnecessary, while teams needing the largest context should look at the base GLM 5.2 or its Fast and Flex variants.
Quick Info
Powered by- Provider
- Neuralwatt
- Model key
- glm-5.2-short-flex
- Release date
- Jun 17, 2026
- Last updated
- Jun 17, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.9425
- Output token cost
- $2.925
Limits
- Output tokens
- 32,000 tokens
- Context window
- 199,984 tokens
Latest news about GLM 5.2 Short Flex
No articles yet. Fetch the latest news to show it here.