Currently listed through these providers:
Model details
GLM-4.5-Flash
GLM-4.5-Flash belongs to Zhipu AI's glm-flash family, a lineage of compact text models that the provider documents at docs.z.ai. The family is positioned as Zhipu's lighter-weight flash tier, and related sibling entries in the same family appear on the Sulat catalog with a the listed knowledge cutoff knowledge horizon. That shared lineage points to a model designed for fast, general-purpose text generation rather than heavyweight reasoning workloads, with Zhipu handling the upstream training and serving while third-party routes like UnoRouter expose the model under their own keys.
Through UnoRouter, GLM-4.5-Flash is surfaced under a free-tier key, making it well suited to prototyping, lightweight agent loops, and high-volume experimentation where cost matters more than frontier accuracy. The model accepts and produces text only, and supports the core developer affordances of the glm-flash family such as reasoning-style prompting, tool/function calling, and temperature control, which together let builders wire it into retrieval or agent pipelines without bespoke glue. In practice it fits projects that need a Zhipu-backed text model on a budget router, with the understanding that newer GLM generations (such as later glm-flash releases) may offer stronger reasoning for more demanding tasks.
Quick Info
Powered by- Provider
- UnoRouter
- Model key
- glm-4.5-flash:free
- Release date
- Jul 28, 2025
- Last updated
- Jul 28, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 98,304 tokens
- Context window
- 131,072 tokens