Currently listed through these providers:
Model details
GLM Flash (latest)
GLM Flash (latest) appears in a third-party model directory under the name "Z.ai: GLM Flash Latest," where it is described as a multimodal API model. That directory explicitly flags vision input and function calling as supported, pointing toward a design intended for applications that combine image and text understanding with tool-driven workflows such as assistants, document analysis, and structured automation pipelines.
Beyond those directory-level capability flags, the supplied evidence does not include any official Z.ai model card, technical report, or benchmark results for GLM Flash (latest). Pricing and context length figures in the directory listing also conflict with the catalog metadata, and no creator-authored documentation is available to confirm training approach, parameter count, or evaluation performance. As a result, the only safe characterization of the model is that it is a Z.ai-attributed multimodal offering with vision and function-calling features, leaving deeper claims about architecture, training scale, and benchmarks unsupported at present.
Quick Info
Powered by- Provider
- Privatemode AI
- Model key
- glm-flash-latest
- Release date
- Aug 26, 2026
- Last updated
- Aug 26, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.8897
- Output token cost
- $4.4718
Limits
- Output tokens
- 131,072 tokens
- Context window
- 256,000 tokens
Transparent token rates
Compare GLM Flash (latest) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM Flash (latest)
No articles yet. Fetch the latest news to show it here.