Currently listed through these providers:
Model details
GLM-4.5V
GLM 4.5V is a vision-language model built on a sophisticated Mixture-of-Experts architecture, designed to bridge the gap between visual understanding and logical text generation. With 106 billion total parameters and 12 billion active parameters per task, the model is engineered to balance computational efficiency with high-level performance. It is specifically optimized for multimodal workflows, including image and video reasoning, document parsing, and GUI agent operations, making it a versatile tool for developers building interactive simulations, web applications, or autonomous research agents.
Rooted in the technical lineage of the GLM-4.5-Air foundation model and the GLM-4.1V-Thinking approach, this model leverages specialized training to achieve state-of-the-art results across dozens of public benchmarks. It offers a flexible operational design, allowing users to toggle between a thinking mode for deep, step-by-step reasoning and a non-thinking mode for rapid, straightforward responses. By supporting extensive multimodal context, the model provides a robust foundation for complex, multi-step workflows that require both visual grounding and precise, long-form text output.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- z-ai/glm-4.5v
- Release date
- Aug 11, 2025
- Last updated
- Aug 11, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $1.80
Limits
- Output tokens
- 16,384 tokens
- Context window
- 65,536 tokens
Transparent token rates
Compare GLM-4.5V pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM-4.5V
No articles yet. Fetch the latest news to show it here.