Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

GLM-4.5V

GLM 4.5V is a vision-language model built on a sophisticated Mixture-of-Experts architecture, designed to bridge the gap between visual understanding and logical text generation. With 106 billion total parameters and 12 billion active parameters per task, the model is engineered to balance computational efficiency with high-level performance. It is specifically optimized for multimodal workflows, including image and video reasoning, document parsing, and GUI agent operations, making it a versatile tool for developers building interactive simulations, web applications, or autonomous research agents.

Rooted in the technical lineage of the GLM-4.5-Air foundation model and the GLM-4.1V-Thinking approach, this model leverages specialized training to achieve state-of-the-art results across dozens of public benchmarks. It offers a flexible operational design, allowing users to toggle between a thinking mode for deep, step-by-step reasoning and a non-thinking mode for rapid, straightforward responses. By supporting extensive multimodal context, the model provides a robust foundation for complex, multi-step workflows that require both visual grounding and precise, long-form text output.

Kilo Gatewayz-ai/glm-4.5vglm

Quick Info

Powered by
Provider
Kilo Gateway
Model key
z-ai/glm-4.5v
Release date
Aug 11, 2025
Last updated
Aug 11, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.60
Output token cost
$1.80

Limits

Output tokens
16,384 tokens
Context window
65,536 tokens

Transparent token rates

Compare GLM-4.5V pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-4.5V

No articles yet. Fetch the latest news to show it here.

Videos about GLM-4.5V

More models around GLM-4.5V