SiliconFlow
GLM-5V-Turbo, Z. AI's multimodal coding foundation model built for vision-based coding and agent-driven tasks, brings native multimodal understanding and balanced visual-programming capabilities.
Model details
GLM-5V-Turbo is positioned as Z.AI's first multimodal coding foundation model, designed to bridge visual understanding and programmatic reasoning in a single system. The model is framed as natively multimodal, taking in images alongside text so it can read screenshots, diagrams, or interface mockups and turn that perception into code or structured tool calls. SiliconFlow describes it as bringing balanced visual-programming capabilities aimed at agent-driven workflows, while the broader Z.AI family places it within a generation of models targeting vision-grounded software tasks rather than purely descriptive outputs.
In practical terms, GLM-5V-Turbo is best suited for developers who want an assistant that can look at a UI or schematic and produce working code, plan multi-step refactors with tool use, or sustain long agentic sessions thanks to its very large working memory. Reporting highlighted in a Medium write-up places it at number five on BridgeBench SpeedBench at roughly 221 tokens per second, suggesting competitive throughput against frontier multimodal systems, though that claim rests on a single secondary source. Hosted availability on SiliconFlow, alongside broader routing through OpenRouter-style gateways, gives teams flexible access at token-based pricing while keeping the same underlying multimodal behavior.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
SiliconFlow
GLM-5V-Turbo, Z. AI's multimodal coding foundation model built for vision-based coding and agent-driven tasks, brings native multimodal understanding and balanced visual-programming capabilities.
SiliconFlow
GLM-5V-Turbo is Z.AI’s first multimodal coding foundation model, built for vision-based coding tasks. It also hit #5 on BridgeBench SpeedBench with 221.2 tokens/sec, faster than Gemini 3.1 Pro…