Sulat.com
AI models
Kilo Gateway logo

Model details

Z.ai: GLM Flash Latest

Z.ai: GLM Flash Latest is positioned as a multimodal API model that accepts image input alongside text and emits text completions, making it suitable for workflows that mix visual references with natural-language instructions. Third-party catalog coverage lists vision input and function calling as supported, which signals a fit for assistant-style applications that need to interpret screenshots, diagrams, or other visual context while orchestrating external tools. The entry sits within Z.ai's GLM family line, giving it the naming pattern shared with other multimodal reasoning endpoints in that family.

In practical terms, the model targets budget-sensitive production traffic: catalog data points to a low per-million-token input rate and a competitive output rate, with a million-scale context window that comfortably fits long documents, extended conversations, or retrieval-augmented pipelines. Function-calling support allows it to slot into agent frameworks where the model decides between tools or hands structured payloads to downstream services. Vision support broadens that footprint to document understanding, UI reasoning, and image-grounded Q&A, while the multimodal input profile makes it a flexible general-purpose option for teams that want one endpoint to cover text-plus-image tasks without sacrificing the price-to-context ratio.

Kilo Gateway~z-ai/glm-flash-latestglm-flash

Quick Info

Powered by
Provider
Kilo Gateway
Model key
~z-ai/glm-flash-latest
Release date
Aug 27, 2026
Last updated
Aug 27, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.07125
Output token cost
$0.2375

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Z.ai: GLM Flash Latest pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Z.ai: GLM Flash Latest

Kilo Gateway

CoverageBenchmark

Z.ai's GLM Flash Latest was released on August 27, 2026, according to the LM Market Cap model page, which serves as a third-party aggregator tracking this Z.ai model routed via the Kilo Gateway. The listing notes that this model "always redirects to the latest model in the GLM Flash family," meaning it functions as a d API pricing for GLM Flash Latest is set at $0.07 per million input tokens and $0.25 per million output tokens in USD, with the provider advising contact for volume and enterprise discounts. The model offers a 1.3M-token context window with a maximum output capacity of approximately 943.7K tokens, and supports text, ima

Videos about Z.ai: GLM Flash Latest

More models around Z.ai: GLM Flash Latest