Sulat.com
AI models
Z.AI logo

Model details

GLM-4.7-FlashX

GLM-4.7-FlashX serves as an extended iteration within the glm-flash family, specifically engineered to provide enhanced performance while maintaining the lightweight deployment characteristics of its predecessor. Built upon a Mixture-of-Experts architecture, this model is designed to handle complex text-based tasks with greater efficiency. Its design intent focuses on delivering robust reasoning and functional capabilities, making it a versatile choice for developers who require a balance between computational speed and the ability to manage extensive information streams.

The model benefits from a lineage that prioritizes operational efficiency, allowing it to maintain high utility in demanding environments. By building on the foundation of the existing Flash MoE framework, it achieves a practical strength in processing large volumes of text, ensuring that it remains responsive even when tasked with intricate reasoning or multi-step instructions. This focus on architectural refinement positions the model as a forward-looking solution for applications that demand consistent, high-quality text generation without the overhead associated with larger, less specialized systems.

Z.AIglm-4.7-flashxglm-flash

Quick Info

Powered by
Provider
Z.AI
Model key
glm-4.7-flashx
Release date
Jan 19, 2026
Last updated
Jan 19, 2026
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.07
Output token cost
$0.40

Limits

Output tokens
131,072 tokens
Context window
200,000 tokens

Transparent token rates

Compare glm-flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-4.7-FlashX

No articles yet. Fetch the latest news to show it here.

Videos about GLM-4.7-FlashX

More models around GLM-4.7-FlashX