Currently listed through these providers:
Model details
GLM-4.7-FlashX
GLM-4.7-FlashX serves as an extended iteration within the glm-flash family, specifically engineered to provide enhanced performance while maintaining the lightweight deployment characteristics of its predecessor. Built upon a Mixture-of-Experts architecture, this model is designed to handle complex text-based tasks with greater efficiency. Its design intent focuses on delivering robust reasoning and functional capabilities, making it a versatile choice for developers who require a balance between computational speed and the ability to manage extensive information streams.
The model benefits from a lineage that prioritizes operational efficiency, allowing it to maintain high utility in demanding environments. By building on the foundation of the existing Flash MoE framework, it achieves a practical strength in processing large volumes of text, ensuring that it remains responsive even when tasked with intricate reasoning or multi-step instructions. This focus on architectural refinement positions the model as a forward-looking solution for applications that demand consistent, high-quality text generation without the overhead associated with larger, less specialized systems.
Quick Info
Powered by- Provider
- Z.AI
- Model key
- glm-4.7-flashx
- Release date
- Jan 19, 2026
- Last updated
- Jan 19, 2026
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.07
- Output token cost
- $0.40
Limits
- Output tokens
- 131,072 tokens
- Context window
- 200,000 tokens
Transparent token rates
Compare glm-flash pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM-4.7-FlashX
No articles yet. Fetch the latest news to show it here.
Videos about GLM-4.7-FlashX
More models around GLM-4.7-FlashX
This exact model name is also listed by 7 other providers.