Sulat.com
AI models
302.AI logo

Model details

glm-4.7-flashx

glm-4.7-flashx belongs to Z.AI's GLM-4.7 family, which the publisher describes as an upgrade over earlier GLM releases in two concrete areas: stronger programming ability and more stable multi-step reasoning and execution. Within that family, the FlashX variant is explicitly positioned as a lightweight, high-speed, and affordable option, sitting alongside a free Flash sibling. The series as a whole targets complex agent tasks while keeping conversational tone natural and front-end code output polished, and the model itself is text-in/text-out, with optional thinking modes, streaming output, function calling, and structured outputs to support tool use and real-time interfaces.

Because weights are released, glm-4.7-flashx can be self-hosted or fine-tuned, which matters for the high-volume, cost-sensitive workloads it is aimed at, such as background automation, batch processing, and pipeline steps where a larger flagship would be wasteful. Third-party trackers highlight the very low per-token price relative to other GLM tiers, and the FlashX tier in particular is described as one of the cheapest inference options in its segment, making it a sensible default when scale and latency matter more than top-end reasoning. In practice this is a model to reach for when you want agent-friendly behavior, tool calling, and structured outputs on a budget, rather than a frontier-reasoning system.

302.AIglm-4.7-flashxglm-flash

Quick Info

Powered by
Provider
302.AI
Model key
glm-4.7-flashx
Release date
Jan 20, 2026
Last updated
Jan 20, 2026
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.0715
Output token cost
$0.429

Limits

Output tokens
131,072 tokens
Context window
200,000 tokens

Transparent token rates

Compare glm-4.7-flashx pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about glm-4.7-flashx

302.AI

Coverage

The UsagePricing blueprint for Zhipu AI (Z.ai) documents the GLM family's published US dollar pricing as of late July 2026, placing GLM-4.7 in the same tier as GLM-4.6 and GLM-4.5 at roughly $0.6 per 1M input tokens and $2.2 per 1M output tokens, which is the most concrete per-token economic reference available for the Z.ai's GLM-4.7-Flash tier is listed as free to call, with cached input available for as low as $0.01 per 1M tokens and cached-input storage marked 'Limited-time Free', making it one of the cheapest Flash-tier inference options in the market and directly relevant to developers evaluating the glm-4.7-flashx SKU on 302.AI

Videos about glm-4.7-flashx

More models around glm-4.7-flashx