Sulat.com
AI models
Vercel AI Gateway logo

Model details

GLM 4.7 Flash

GLM 4.7 Flash is positioned within Z.ai's GLM-4.7 family as a lightweight, high-speed variant aimed at developers who need practical coding and reasoning performance without heavy infrastructure. According to coverage of the release, it is built on a 31-billion-parameter architecture and is distributed under an MIT license with free API access, making it straightforward to integrate into agentic pipelines and local toolchains. It targets complex agent tasks, multi-step reasoning, and front-end coding scenarios where quick iteration matters more than maximum model depth.

Benchmark coverage highlights GLM 4.7 Flash as a capable coding assistant for its size, with reported scores of 59% on a Software Engineering benchmark, 79.5% on TA2 agentic tasks, and 75.2% on GPQA. Multiple thinking modes are available so teams can balance latency against reasoning depth, and streaming output plus function calling make it well suited to interactive developer environments. Its combination of an open-style license, long context window, and very low pricing makes it a strong fit for prototyping coding agents, building cost-sensitive production tools, and running local inference where a lightweight yet reasoning-capable model is required.

Vercel AI Gatewayzai/glm-4.7-flashglm-flash

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
zai/glm-4.7-flash
Release date
Jan 19, 2026
Last updated
Jan 19, 2026
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.07
Output token cost
$0.40

Limits

Output tokens
131,000 tokens
Context window
200,000 tokens

Transparent token rates

Compare glm-flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM 4.7 Flash

Vercel AI Gateway

CoverageBenchmark

Z.AI's GLM-4.7 Flash is a 31-billion-parameter open-source model released for coding, reasoning, and agentic workflows, offering free API access and local deployment. It records 59% on Software Engineering Bench, 79.5% on TA2 agentic tasks, and 75.2% on GPQA, while pricing lists $0.07 input, $0.01 cached input, and $0.

Vercel AI Gateway

CoverageBenchmark

GLM-4.7 Flash packs 31B parameters and an MIT license with free API access, helping you test ideas and ship tools on a tiny budget.

Vercel AI Gateway

Coverage

The news blog specialized in Japanese culture, odd news, gadgets and all other funny stuffs. Updated everyday.

Vercel AI Gateway

CoverageRelease Notes

GLM-4.7-Flash is a new member of the GLM 4.7 family and targets developers who want strong coding and reasoning performance in a model that is practical to...

Videos about GLM 4.7 Flash

More models around GLM 4.7 Flash