Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Qwen3.7 Flash

Qwen3.7 Flash sits at the cost- and speed-oriented end of Alibaba's Qwen3.7 series, positioned as a vision-language reasoning model that accepts interleaved text and image input and produces text output. Like its Qwen3.7, Qwen3.6, and Qwen3.5 siblings served through Alibaba Cloud Model Studio, it is built as a hybrid thinking model: it can either emit an explicit reasoning trace before answering or respond directly, with the reasoning behavior controlled by an enable_thinking switch that defaults to on for this generation. Weights are not published, so it is deployed as a proprietary endpoint rather than an open download.

In practice the model is tuned for multimodal agent workloads rather than open-ended chat, with reported strengths in object recognition, spatial understanding, and perception of real-world scenes, as well as visual coding, search, and computer-use style tasks where the model reads screen content and reasons over interface state. A Qwen3.7 Flash Thinking variant is offered for deeper multimodal reasoning, multi-step task execution, and longer agent trajectories. The combination of roughly a one-the cataloged API limit with a 65,536-token generation ceiling lets it hold long multi-image sequences, long documents, or extended tool traces in a single request, and on the NanoGPT Auto route it is listed with sub-two-second latency at around 73 tokens per second, making it a sensible fit for production pipelines that need multimodal reasoning at a low per-token cost.

OpenRouterqwen/qwen3.7-flashqwen

Quick Info

Powered by
Provider
OpenRouter
Model key
qwen/qwen3.7-flash
Release date
Jul 15, 2026
Last updated
Jul 15, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.03
Output token cost
$0.13

Limits

Input tokens
991,808 tokens
Output tokens
65,536 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen3.7 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.7 Flash

OrcaRouter

Coverage

Qwen3.7 Flash is an Alibaba vision-language model in the Qwen3.7 family, and a Spheron Network explainer confirms it shipped via QwenCloud with a one-paragraph first-party changelog entry dated July 25, 2026 — with OpenRouter listing it two days later on July 27. Alibaba's own description, quoted in the article, calls The same Spheron post documents an important distribution shift for the Qwen3.7 line: unlike Qwen 3, 3.5, and the open side of 3.6 — which shipped downloadable weights on day one — the entire Qwen3.7 series (Flash, Plus, and Max) is closed weight and served only through Alibaba's QwenCloud API plus aggregator resellers

CrossModel

Coverage

According to the Krater.ai catalog page, Qwen3.7 Flash is a vision-language reasoning model attributed to Alibaba Qwen, released on July 27, 2026. The page states that the model supports a 1,000,000-token context window with a maximum output length of 65,536 tokens, and accepts text, image, and video as input while pro The same Krater page characterizes Qwen3.7 Flash as a multimodal reasoning model from the Qwen family, designed to combine image understanding with long-context text handling and tool-using workflows. The catalog entry is a third-party aggregation of model metadata rather than an Alibaba/Qwen first-party announcement,

Videos about Qwen3.7 Flash

More models around Qwen3.7 Flash