Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
302.AI logo

Model details

Qwen3.7 Max

Qwen3.7 Max is Alibaba's frontier Mixture-of-Experts language model announced in May 2026, arriving roughly six weeks after the Qwen 3.6 Plus release and continuing the same Qwen family lineage. It is positioned as a next-generation step in Alibaba's open-weight frontier lineup, sitting between the Qwen 3.6 Plus predecessor and the larger Qwen3.8-Max successor. The model is described as an agent-oriented system intended for coding and autonomous task workflows, aligning with Alibaba's broader push toward agent-first model design.

Architecturally, Qwen3.7 Max retains the MoE foundation from the prior generation while introducing updated expert routing refinements, which are reported to yield improvements across standard benchmark suites compared with Qwen 3.6 Plus. One of the most notable practical strengths is the retained the cataloged API limit token context window, enabling long-document analysis, extended multi-turn agent traces, and large repository-scale coding sessions without aggressive truncation. Compared with predecessor attention behavior, the refined MoE routing is expected to improve inference efficiency for agentic workloads, making the model a strong fit for developers building long-horizon coding assistants, research agents, and retrieval-augmented pipelines that benefit from extended context.

302.AIqwen3.7-maxqwen

Quick Info

Powered by
Provider
302.AI
Model key
qwen3.7-max
Release date
May 21, 2026
Last updated
May 21, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.80
Output token cost
$5.30

Limits

Output tokens
65,536 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen3.7 Max pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.7 Max

302.AI

Coverage

Alibaba released Qwen3.7 Max on 21 May 2026 as a proprietary, text-only flagship with a one-million-token context window and up to 65,536 output tokens. Beam AI emphasizes that the distinction between the proprietary Max model and the broader open Qwen ecosystem matters: a buyer selecting Max is purchasing a managed Al Qwen3.7 Max's strongest evidence is a 35-hour autonomous kernel-optimization run with more than 1,000 tool calls, testing the operating pattern of maintaining an objective across a long sequence of actions. The provider-reported benchmark suite is unusually broad: Terminal-Bench 2.0 69.7, SWE-bench Pro 60.6, SWE-bench

Videos about Qwen3.7 Max

More models around Qwen3.7 Max