Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Qwen 3.8 Max

Qwen 3.8 Max represents Alibaba's most ambitious release in the Qwen family to date, scaling to 2.4 trillion parameters on the architectural foundation laid by Qwen 3.5. The model is explicitly positioned as a new bar for coding and agentic "Cowork" tasks, reflecting a strategic emphasis on work-oriented workloads where sustained reasoning and tool interaction matter more than open-ended chat. Marking the first time a Max-class Qwen will have its weights opened, the release signals a deliberate shift by the lab toward closed API access accompanied by a flagship open-weights follow-up, giving teams a path to evaluate through the API while anticipating a self-hosted option.

For practitioners, the practical story is a frontier-scale coder and workhorse rather than a lightweight generalist. The 2.4T parameter footprint, combined with lineage inherited from Qwen 3.5, points to a system tuned for long-horizon coding, multi-step agent behavior, and complex workflow execution where context retention across many turns is essential. Teams building developer tools, automated engineering pipelines, or agentic coworker experiences will find the model designed for exactly those scenarios, while those seeking casual or single-turn chat capability can rely on lighter Qwen variants. The combination of a closed-hosted API for immediate use and a promised open-weight release makes it a flexible option for both production deployment and deeper customization.

Vercel AI Gatewayalibaba/qwen3.8-maxqwen

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
alibaba/qwen3.8-max
Release date
Jul 19, 2026
Last updated
Jul 19, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.00
Output token cost
$6.00

Limits

Output tokens
128,000 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen 3.8 Max pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen 3.8 Max

Vercel AI Gateway

Official sourceRelease Notes

On September 1, 2026, Vercel shipped a new dated snapshot, Qwen 3.8 Max 0902, on the AI Gateway. The changelog entry states the gains are concentrated in coding on larger projects, long-horizon tasks that run without supervision, and agent runs, with vision handling described as more accurate on charts and dense docume Existing traffic on `alibaba/qwen3.8-max` can be moved onto the 0902 snapshot without a code change by adding a rewrite routing rule in the gateway. The entry lists supported surfaces including the AI SDK, Claude Code, Codex, Hermes, and Chat Completions, and points developers to the coding agents guide for agent setup

Vercel AI Gateway

Coverage

This Substack analysis covers Alibaba's Qwen 3.8 Max launch, confirmed at the World AI Conference in Shanghai on July 19, 2026 with official release on August 3, 2026. The piece details that the model is a sparse Mixture-of-Experts architecture with 2.4 trillion total parameters and approximately 95 billion active per The article argues the practical significance lies less in the 2.4 trillion parameter flagship—which is a multi-node datacentre artifact—and more in the companion Qwen3.8-27B open-weight release that fits standard on-premise GPU hardware. It characterizes the timing of Qwen 3.8 Max's release, days after Kimi K3, as par

Vercel AI Gateway

CoverageBenchmark

Yotta Labs provides a structured spec sheet for Qwen 3.8-Max, confirming the July 19, 2026 preview at WAIC Shanghai and the August 3, 2026 official launch with a 2.4 trillion parameter sparse MoE architecture activating roughly 95 billion parameters per token and a one-million-token context window with up to 128k outpu The article notes Alibaba's own framing that Qwen 3.8-Max is "second only to Fable 5," with no benchmark table yet published by Alibaba, and reports standard Alibaba API pricing at $2 per million input tokens, $6 output, and $0.25 cached, with OpenAI and Anthropic spec compatibility. It positions Qwen 3.8-Max as the ne

Vercel AI Gateway

CoverageBenchmark

TimesOfAI provides a technical breakdown of Qwen 3.8-Max launched on August 3, 2026, as a 2.4-trillion-parameter model with sparse Mixture-of-Experts architecture that activates 95 billion parameters during live inference. It confirms a one-million-token multimodal context window and emphasizes the model's design for l The article reports an Alibaba internal-validation claim that Qwen 3.8-Max operated autonomously within a live GitHub repository for 10 consecutive days, executing 265 code commits and 127 pull requests. It describes the model's intended workloads as coding, computer vision, and multi-step programming tasks, noting pre

Vercel AI Gateway

Official sourceRelease Notes

Vercel announced general availability of Alibaba's Qwen 3.8 Max on the Vercel AI Gateway on August 2, 2026. According to the changelog entry, Qwen 3.8 Max is a single model handling both text-only and vision-language workloads, built with 2.4 trillion parameters and a context window of up to 1 million tokens. Vercel po Integration is exposed through the unified AI Gateway API using the model ID `alibaba/qwen3.8-max`. Developers can also wire it into coding agents (Claude Code, Codex, OpenCode, Pi) via `vercel ai-gateway coding-agents setup`. The changelog confirms gateway features including custom reporting, Zero Data Retention suppo

Videos about Qwen 3.8 Max

More models around Qwen 3.8 Max