Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Venice AI logo

Model details

Qwen 3.8 Max

Qwen 3.8 Max is positioned by Alibaba's Qwen team as the most capable model in the Qwen family to date, built on the architectural foundation of Qwen 3.5 and scaled to roughly 2.4 trillion parameters. It is presented as a Qwen-Max-class flagship focused on comprehensive improvements across coding and collaborative work scenarios, an area the team describes as a new bar for these tasks. Because the official release entry on the Qwen research page carries a date stamp of 2026/08/26, this listing represents the post-preview, officially released state of the model rather than the earlier qwen3.8-max-preview that circulated through Token Plan and Qoder channels.

In practical terms, the model is aimed at developers and teams who need long-context reasoning, vision understanding, and tool-driven workflows in a single system. Coursiv's preview-stage write-up noted that Qwen 3.8 already exposed a real endpoint with documented context and output limits, reasoning controls, and vision input through subscription access, which lines up with the multimodal, agentic direction emphasized in the official release. The team also signals a forward-looking shift by highlighting Qwen3.8-Flash-Next alongside it, a multimodal MoE sibling that previews the hybrid Gated DeltaNet plus Gated Attention architecture intended to carry forward into Qwen4, suggesting that Max-class users are also getting a glimpse of the next-generation design lineage.

Venice AIqwen-3-8-maxqwen

Quick Info

Powered by
Provider
Venice AI
Model key
qwen-3-8-max
Release date
Jul 22, 2026
Last updated
Jul 19, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.50
Output token cost
$7.50

Limits

Output tokens
131,072 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen 3.8 Max pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen 3.8 Max

Venice AI

Official sourceAnnouncement

Alibaba officially announced Qwen3.8-Max on August 3, 2026 as the most powerful model in its Qwen series, featuring 2.4 trillion total parameters and a context window of up to 1 million tokens. The multimodal model supports visual intelligence natively and ranks fifth on Text Arena and second on Vision Arena. Built on Qwen 3.5's foundation, Qwen3.8-Max uses a Sparse Mixture-of-Experts architecture with a hybrid attention mechanism, activating only 95 billion of its 2.4 trillion parameters per inference pass to balance scale with efficiency. Developers can access it via Alibaba Cloud Model Studio APIs, with model weights scheduled for release the following week.

Venice AI

Coverage

Alibaba first unveiled Qwen3.8-Max, a 2.4-trillion-parameter vision-language model for long-running coding and knowledge work, on August 2, 2026, then released its open weights within a week, a first for a Max-tier Qwen model. The open-weights version is text-only and limited versus the full API model. The mixture-of-experts transformer activates 95 billion parameters per token, supports 1-million-token inputs and 131,000-token outputs at 77.6 tokens per second, and exposes reasoning levels, function calling, structured output, and prefix completion. It ranks fifth on Artificial Analysis' Intelligence Index and second on Arena.ai's Vision Arena, with API pricing of $2/$0.25/$6 per million input/cached/output tokens.

Venice AI

CoverageBenchmark

Alibaba made Qwen3.8-Max accessible to developers worldwide from August 3, 2026 through Model Studio APIs and the QwenWork platform in public beta on web and desktop. The model was previewed at the World AI Conference in Shanghai on July 19, 2026, reachable only through Alibaba's Token Plan and Qoder platforms. The vendor-stated specifications include 2.4 trillion total parameters using a Mixture-of-Experts architecture, context up to 1 million tokens, and maximum output of 131,072 tokens. Open-weight release was announced for the following week on Hugging Face and ModelScope, with licensing details still undefined at launch.

Venice AI

CoverageBenchmark

On September 2, 2026, Alibaba released Qwen3.8-Max-0902, a post-training upgrade focused on coding and long-horizon agent work, with architecture and pricing unchanged from the original. It tops Code Arena WebDev at 1,691 points, beating Claude Opus 5 Max by three and the prior Qwen3.8-Max by 22. All eight coding benchmarks improved, with TerminalBench 3.0 jumping from 11.3 to 29.0 and ProgramBench Almost Solved rising from 10.5 to 28.0. The upgrade also leads Claude Opus 5 on MLS-Bench-Lite, SWE-Atlas QnA, QwenSWEBench V2, WorkArena, and both published multimodal evaluations.

Videos about Qwen 3.8 Max

More models around Qwen 3.8 Max