Sulat.com
AI models
Get 10-25% off
Get 10-25% off from Qwen
Alibaba logo

Model details

Qwen3.8 Max

Qwen3.8 Max is positioned as the most capable entry in the Qwen family, built on the architectural foundation of the earlier Qwen 3.5 generation and scaled up to a 2.4-trillion-parameter Mixture-of-Experts design with roughly 95 billion active parameters per pass. That lineage is leveraged to push improvements across coding, research-oriented work, and long-horizon tasks, with the release framed around end-to-end delivery of complex jobs rather than single-turn answers. The model is also described as a native vision-language system, reflecting its text, image, video, and PDF input handling, and is presented as the first Qwen-Max-class checkpoint whose weights will be opened, an unusual move for a flagship tier.

In practical terms, Qwen3.8 Max is aimed at developers and teams who need a single model that can carry a multi-day coding project from an empty folder to a finished result, while also handling agent-style workflows and visual understanding on the side. Its hybrid thinking mode is enabled by default, and it exposes function calling, built-in tools, and structured outputs, making it suitable for tool-using pipelines that need both reasoning and reliable machine-readable responses. The combination of a very large context window, multimodal input, and tool orchestration makes it a fit for ambitious coding, research assistance, and complex task automation where dependability of the final deliverable matters as much as raw answer quality.

Alibabaqwen3.8-maxqwen

Quick Info

Powered by
Provider
Alibaba
Model key
qwen3.8-max
Release date
Aug 3, 2026
Last updated
Aug 3, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.00
Output token cost
$6.00

Limits

Output tokens
131,072 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen3.8 Max pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.8 Max

Alibaba

Official sourceAnnouncement

Alibaba officially unveiled Qwen3.8-Max on August 3, 2026 as the most powerful model in its Qwen series to date. The flagship features 2.4 trillion total parameters, activates just 95 billion parameters per forward pass, and supports a context window of up to 1 million tokens. The model ranks fifth in Text Arena and se Built on Qwen 3.5 with a Sparse Mixture-of-Experts architecture combined with a hybrid attention mechanism, Qwen3.8-Max balances massive scale with inference efficiency. It is accessible via APIs on Alibaba Cloud Model Studio for global developers, with model weights scheduled for release the following week, and can al

Cloudflare AI Gateway

Official sourceAnnouncement

Alibaba officially launched Qwen3.8-Max on August 3, 2026 in Hangzhou, describing it as the most powerful model in its Qwen series to date. The model uses a Sparse Mixture-of-Experts architecture with a hybrid attention mechanism, totaling 2.4 trillion parameters while activating only 95 billion at inference, and supports a 1 million token context window. Qwen3.8-Max ranks fifth on Text Arena, second on Vision Arena, and fourth on Frontend Code Arena, and is available via APIs on Alibaba Cloud Model Studio with weights scheduled for release the following week. The model demonstrates multimodal capabilities and strong long-horizon autonomous coding performance; Alibaba reports Qwen3.8-Max autonomously executed a real-world software engineering project over a 16-day period. Internal testing emphasized strengths in coding, real-life work, research, and long-horizon tasks, with the model also accessible through QwenWork, Alibaba's all-in-one workplace AI agent platform. Open-weight versions are planned for distribution through Alibaba Cloud's Model Studio.

Alibaba

CoverageBenchmark

Alibaba pushed open weights for Qwen3.8-Max onto Hugging Face and ModelScope in mid-August 2026, with the open-weight variant Qwen3.8-2.4T-A95B landing on August 13, 2026, and a smaller dense sibling Qwen3.8-27B shipping its own open weights the next day on August 14. The release came after the August 2, 2026 API debut The Qwen3.8-Max architecture is a sparse mixture-of-experts design with 2.4 trillion total parameters and only 95 billion active on any given forward pass, making self-hosting feasible without a small data center. Alibaba positioned the model under the headline "A New Bar for Coding and Cowork" on its Qwen.ai blog, tar

Alibaba

Coverage

Qwen3.8-Max-0902 (alias qwen3.8-max-2026-09-02) went live on QwenCloud on September 1, 2026, as an upgraded snapshot of the 2.4-trillion-parameter Qwen3.8-Max base model. The variant retains the same 1M context window, thinking mode, and tool ecosystem, but is further post-trained on what Qwen calls "Coding & Cowork" t Pricing for the 0902 snapshot remains unchanged from the base Qwen3.8-Max at $2 per million input tokens and $6 per million output tokens, with cache reads at $0.25 per million on implicit cache hits, $0.17 on explicit cache reads, and explicit cache creation at $2.50 per million. On Qwen's published benchmark table, C

Alibaba

Coverage

A third-party Medium explainer dated August 14, 2026 reiterates that Qwen3.8-Max was officially released on August 3, 2026 as Alibaba's largest and most capable model to date, and describes it as the first time Alibaba has open-sourced a Max-level flagship. The article restates the Sparse MoE architecture at 2.4 trilli The piece frames Qwen3.8-Max as a milestone for open-source LLMs reaching a global frontier level and pivots into developer-tooling discussion, arguing that integrating a comprehensive AI Gateway or MCP server lets teams switch between models like Qwen3.8-Max without modifying application code. It notes this is the fir

Cloudflare AI Gateway

Coverage

VentureBeat reported on August 3, 2026 that Alibaba's Qwen team unveiled Qwen3.8-Max as a flagship 2.4-trillion-parameter MoE multimodal LLM aimed at autonomous software engineering and long-horizon enterprise work. The article cites Alibaba's claim that Qwen3.8-Max scored 86.1 on the OSWorld-Verified benchmark, ahead The piece frames the launch as a strategic shift because Alibaba indicated open weights would be released the following week alongside Qwen3.8-27B, which would mark the first time a Max-class Qwen model became available for self-hosted deployment. VentureBeat notes that Alibaba had not yet disclosed the licensing terms

Alibaba

Coverage

Alibaba's Qwen team officially released Qwen3.8-Max on August 3, 2026, via QwenCloud as the new flagship of the Qwen family. According to the official Qwen blog, the model uses a 2.4 trillion-parameter Mixture-of-Experts architecture with 95 billion active parameters per inference, is built on the Qwen 3.5 foundation, The launch post emphasizes end-to-end task execution rather than benchmark tables, highlighting a 10+ day autonomous coding run that produced the "oh-my-cli" self-evolving harness on GitHub, plus research-method refinement and competition-leaderboard-climbing workflows. The blog also notes the model is callable via the

Cloudflare AI Gateway

Coverage

InfoWorld reported on August 3, 2026 that Alibaba introduced Qwen3.8-Max as its largest AI model yet, expanding its enterprise AI portfolio with an open-weight model targeted at software engineering, multimodal reasoning, and knowledge-intensive workloads. The article confirms the 2.4-trillion-parameter MoE architecture activating about 95 billion parameters at inference, with open-weight versions scheduled for release the following week via Alibaba Cloud's Model Studio. Alibaba positioned Qwen3.8-Max as comparable to leading frontier models, second only to Fable 5. The launch included internal benchmark comparisons against Claude Opus 4.8, Claude Fable 5, and OpenAI's GPT-5.6 Sol on coding tests such as SWE-Bench Pro and Alibaba's proprietary NL2Repo-Bench, using each vendor's own harness. Forrester VP Charlie Dai noted Alibaba is narrowing the gap with proprietary leaders, framing the larger story as rapid maturation of open-weight models offering enterprises credible alternatives for software engineering and cost-sensitive deployments.

Alibaba

CoveragePreview

EvoLink's documentation, originally posted July 21 and updated August 3, 2026, confirms Qwen3.8-Max is live on QwenCloud and exposed through EvoLink's gateway as the production route qwen3.8-max. The guide lists the QwenCloud model catalog features: 1 million token context window, hybrid Thinking enabled by default, Fu The article explicitly disambiguates Qwen3.8 Max (the 2026 flagship) from the older Qwen3-8B checkpoint, warning developers not to conflate local deployment guides or pricing for Qwen3-8B with the new Max release. It also flags that EvoLink's documentation URL still retains a legacy "Preview" slug, advising developers

Cloudflare AI Gateway

CoverageBenchmark

The AI Release Tracker records Qwen3.8-Max as released by Qwen on August 3, 2026, 63 days after Qwen3.7-Plus, listing it as a 2.4T parameter model with a 1M token context window. The page compiles third-party benchmark scores spanning BullshitBench v2, PaperBench, OSWorld-Verified, BabyVision, SWE-Bench Pro, DeepSWE 1.1, and additional tests, providing a comparative reference set for the model's coding and reasoning performance against other frontier models. Benchmark highlights include PaperBench at 93%, SWE-Bench Pro at 67.7% (5th of 23, behind Claude Fable 5.1's 81.2%), DeepSWE 1.1 at 56.6% (19th of 29), and NL2Repo-Bench at 55.9% (4th of 6, behind DeepSeek-V4.1-Flash's 65.4%). The tracker lists Alibaba as a provider with input pricing at $2.00 and output at $6.00 per million tokens, with rates fetched from openrouter.ai on September 25, 2026.

Videos about Qwen3.8 Max

More models around Qwen3.8 Max