Qwen3.8-Max is Alibaba's top-tier mixture-of-experts model, scaled to 2.4 trillion total parameters with 95 billion active per pass and built on the architectural foundation of Qwen 3.5. It was teased at the World AI Conference in Shanghai on July 19, 2026 with a claim of ranking "second only to Fable 5," then formally released a few weeks later and made available through QwenCloud. The official launch positioned it as a flagship for tackling extended, multi-step projects end-to-end rather than just answering isolated prompts, and the team announced that open weights of a Max-class model would follow the release — the first time Alibaba has open-sourced a model at this tier.
In practice, the model is aimed at developers and research teams who need long-context reasoning and reliable execution on real engineering work. An independent benchmark explainer reports strong showings on coding and agent-style suites such as PaperBench (93.0), Terminal Bench 2.1 (86.6), MathVision (95.2), IFBench (82.8), and OSWorld-Verified (86.1), while flagging softer results on HLE (43.6, last among four compared flagships) and SWE-bench Pro (67.7, about 12 points behind Fable 5). Combined with its the cataloged API limit working window, that profile makes Qwen3.8-Max a natural fit for long-horizon coding agents, codebase-scale analysis, and research workflows where sustained, end-to-end task delivery matters more than single-turn question answering.