Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Opper logo

Model details

Qwen3.8 Max

Alibaba's Qwen team introduced Qwen3.8-Max as the most capable model in its Qwen family, building on the architectural foundation of Qwen 3.5 and scaling to 2.4 trillion total parameters with 95 billion active per pass — a Mixture-of-Experts design that points to substantial capacity for complex reasoning while keeping inference manageable. The launch marked a notable shift in openness for the lineup, with Alibaba announcing that weights for a Qwen-Max-class model would be released openly for the first time shortly after the debut. The model is positioned as an end-to-end workhorse: rather than answering isolated questions, it is designed to take a real project from an empty folder to a finished deliverable on its own, with emphasis on dependable execution of multi-step coding, research, and general workplace workflows.

The intended practical fit for Qwen3.8-Max is sustained, agent-like productivity on long-horizon tasks — exactly the kind of multi-day coding and research work that requires holding context across many turns and tools. Its broad capability profile spans coding, real-life work, and research scenarios, and the flagship positioning is reinforced by strong independent leaderboard showings in both text and vision evaluations shortly after release. For teams evaluating it, the model reads as a general-purpose foundation that can drive both chat-style interactions and longer autonomous workflows, with an unusually large context window that suits projects where files, codebases, and background research need to stay in view throughout a session.

Opperqwen3.8-maxqwen

Quick Info

Powered by
Provider
Opper
Model key
qwen3.8-max
Release date
Aug 3, 2026
Last updated
Aug 3, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.00
Output token cost
$6.00

Limits

Output tokens
131,072 tokens
Context window
983,616 tokens

Transparent token rates

Compare Qwen3.8 Max pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.8 Max

Opper

CoverageBenchmark

Alibaba's Qwen3.8-Max debuted via the QwenCloud API on August 2, 2026, and was published as an open-weight MoE checkpoint (Qwen3.8-2.4T-A95B, 2.4T total / 95B active parameters) on Hugging Face and ModelScope on August 13, with a dense Qwen3.8-27B sibling following on August 14. The release framing on Alibaba's Qwen.ai Coverage notes that Qwen3.8-Max landed in a crowded September release window, with OpenAI, Anthropic, and Google all shipping competing frontier models within days of each other. Reported benchmark highlights include 86.6 on Terminal-Bench 2.1 and a 1M-token context window, framed as evidence that open-weight releases

Opper

CoverageBenchmark

This guide covers Qwen3.8-Max as a 2.4T-parameter / 95B-active MoE with 512 experts (10 routed plus 1 shared), 92 layers, and a hybrid Gated-DeltaNet plus full-attention architecture built on the Qwen3.5 stack, previewed July 19, 2026 and released August 3, 2026. Context is listed as 262,144 tokens native, extensible t On licensing, the article describes the Max terms as "close to verbatim MIT" but with two thresholds bolted on: attribution is required above 100M monthly active users or $20M monthly revenue, and a separate licence is required only for Model-as-a-Service or "AI Work Assistant" businesses above $50M in twelve months, a

Opper

CoverageBenchmark

This explainer describes Qwen3.8-Max as a 2.44-trillion-parameter sparse Mixture-of-Experts model that activates approximately 95 billion parameters per token, launched August 3, 2026 after a July 19 preview, with a 1 million-token context window and up to 128k output tokens. The piece reports Qwen3.8-Max scoring 86.6 The guide frames the realistic deployment choice as "API-vs-API" rather than self-hosted-versus-API, arguing that for most teams the open-weight release changes licensing and auditability rather than hosting economics. It highlights the Qwen3.8-27B as the variant most teams will actually be able to run locally, disting

Opper

Coverage

Alibaba's official August 3, 2026 announcement describes Qwen3.8-Max as the largest and most capable model in its Qwen series to date, with 2.4 trillion total parameters, a context window of up to 1 million tokens, and multimodal visual intelligence support. The model ranked fifth on Text Arena and second on Vision Are The post details the architecture as a Sparse Mixture-of-Experts design with a hybrid attention mechanism, activating just 95 billion parameters despite the 2.4T total, which Alibaba frames as the key to delivering frontier-level intelligence at reduced compute and latency versus dense models of similar scale. Capabili

Opper

CoverageBenchmark

Written on the August 3, 2026 launch day, this guide treats Qwen3.8 Max as Qwen's flagship for coding, professional work, multimodal agents, research, and long-horizon execution, describing it as a 2.4-trillion-parameter MoE with 95B active parameters per token. It emphasizes a "feedback-loop" product thesis: the model The launch evidence cited includes a 16-day autonomous coding project, a five-day research reproduction, a 24-hour multimodal competition, a closed-loop chip-design run, hundreds of professional-work tests, and broad benchmark gains. The piece clearly distinguishes Qwen3.8 Max from the separately announced Qwen3.8 27B

Videos about Qwen3.8 Max

More models around Qwen3.8 Max