Model details
Qwen3.8 Max
Qwen3.8-Max is the current flagship of the Qwen family, positioned by its developers as a step up for coding, research, and long-running agent tasks rather than a routine refresh. Built on the architectural foundation of Qwen 3.5, it scales to 2.4 trillion total parameters with 95 billion active in a Mixture-of-Experts design, and the release post frames the model around one idea: taking real multi-day projects from an empty folder to a finished result on its own. The the cataloged API limit context window, hybrid thinking enabled by default, and structured output support make it well suited to workloads that mix large document analysis, tool use, and iterative code generation in a single session.
For teams evaluating it in production, Qwen3.8-Max is most interesting where reliability on long-horizon tasks matters more than single-turn latency, such as autonomous coding agents, multi-step research pipelines, and document-heavy reasoning jobs that benefit from the the cataloged API limit. The official announcement highlights end-to-end task completion and "dependable deliverables" as the practical differentiator, and the model is distributed through QwenCloud with third-party access points exposing the qwen3.8-max identifier. Qwen also indicated that the weights for this Max-class model would be released openly shortly after launch, an unusual move for the lineup that lowers the barrier for self-hosting and fine-tuning once those checkpoints become available.
Quick Info
Powered by- Provider
- OpenCode Go
- Model key
- qwen3.8-max
- Release date
- Aug 3, 2026
- Last updated
- Aug 3, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $2.00
- Output token cost
- $6.00
Limits
- Output tokens
- 131,072 tokens
- Context window
- 1,000,000 tokens