Currently listed through these providers:
Model details
Qwen3.8 Max
Alibaba's Qwen3.8 Max is the general-availability flagship of the Qwen3.8 series, succeeding the earlier Preview release and positioned as a multimodal reasoning model built for complex reasoning, visual understanding, coding, and agentic workflows. It runs on a Mixture-of-Experts-style architecture with 2.4T total parameters and 95B active parameters, paired with a 1M-token context window that lets it handle very long documents and multi-step agent tasks in a single session. Alibaba released it through QwenCloud with an API-first rollout, and open weights were promised for shortly after launch, giving teams both a hosted path and a near-term self-hosted option.
In practice the model is aimed at workloads where depth of reasoning and long context matter more than minimal latency, including codebase-scale analysis, multi-document research, and tool-using agents that chain many steps together. The very wide context window makes it well suited to ingesting entire repositories, transcripts, or knowledge bases without aggressive truncation, while the active-parameter footprint keeps inference economical relative to its total capacity. Teams evaluating it should benchmark their low, medium, and xhigh reasoning settings against the published rates once Alibaba documents them, since prior Qwen3.7-Max pricing does not automatically carry over and cost per accepted task is the practical comparison point against other frontier APIs.
Quick Info
Powered by- Provider
- AIHubMix
- Model key
- qwen3.8-max
- Release date
- Aug 3, 2026
- Last updated
- Aug 3, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.69
- Output token cost
- $5.07
Limits
- Output tokens
- 128,000 tokens
- Context window
- 991,000 tokens
Latest news about Qwen3.8 Max
Videos about Qwen3.8 Max
Recent tweets and retweets from AIHubMix
More models around Qwen3.8 Max
This exact model name is also listed by 26 other providers.