Currently listed through:
Model details
Qwen 3.8 Max Prime
Qwen 3.8 Max Prime is a Qwen-family large language model from Alibaba positioned as a high-speed variant of Qwen 3.8 Max, retaining the underlying MoE architecture while prioritizing elevated output throughput for production scenarios. Source excerpts describe it as carrying a 2.4-trillion-parameter mixture-of-experts design and being explicitly engineered for coding, office automation, and long-running agent workflows. Its capability profile includes reasoning, structured output, tool use, and vision, making it suitable for multimodal assistants that must both parse visual input and call external functions in extended sessions.
Practically, the model fits developer and enterprise pipelines that demand sustained agent execution, since the high-throughput framing targets the kind of repeated tool invocations and multi-step coding runs that slow down slower models. A third-party reasoning index scores Qwen 3.8 Max Prime at the top of its reasoning comparison set, underscoring its strength in deliberate, multi-step problem solving alongside its image and document understanding. This makes it a sensible pick when chain-of-thought depth and tool orchestration matter more than minimizing per-token spend.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- alibaba/qwen3.8-max-prime
- Release date
- Sep 23, 2026
- Last updated
- Sep 23, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $4.00
- Output token cost
- $12.00
Limits
- Output tokens
- 131,072 tokens
- Context window
- 1,000,000 tokens
Transparent token rates
Compare Qwen 3.8 Max Prime pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen 3.8 Max Prime
No articles yet. Fetch the latest news to show it here.