Currently listed through these providers:
Model details
Qwen3.5 397B-A17B
Qwen3.5 397B A17B is the first open-weight release in the Qwen3.5 series, designed as a native vision-language model for developers and enterprises building productive agents. It pairs multimodal understanding with strong reasoning, coding, and tool-use behavior, making it well suited to applications that blend perception, planning, and code generation rather than pure text tasks. The model also broadens linguistic reach by expanding language and dialect support from 119 to 201, which helps teams serve a wider range of users in a single deployment. The architecture behind the model blends linear attention through Gated Delta Networks with a sparse mixture-of-experts design, allowing it to keep total parameter capacity high while activating only a fraction of those parameters per token. With 397 billion total parameters and 17 billion activated on each forward pass, the design aims to deliver the quality of a very large model while keeping inference cost and latency closer to a mid-sized one. The same post announcing the release reports outstanding results across the full benchmark suite it evaluates, covering reasoning, coding, agent capabilities, and multimodal understanding, positioning the model as a capable base for general-purpose agentic workloads.
In practice, this balance of a heavy expert pool with a small active footprint makes the model attractive for workloads that demand long-horizon reasoning and tool orchestration without paying full dense-model inference costs. The hybrid attention choice helps sustain efficiency on long inputs, which matters for agent loops, document analysis, and code repositories that exceed typical short-context limits. Because the weights are openly released, teams can self-host the model, fine-tune it for domain-specific agents, or run it through hosted offerings such as Qwen3.5-Plus on Alibaba Cloud Model Studio. That combination of open availability, a multilingual footprint, and an efficiency-oriented MoE design makes the model a flexible foundation for organizations building production-grade multimodal assistants.
Quick Info
Powered by- Provider
- Cortecs
- Model key
- qwen3.5-397b-a17b
- Release date
- Feb 15, 2026
- Last updated
- Feb 15, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.668
- Output token cost
- $4.01
Limits
- Output tokens
- 250,000 tokens
- Context window
- 262,000 tokens