Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Deep Infra logo

Model details

Qwen 3.5 397B A17B

Qwen3.5-397B-A17B marks a significant step in Alibaba's open-weight strategy by combining a sparse Mixture-of-Experts architecture with native multimodal training from the ground up rather than bolting on vision capabilities. The hybrid design layers Gated Delta Networks linear attention alongside sparse expert routing, allowing only 17 billion parameters to activate per forward pass despite the model's massive total scale—meaning developers get flagship-tier reasoning and generation at roughly 4% of the compute cost. The architecture also brings impressively fast decoding that reaches 8.6 times the throughput of earlier Qwen3-Max at standard context lengths and scales to 19 times faster at 256K context, making large workload handling substantially more practical than with earlier dense models.

The open-weight release reflects Alibaba's push into the agentic AI era, with Qwen3.5-397B-A17B positioned as the most capable model in the Qwen3.5 series and designed to handle autonomous task execution across desktop and mobile interfaces. Built on reinforcement learning scale principles, it brings reasoning and thinking modes alongside tool calling with MCP integration. Its broad language support spanning over 200 languages and YaRN-extendable context window up to 1 million tokens underscore a design philosophy centered on both global accessibility and long-horizon reasoning. For developers and enterprises, the combination of native multimodal fusion, sparse efficiency, and open deployment options on standard hardware like consumer GPUs makes this model a practical bridge between research-grade capability and real-world accessibility.

Deep InfraQwen/Qwen3.5-397B-A17Bqwen

Quick Info

Powered by
Provider
Deep Infra
Model key
Qwen/Qwen3.5-397B-A17B
Release date
Feb 1, 2026
Last updated
Apr 20, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.45
Output token cost
$3.00

Limits

Output tokens
81,920 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen 3.5 397B A17B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen 3.5 397B A17B

Deep Infra

Official sourceOfficial

Qwen3.5-397B-A17B is Alibaba's most capable Qwen3.5 model, a Mixture-of-Experts architecture with 397B total parameters and 17B activated per token. It features a 262K token context window (extensible to 1M with YaRN), thinking/reasoning mode, tool calling with MCP integration, and support for 201 languages. Sets state

Videos about Qwen 3.5 397B A17B

More models around Qwen 3.5 397B A17B