Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
EmpirioLabs AI logo

Model details

Qwen3.5 397B-A17B

Qwen3.5-397B-A17B is a large-scale Mixture-of-Experts model in the Qwen3.5 family, pairing 397B total parameters with a much smaller active footprint per token. The architecture is explicitly positioned as multimodal, letting a single deployment handle text alongside image and video inputs while still producing text outputs. A vLLM-Ascend deployment tutorial describes it as combining multimodal capability, long-context inference, MTP speculative decoding, and W8A8 quantized deployment for production serving on Ascend hardware, signaling that the model is meant to be dropped into production inference stacks rather than treated as a research artifact.

Because the active parameter count is kept low relative to total size, the model is shaped for efficient inference at scale, and the MTP speculative decoding plus W8A8 quantization path reinforces that focus on throughput-friendly serving. Long-context inference is a first-class use case, making it well suited to workloads such as document analysis, video understanding, and retrieval-heavy reasoning where extended inputs are common. NVIDIA's NGC catalog lists the model under the qwen organization for NIM-based deployment, and the vLLM-Ascend project recorded first support in v0.17.0rc1, giving teams a validated open serving path on Ascend accelerators with tooling that covers single-node, multi-node, and Prefill-Decode disaggregated setups.

EmpirioLabs AIqwen3-5-397b-a17bqwen

Quick Info

Powered by
Provider
EmpirioLabs AI
Model key
qwen3-5-397b-a17b
Release date
Feb 15, 2026
Last updated
Feb 15, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.172
Output token cost
$1.032

Limits

Output tokens
64,000 tokens
Context window
256,000 tokens

Transparent token rates

Compare Qwen3.5 397B-A17B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5 397B-A17B

EmpirioLabs AI

Coverage

On February 16, 2026, Alibaba officially open-sourced Qwen3.5-397B-A17B (also known as "Qwen3.5-Plus"), the first release in the Qwen3.5 series. The post describes the model as a natively multimodal foundation model with strong capabilities in reasoning, coding, agent workflows, and multimodal understanding, explicitly Technically, Qwen3.5-397B-A17B is trained as a natively multimodal model on trillions of vision-language tokens spanning multilingual text, images, videos, STEM, and reasoning data, and can process text, image, and video while generating text. Language support expanded to 201 languages and dialects (up from 119 in Qwen

EmpirioLabs AI

CoverageBenchmark

An independent technical analysis characterizes Qwen3.5-397B-A17B as the first open-weights release in Alibaba's Qwen3.5 series, announced as a native vision-language foundation model with 397 billion total parameters and 17 billion activated per forward pass. Weights were published on Hugging Face as a post-trained ch The analysis highlights the model's agentic, multimodal positioning targeting reasoning, coding, agent capabilities, and multimodal understanding, with early-fusion multimodal training achieving cross-generational parity with Qwen3 while outperforming the Qwen3-VL line on reasoning, coding, agents, and visual understan

EmpirioLabs AI

Coverage

According to a Baidu Baike entry, Qwen3.5 is a flagship large language model series launched by Alibaba on February 16, 2026, comprising two versions: Qwen3.5-Plus and Qwen3.5-397B-A17B. The series supports text and multimodal tasks, using a hybrid attention mechanism and a sparse Mixture-of-Experts architecture, and i Qwen3.5-397B-A17B has 397 billion total parameters with 17 billion activated, reducing GPU memory consumption for deployment by 60%. Alibaba announced the development plan in January 2026, and on February 9 a code merge request appeared on the Hugging Face open-source project page. The Qwen3.5-Plus API was priced at 0.

Videos about Qwen3.5 397B-A17B

More models around Qwen3.5 397B-A17B