Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Ofox logo

Model details

Qwen3.5 397B-A17B

Qwen3.5-397B-A17B is the first open-weights release in Alibaba's Qwen3.5 series, introduced by the Qwen team in mid-February 2026 as a native vision-language foundation model. The Qwen blog and InferenceX both describe it as combining linear attention via Gated Delta Networks with a sparse mixture-of-experts layout, giving 397 billion total parameters while activating only 17 billion per forward pass. The team frames the model for reasoning, coding, agent-style tasks, and multimodal understanding, and positions it as the open counterpart to the hosted Qwen3.5-Plus service that offers a 1M-token default context and adaptive tool use.

The model is distributed as a post-trained checkpoint on Hugging Face under Apache License 2.0, with a mirror on ModelScope and compatibility with Transformers, vLLM, SGLang, and KTransformers, making it practical for self-hosted and research deployments. Its native multimodal framing extends beyond text and images to video and audio understanding along with GUI interaction, and language and dialect support has been expanded from 119 to 201 languages. Together, those traits make Qwen3.5-397B-A17B well suited for agentic applications, long-context reasoning, and multilingual multimodal pipelines where an open-weights checkpoint with a strong activation-efficient MoE design is preferred over a closed hosted endpoint.

Ofoxqwen/qwen3.5-397b-a17bqwen

Quick Info

Powered by
Provider
Ofox
Model key
qwen/qwen3.5-397b-a17b
Release date
Feb 15, 2026
Last updated
Feb 15, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.55
Output token cost
$3.50

Limits

Output tokens
64,000 tokens
Context window
256,000 tokens

Transparent token rates

Compare Qwen3.5 397B-A17B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5 397B-A17B

Kilo Gateway

Coverage

Alibaba announced the open-source release of Qwen3.5 on February 16, 2026, with the first model in the series being Qwen3.5-397B-A17B, also branded as the hosted "Qwen3.5-Plus". The model is described as a natively multimodal foundation model trained from scratch on trillions of vision-language tokens spanning multilingual text, images, videos, STEM, and reasoning data, with support for 201 languages and dialects, up from 119 in the Qwen3 series, including low-resource languages such as Hawaiian, Fijian, and Niger-Congo languages. According to Alibaba's announcement, Qwen3.5-397B-A17B targets strong performance across language understanding and reasoning, code generation, agentic workflows, image and video comprehension, and GUI interaction, with results rivaling leading frontier models in versatility and capability. A central design priority is inference efficiency: the model pairs a linear attention mechanism with a sparse mixture-of-experts (MoE) architecture to lower compute requirements while remaining suitable for broad real-world deployment of multimodal agentic applications.

OpenRouter

Coverage

On February 16, 2026, Alibaba officially open-sourced Qwen3.5, with the headline release being the Qwen3.5-397B-A17B variant (also branded "Qwen3.5-Plus"). The model is described as a natively multimodal foundation model that combines a linear attention mechanism with a sparse Mixture-of-Experts (MoE) design, activatin The Qwen3.5-397B-A17B was trained on trillions of vision-language tokens spanning multilingual text, images, videos, STEM, and reasoning data, supports 201 languages and dialects (up from 119 in Qwen3), and processes text, image, and video while generating text. According to Alibaba's announcement, it delivers strong p

TokenGo

Coverage

The Baidu encyclopedia entry for Qwen3.5 names both Qwen3.5-Plus and Qwen3.5-397B-A17B as versions in Alibaba's February 16, 2026 flagship release and describes the underlying architecture as a Hybrid Attention Mechanism combined with a Sparse Mixture-of-Experts design optimized for logical reasoning, mathematical comp The same Baidu entry documents rapid ecosystem adoption following the February 16 launch, noting that by February 18, 2026, international hardware and framework vendors including NVIDIA, AMD, and Apple had completed adaptation for Qwen3.5, while domestic Chinese GPU and platform providers such as Huawei Ascend, Moore T

SiliconFlow

CoverageBenchmark

Qwen3.5-397B-A17B is the first open-weights release in Alibaba's Qwen3.5 series, announced by the Qwen team as a native vision-language foundation model with 397B total parameters and 17B activated per forward pass, per the Alibaba Cloud "Qwen3.5: Towards Native Multimodal Agents" post. The Qwen blog dates the official The model is positioned for reasoning, coding, agent tasks and multimodal understanding, with early-fusion training achieving parity with Qwen3 while outperforming Qwen3-VL on reasoning, coding, agents and visual understanding, per the Hugging Face model card. Independent evaluation placed it at 45 on the Artificial An

Videos about Qwen3.5 397B-A17B

More models around Qwen3.5 397B-A17B