Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NEAR AI Cloud logo

Model details

Qwen3.5 397B-A17B

Qwen3.5-397B-A17B is the first open-weights release in Alibaba's Qwen3.5 series, introduced by the Qwen team in mid-February 2026 as a native vision-language foundation model. The Qwen blog and InferenceX both describe it as combining linear attention via Gated Delta Networks with a sparse mixture-of-experts layout, giving 397 billion total parameters while activating only 17 billion per forward pass. The team frames the model for reasoning, coding, agent-style tasks, and multimodal understanding, and positions it as the open counterpart to the hosted Qwen3.5-Plus service that offers a 1M-token default context and adaptive tool use.

The model is distributed as a post-trained checkpoint on Hugging Face under Apache License 2.0, with a mirror on ModelScope and compatibility with Transformers, vLLM, SGLang, and KTransformers, making it practical for self-hosted and research deployments. Its native multimodal framing extends beyond text and images to video and audio understanding along with GUI interaction, and language and dialect support has been expanded from 119 to 201 languages. Together, those traits make Qwen3.5-397B-A17B well suited for agentic applications, long-context reasoning, and multilingual multimodal pipelines where an open-weights checkpoint with a strong activation-efficient MoE design is preferred over a closed hosted endpoint.

NEAR AI Cloudqwen/qwen3.5-397b-a17bqwen

Quick Info

Powered by
Provider
NEAR AI Cloud
Model key
qwen/qwen3.5-397b-a17b
Release date
Feb 15, 2026
Last updated
Feb 15, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.50
Output token cost
$3.30

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.5 397B-A17B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5 397B-A17B

Kilo Gateway

Coverage

Alibaba's Qwen team officially released Qwen3.5-397B-A17B, the first open-weight model in the Qwen3.5 series. It is a native vision-language model delivering strong results across reasoning, coding, agent, and multimodal benchmarks. The hybrid architecture fuses linear attention via Gated Delta Networks with a sparse mixture-of-experts, running 397 billion total parameters with 17 billion activated per forward pass for inference efficiency. Language and dialect support expanded from 119 to 201, broadening global accessibility. On Qwen3.5-397B-A17B's leaderboard evaluations, the model posts 87.8 on MMLU-Pro, 88.4 on GPQA, 76.4 on SWE-bench Verified, 79.0 on MMMU-Pro, 72.9 on the Berkeley Function Calling Leaderboard, 95.6 on τ²-bench Telecom, and 93.33 on AIME 2026, with an AA-Omniscience index of -30 signaling notable hallucination weakness. The companion Qwen3.5-Plus is a separate hosted variant on Alibaba Cloud Model Studio with a 1M-token context and built-in adaptive tools.

Kilo Gateway

CoverageBenchmark

evals.report independently tracks 22 reported benchmark scores for Qwen3.5-397B-A17B released on February 16, 2026, with per-benchmark source-status labels (Official, Verified, or Unverified). Coverage spans SWE-bench Verified, GPQA Diamond, Humanity's Last Exam, Berkeley Function Calling Leaderboard, MMMU-Pro, Artificial Analysis and Epoch Capabilities indices, τ²-bench Telecom, AIME 2026, GDPval, and WebDev Arena. No single combined leaderboard score is published, reflecting a per-benchmark rather than aggregated view. Highlights from the tracker include SWE-bench Verified at 76.4 percent, GPQA Diamond at 88.4 percent, MMMU-Pro at 79.0 percent, AIME 2026 at 93.33 percent, WebDev Arena at 1393 Elo, and an Epoch Capabilities Index of 146.1. The AA-Omniscience score of -30 stands out as a notable hallucination weakness for the variant. The page also flags many scores as Unverified, signaling that independent reproduction is still pending for several measures.

Kilo Gateway

Coverage

Alibaba announced the open-source release of Qwen3.5 on February 16, 2026, with the first model in the series being Qwen3.5-397B-A17B, also branded as the hosted "Qwen3.5-Plus". The model is described as a natively multimodal foundation model trained from scratch on trillions of vision-language tokens spanning multilingual text, images, videos, STEM, and reasoning data, with support for 201 languages and dialects, up from 119 in the Qwen3 series, including low-resource languages such as Hawaiian, Fijian, and Niger-Congo languages. According to Alibaba's announcement, Qwen3.5-397B-A17B targets strong performance across language understanding and reasoning, code generation, agentic workflows, image and video comprehension, and GUI interaction, with results rivaling leading frontier models in versatility and capability. A central design priority is inference efficiency: the model pairs a linear attention mechanism with a sparse mixture-of-experts (MoE) architecture to lower compute requirements while remaining suitable for broad real-world deployment of multimodal agentic applications.

OpenRouter

Coverage

On February 16, 2026, Alibaba officially open-sourced Qwen3.5, with the headline release being the Qwen3.5-397B-A17B variant (also branded "Qwen3.5-Plus"). The model is described as a natively multimodal foundation model that combines a linear attention mechanism with a sparse Mixture-of-Experts (MoE) design, activatin The Qwen3.5-397B-A17B was trained on trillions of vision-language tokens spanning multilingual text, images, videos, STEM, and reasoning data, supports 201 languages and dialects (up from 119 in Qwen3), and processes text, image, and video while generating text. According to Alibaba's announcement, it delivers strong p

TokenGo

Coverage

The Baidu encyclopedia entry for Qwen3.5 names both Qwen3.5-Plus and Qwen3.5-397B-A17B as versions in Alibaba's February 16, 2026 flagship release and describes the underlying architecture as a Hybrid Attention Mechanism combined with a Sparse Mixture-of-Experts design optimized for logical reasoning, mathematical comp The same Baidu entry documents rapid ecosystem adoption following the February 16 launch, noting that by February 18, 2026, international hardware and framework vendors including NVIDIA, AMD, and Apple had completed adaptation for Qwen3.5, while domestic Chinese GPU and platform providers such as Huawei Ascend, Moore T

SiliconFlow

CoverageBenchmark

Qwen3.5-397B-A17B is the first open-weights release in Alibaba's Qwen3.5 series, announced by the Qwen team as a native vision-language foundation model with 397B total parameters and 17B activated per forward pass, per the Alibaba Cloud "Qwen3.5: Towards Native Multimodal Agents" post. The Qwen blog dates the official The model is positioned for reasoning, coding, agent tasks and multimodal understanding, with early-fusion training achieving parity with Qwen3 while outperforming Qwen3-VL on reasoning, coding, agents and visual understanding, per the Hugging Face model card. Independent evaluation placed it at 45 on the Artificial An

Videos about Qwen3.5 397B-A17B

More models around Qwen3.5 397B-A17B