Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
SiliconFlow (China) logo

Model details

Qwen/Qwen3.5-397B-A17B

Qwen3.5-397B-A17B is a large-scale open-weight foundation model built on a sparse Mixture-of-Experts architecture, giving it 397 billion total parameters while activating only a subset of them during any given forward pass. This dynamic parameter activation means developers can tap into large-model intelligence without the full computational burden of a dense 400B model at every step. The model pairs this MoE design with a hybrid architecture that blends linear attention mechanisms, enabling high-throughput inference with reduced latency and cost overhead. It functions as a native vision-language model, capable of processing text, images, and video inputs within a unified framework that was trained using early fusion techniques across multimodal tokens, achieving cross-generational parity with the standalone Qwen3 series.

The development lineage traces back to Qwen's emphasis on integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and accessibility for developers and enterprises. Qwen3.5 draws on post-training work that has produced artifacts compatible with Hugging Face Transformers, vLLM, SGLang, and other inference engines, making it practical to deploy across varied infrastructure. Its design specifically targets complex AI workflows involving reasoning, coding, and multimodal tasks, where its open-weight nature and efficient hybrid design make it viable for teams that want the power of a frontier-scale model without the typical deployment constraints.

SiliconFlow (China)Qwen/Qwen3.5-397B-A17Bqwen

Quick Info

Powered by
Provider
SiliconFlow (China)
Model key
Qwen/Qwen3.5-397B-A17B
Release date
Feb 16, 2026
Last updated
Feb 16, 2026
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.29
Output token cost
$1.74

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen/Qwen3.5-397B-A17B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen/Qwen3.5-397B-A17B

SiliconFlow (China)

CoverageBenchmark

Benchmark Qwen3.5 397B A17B API latency, throughput, and cost efficiency. Compare response speed, token output, and pricing for scalable AI workloads.

Videos about Qwen/Qwen3.5-397B-A17B

More models around Qwen/Qwen3.5-397B-A17B