Currently listed through these providers:
Model details
Qwen3.5 35B-A3B
Qwen3.5 35B-A3B is a sparse mixture-of-experts model in the Qwen3.5 family, designed as a reasoning vision-language model that supports tool use. It carries 35 billion total parameters with only 3 billion activated per token, a design that lets the model deliver substantial capability while keeping active compute low. According to third-party listings, it is pitched as outperforming previous-generation models more than six times its active size, signaling an efficiency-over-parameter emphasis rather than raw scale. The LM Studio listing for the model tags it with Vision Input, reasoning, and trained tool-use capabilities, and indicates a minimum system memory of around 21 GB, which is consistent with the 3B active footprint in practice.
The Qwen team's own communications later position Qwen3.5 35B-A3B as the direct predecessor to the open-sourced Qwen3.6-35B-A3B, which is described as a sparse MoE with the same 35B total / 3B active shape and multimodal thinking plus non-thinking modes. That successor is said to surpass Qwen3.5 35B-A3B by a wide margin on agentic coding and to rival much larger dense models, suggesting this model family is optimized for code agents and tool-driven workflows rather than long-tail open chat. For practitioners, Qwen3.5 35B-A3B fits workloads that need reasoning and tool use without paying for full dense inference, while teams planning ahead may want to track the newer 3.6 release for the freshest agentic coding gains.
Quick Info
Powered by- Provider
- SiliconFlow
- Model key
- Qwen/Qwen3.5-35B-A3B
- Release date
- Feb 23, 2026
- Last updated
- Feb 23, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.24
- Output token cost
- $1.80
Limits
- Output tokens
- 262,144 tokens
- Context window
- 262,144 tokens