Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OrcaRouter logo

Model details

Qwen3.6 35B-A3B

Qwen3.6 35B-A3B is positioned as a mid-sized open-weight model in the Qwen family that emphasizes stability and real-world utility rather than headline-grabbing scale. The naming and the "35moe" tag indicate a Mixture-of-Experts design at the 35B tier, suggesting a routing structure where only a subset of experts is active per token. On the LM Studio hub, where the model is publicly distributed and has been downloaded millions of times, the listing frames it as offering developers a more intuitive, responsive, and genuinely productive coding experience, reinforcing that practical coding workflows are the primary target rather than pure research benchmarks.

Beyond its coding focus, Qwen3.6 35B-A3B is presented as a multimodal-capable assistant with vision input alongside tool-use and reasoning support, making it suitable for workflows that combine screenshots, diagrams, or other visual context with text-driven reasoning and agent-style tool calls. Local deployment is approachable for serious hobbyists and small teams, with a stated minimum system memory of roughly 20GB, while the long context window shown in the catalog supports extended codebases, documentation retrieval, and multi-step agent sessions. For teams already invested in open-weight stacks who want a coding-oriented MoE model with multimodal and tool-calling capability, this variant offers a pragmatic balance between footprint, capability, and day-to-day developer usefulness.

OrcaRouterqwen/qwen3.6-35b-a3bqwen

Quick Info

Powered by
Provider
OrcaRouter
Model key
qwen/qwen3.6-35b-a3b
Release date
Apr 17, 2026
Last updated
Apr 17, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.248
Output token cost
$1.485

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.6 35B-A3B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.6 35B-A3B

OpenRouter

CoverageBenchmark

The OpenRouter listing for qwen/qwen3.6-35b-a3b confirms the model as an open-weight multimodal release from Alibaba Cloud with 35 billion total parameters and 3 billion active per token, using a hybrid sparse MoE that combines Gated DeltaNet linear attention with standard gated attention layers for efficient inference The listing records the model's OpenRouter release date as April 27, 2026, with listed pricing of $0.05 per 1M input tokens and $0.70 per 1M output tokens. OpenRouter also displays a weighted-average effective price based on what customers actually pay across its routed providers, along with a price-history chart spann

OpenRouter

Coverage

On April 2, 2026, Alibaba's Qwen team open-sourced Qwen3.6-35B-A3B as the first open-weight variant of the Qwen3.6 generation, releasing it under Apache 2.0 on Hugging Face alongside the proprietary Qwen3.6-Plus API model. The page frames the release with the tagline "Agentic Coding Power, Now Open to All," noting the Architecturally, Qwen3.6-35B-A3B is a sparse MoE with 256 experts where 8 routed plus 1 shared expert activate per token, yielding 3B active parameters out of 35B total. The 40-layer stack uses a repeating block of three Gated DeltaNet (linear attention) layers followed by one Gated Attention layer, each paired with an

Videos about Qwen3.6 35B-A3B

More models around Qwen3.6 35B-A3B