Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
ZenMux logo

Model details

Qwen3.6-Plus

Qwen3.6-Plus is a multimodal agentic model from Alibaba's Qwen team that succeeds the Qwen 3.5 Plus series with a hybrid architecture combining efficient linear attention with sparse Mixture of Experts routing. This design choice is central to its ability to handle a the cataloged API limit that quick-info value window while remaining computationally practical. The model ships with chain-of-thought reasoning that is always active, and supports tool use and function calling natively, making it purpose-built for workflows where a model needs to plan, execute, and revise across multiple steps. It processes not just text but also images, PDFs, and video, enabling it to work with UI screenshots, design mockups, and document-heavy pipelines that developers commonly encounter when building AI-driven applications.

The model demonstrates particularly strong performance in agentic coding scenarios, achieving 78.8% on SWE-Bench Verified and 61.6% on Terminal-Bench 2.0, alongside 90.4% GPQA Diamond for scientific reasoning and 86.0% on MMMU for multimodal understanding. A standout feature for developers is its preserve thinking capability, which maintains full reasoning that quick-info value across multi-turn agent sessions, preventing the model from losing track of intermediate conclusions in long workflows. This makes Qwen3.6-Plus well suited for building code review tools, autonomous agents that navigate repositories, or applications that need to reason deeply over large inputs in a single request.

ZenMuxqwen/qwen3.6-plus

Quick Info

Powered by
Provider
ZenMux
Model key
qwen/qwen3.6-plus
Release date
Mar 30, 2026
Last updated
Mar 30, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.50
Output token cost
$3.00

Limits

Output tokens
64,000 tokens
Context window
1,000,000 tokens

Latest news about Qwen3.6-Plus

ZenMux

CoverageBenchmark

Qwen 3.6 Plus is a hybrid-architecture large language model from Alibaba that combines efficient linear attention with sparse mixture-of-experts routing, succeeding the Qwen 3.5 Plus series with major gains in agentic coding, front-end development, and overall reasoning. According to its OpenRouter listing (released Ap On OpenRouter the model is hosted solely by Alibaba Cloud International, with listed pricing of $0.325 per 1M input tokens and $1.95 per 1M output tokens, plus cache-create at $0.4063/M and cache-read at $0.0325/M. Measured P50 latency is 0.82 seconds with throughput of 39 tokens per second, and the weighted-average ef

Videos about Qwen3.6-Plus