Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Qwen3.5-Flash

The Qwen3.5-Flash model is built on a hybrid architecture that combines linear attention with a sparse mixture-of-experts design, a structural choice that prioritizes inference efficiency without sacrificing capability. It represents the production-hosted, closed-source API version of the Qwen3.5-35B-A3B model, meaning it shares the same intelligence foundation as that open-weight counterpart while being served through Alibaba Cloud Model Studio for immediate API access. Compared to the earlier Qwen3 series, this generation marks a meaningful leap forward in both pure text and multimodal performance, handling text, image, and video inputs natively. The Flash tier is specifically optimized for speed and throughput, making it well-suited for agentic workflows where low latency and fast response times matter.

The model's alignment with the Qwen3.5-35B-A3B checkpoint means it inherits the capabilities developed through the Qwen3.5 training pipeline while being packaged for production use with built-in tool support and function calling. It carries a default context configuration that enables working with large documents and codebases without additional setup, and it supports over 200 languages for global use cases. Performance benchmarks show particular strength in finance and programming domains, positioning it as a practical choice for knowledge work and technical tasks. At a fraction of the cost of comparable flagship models, Qwen3.5-Flash targets developers and teams seeking frontier-adjacent intelligence in a cost-effective, ready-to-deploy package rather than self-hosted infrastructure.

OpenRouterqwen/qwen3.5-flash-02-23qwen

Quick Info

Powered by
Provider
OpenRouter
Model key
qwen/qwen3.5-flash-02-23
Release date
Feb 25, 2026
Last updated
Feb 25, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.065
Output token cost
$0.26

Limits

Output tokens
65,536 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen3.5-Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5-Flash

OpenRouter

CoverageBenchmark

Alibaba's Qwen team (Tongyi Lab) announced the Qwen 3.5 Medium Model Series on February 25, 2026, explicitly naming Qwen3.5-Flash alongside Qwen3.5-122B-A10B, Qwen3.5-27B, and Qwen3.5-35B-A3B as four new open models. The article embeds official posts from @Alibaba_Qwen and @Ali_TongyiLab confirming the release and the positioning of more intelligence with less compute. The launch notes highlight Qwen3.5-35B-A3B surpassing the prior Qwen3-235B-A22B-2507 and Qwen3-VL-235B-A22B models as evidence of architectural and data quality improvements. Coverage is dated February 26, 2026, roughly seven months before the current date, and includes machine-translated English text from the original Japanese report.

OpenRouter

Official sourceComparison

Compare Qwen3.5-Flash from Qwen and Qwen3.6 Plus Preview from Qwen on key metrics including benchmarks, price, context length, and other model features.

OpenRouter

Official sourceComparison

Compare Qwen3.5-Flash from Qwen and Grok 4.1 Fast from xAI on key metrics including benchmarks, price, context length, and other model features.

OpenRouter

Official sourceComparison

Compare Qwen3.5-Flash from Qwen and Qwen3.5 Plus 2026-02-15 from Qwen on key metrics including benchmarks, price, context length, and other model features.

OpenRouter

CoverageBenchmark

Detailed breakdown of Qwen3.5-Flash including features, pricing, benchmarks, and performance analysis. Last updated in April 2026.

Videos about Qwen3.5-Flash

More models around Qwen3.5-Flash