Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

Qwen3.5 27B

Qwen3.5 27B is a native vision-language dense model that incorporates a linear attention mechanism, an architectural choice intended to deliver quicker response times and a better balance between inference speed and overall performance than a purely quadratic attention design would allow. The model is described as part of the Qwen3.5 family, with OpenRouter positioning its overall capabilities as comparable to the larger Qwen3.5-122B-A10B variant, suggesting that the 27B dense configuration aims to approximate the practical reach of a much bigger mixture-of-experts sibling while staying compact enough to deploy on a single high-memory workstation. Weights are openly published on Hugging Face under the Qwen/Qwen3.5-27B repository, which makes the model accessible for self-hosted experimentation, fine-tuning, and integration into private pipelines.

In practical terms, Qwen3.5 27B is shaped for workloads that mix image and text understanding with long-context reasoning, since it supports multimodal input alongside a text-only output stream and operates within a 262K-token context window. The linear attention design is the main qualitative differentiator: it should make the model attractive for interactive assistants, document and screenshot analysis, and agent-style tasks where low latency and predictable throughput matter more than chasing the largest possible parameter count. For teams comparing options inside the Qwen3.5 lineup, the 27B dense variant is best understood as the efficiency-focused member of the family rather than a flagship reasoning model.

Kilo Gatewayqwen/qwen3.5-27bqwen

Quick Info

Powered by
Provider
Kilo Gateway
Model key
qwen/qwen3.5-27b
Release date
Feb 23, 2026
Last updated
Feb 23, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.195
Output token cost
$1.56

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.5 27B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5 27B

Ofox

CoverageBenchmark

Alibaba's Qwen team (Tongyi Lab) announced the Qwen 3.5 Medium Model Series on February 25, 2026, with Qwen3.5-27B explicitly named as one of four medium-tier variants released alongside Qwen3.5-Flash, Qwen3.5-35B-A3B, and Qwen3.5-122B-A10B. The announcement, originally posted on X by both the @Alibaba_Qwen and @Ali_To The GIGAZINE coverage positions Qwen3.5-27B as part of an open model family, with the related claim that Qwen3.5-35B-A3B surpasses the prior Qwen3-235B-A22B-2507 and Qwen3-VL-235B-A22B checkpoints — evidence of better architecture and data quality at smaller parameter counts. While the article is a third-party (machine

Videos about Qwen3.5 27B

More models around Qwen3.5 27B