Currently listed through these providers:
Model details
Qwen3.5 Plus
Qwen3.5 Plus sits at the top of Alibaba's Qwen3.5 family as a premium tier explicitly framed for the agentic AI era, with a design that openly takes aim at Gemini-class competitors. Rather than splitting language and vision into separate models, the series unifies them into a single native vision-language foundation, and the Plus variant extends this with a one-the cataloged API limit suitable for long-running agent sessions, large codebases, and extended multimodal documents. Its intended role is therefore that of a general-purpose flagship rather than a narrow specialist, aimed at developers building assistants, search, and tool-using workflows where breadth and context depth matter.
Under the hood, the Plus tier is described as using a hybrid architecture that pairs linear attention with sparse mixture-of-experts routing, a combination positioned to lift inference efficiency while preserving the capacity needed for multimodal reasoning. This shift from the preceding generation represents a meaningful step forward in both pure-text and vision-language task evaluations, with Alibaba highlighting performance comparable to current frontier systems. Practically, the model fits teams that want long-context multimodal understanding for agentic pipelines and are willing to trade the openness of earlier Qwen releases for a hosted premium endpoint.
Quick Info
Powered by- Provider
- LLMTR
- Model key
- qwen/qwen3.5-plus
- Release date
- Feb 16, 2026
- Last updated
- Feb 16, 2026
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.40
- Output token cost
- $2.40
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,000,000 tokens