Currently listed through these providers:
Model details
Qwen3.5 122B-A10B
The Qwen3.5-122B-A10B is a hybrid mixture-of-experts model built to handle text, image, and video inputs within a single architecture. Its design blends linear attention with sparse expert routing, activating a targeted subset of its 10 billion parameters per token while keeping the full 122 billion parameter count available for capacity. This architecture supports a 262K-token context window, enabling complex long-document reasoning and extended multimodal conversations. Grouped-query attention with 32 query heads and 2 key-value heads keeps memory and compute manageable at scale, while RoPE position embedding with a high theta value supports the model's long-context capabilities. As an open-weight model released under Apache 2.0, it is designed to be accessible for developers building native multimodal agent applications.
Positioned as the mid-tier offering in the Qwen3.5 series, this model sits just below the larger 397B variant and belongs to a February 2026 release of four lightweight models in the series. Performance benchmarks indicate text capabilities that significantly surpass earlier Qwen3 models, along with visual reasoning that exceeds comparable vision-language models, making this tier particularly well-suited for developers building native multimodal agent applications. Community testing on single DGX Spark hardware has demonstrated throughput reaching 51 tokens per second, showing that the model delivers strong capability without requiring prohibitively expensive infrastructure. This balance of high performance, open accessibility, and practical deployment requirements positions the model as a compelling choice for teams building next-generation multimodal AI systems.
Quick Info
Powered by- Provider
- Cortecs
- Model key
- qwen3.5-122b-a10b
- Release date
- Feb 23, 2026
- Last updated
- Feb 23, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.495
- Output token cost
- $3.46
Limits
- Output tokens
- 262,144 tokens
- Context window
- 262,144 tokens