Model details
Qwen3 VL 235B A22B Thinking
Qwen3 VL 235B A22B Thinking sits at the top of the Qwen3-VL family as the largest vision-language offering in the series, described by platform listings as the most capable model in the Qwen lineup to date. It uses a Mixture-of-Experts architecture identified as qwen3vlmoe with a total of 236 billion parameters, packaged on community distributions under the Apache License 2.0. A thinking-oriented post-training variant places it alongside a sibling instruct release, giving users two complementary modes for visual and textual reasoning tasks.
In practical terms, the model targets workflows that pair image understanding with deliberate multi-step reasoning, drawing on the Qwen3-VL lineage's expanded capabilities in perception and text generation. The open-weight Apache 2.0 release lowers the barrier for self-hosting and fine-tuning, while the MoE design lets the full 236B-parameter model activate only a subset of experts per token, balancing capacity with inference cost. Its fit is strongest for teams that need a vision-language backbone with explicit reasoning behavior and the freedom to run or adapt the weights outside of closed APIs.
Quick Info
Powered by- Provider
- NovitaAI
- Model key
- qwen/qwen3-vl-235b-a22b-thinking
- Release date
- Sep 24, 2025
- Last updated
- Sep 24, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.98
- Output token cost
- $3.95
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Latest news about Qwen3 VL 235B A22B Thinking
No articles yet. Fetch the latest news to show it here.