Currently listed through these providers:
Model details
Qwen3 VL 235B A22B Thinking
Qwen3 VL 235B A22B Thinking sits at the top of the Qwen3-VL family as the largest vision-language offering in the series, described by platform listings as the most capable model in the Qwen lineup to date. It uses a Mixture-of-Experts architecture identified as qwen3vlmoe with a total of 236 billion parameters, packaged on community distributions under the Apache License 2.0. A thinking-oriented post-training variant places it alongside a sibling instruct release, giving users two complementary modes for visual and textual reasoning tasks.
In practical terms, the model targets workflows that pair image understanding with deliberate multi-step reasoning, drawing on the Qwen3-VL lineage's expanded capabilities in perception and text generation. The open-weight Apache 2.0 release lowers the barrier for self-hosting and fine-tuning, while the MoE design lets the full 236B-parameter model activate only a subset of experts per token, balancing capacity with inference cost. Its fit is strongest for teams that need a vision-language backbone with explicit reasoning behavior and the freedom to run or adapt the weights outside of closed APIs.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- qwen/qwen3-vl-235b-a22b-thinking
- Release date
- Sep 23, 2025
- Last updated
- Sep 23, 2025
- Knowledge cutoff
- 2025-03-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.40
- Output token cost
- $4.00
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen3 VL 235B A22B Thinking pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 VL 235B A22B Thinking
No articles yet. Fetch the latest news to show it here.