Currently listed through these providers:
Model details
Qwen3 VL 235B A22B Thinking
Qwen3 VL 235B A22B Thinking is the reasoning-tuned edition of the Qwen3-VL vision-language family, engineered for tasks that require careful, step-by-step analysis rather than quick surface responses. Its MoE design pairs a large total parameter count with a much smaller active subset per token, which lets the model carry broad visual and textual knowledge while keeping inference compute closer to a mid-sized model. SiliconFlow describes the variant as achieving state-of-the-art results across multimodal reasoning benchmarks, particularly in STEM, mathematics, causal analysis, and evidence-based answers, positioning it for use cases where chain-of-thought quality matters more than raw throughput.
In practical deployments the model behaves like a high-end multimodal assistant that can ingest images alongside text and emit structured, explained answers, which suits workflows such as scientific figure analysis, document and diagram reasoning, visual code debugging, and other tasks that blend perception with logical deduction. The weights are openly published under Apache 2.0, and third-party listings confirm a Q4_K_M quantized build near 143 GB for self-hosted local runtimes, giving organizations a path to run advanced multimodal reasoning on their own infrastructure alongside managed API access. This combination of scale, open availability, and explicit reasoning focus makes the model a strong fit for teams that need transparent, auditable multimodal reasoning rather than a general-purpose chatbot.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- qwen/qwen3-vl-235b-a22b-thinking
- Release date
- Sep 23, 2025
- Last updated
- Sep 23, 2025
- Knowledge cutoff
- 2025-03-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.40
- Output token cost
- $4.00
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens