Currently listed through these providers:
Model details
Qwen3 VL 235B A22B Thinking
This reasoning-focused member of the Qwen3-VL family uses a Mixture-of-Experts design, with distribution listings reporting roughly 236B parameters. It is intended for work that combines written instructions with visual information, including visual question answering, GUI-oriented agent tasks, and code generation from screenshots or video. Unsloth also provides a chat-template-fixed GGUF distribution for llama.cpp, using its Dynamic 2.0 quantization path.
The model’s practical strengths are deeper visual perception, multimodal reasoning, spatial and grounding tasks, and long-form content analysis. Its distribution materials describe improved text understanding, stronger spatial perception, expanded context handling, and the ability to process books or hours-long video. The Thinking edition is a good fit for applications that can benefit from explicit reasoning and tool interaction, while the Apache 2.0 licensing supports open experimentation and deployment.
Quick Info
Powered by- Provider
- Jalapeno Cloud
- Model key
- Qwen3-VL-235B-A22B-Thinking
- Release date
- Sep 23, 2025
- Last updated
- Sep 23, 2025
- Knowledge cutoff
- 2025-03-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.98
- Output token cost
- $3.95
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen3 VL 235B A22B Thinking pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 VL 235B A22B Thinking
No articles yet. Fetch the latest news to show it here.
