Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Jalapeno Cloud logo

Model details

Qwen3 VL 235B A22B Thinking

This reasoning-focused member of the Qwen3-VL family uses a Mixture-of-Experts design, with distribution listings reporting roughly 236B parameters. It is intended for work that combines written instructions with visual information, including visual question answering, GUI-oriented agent tasks, and code generation from screenshots or video. Unsloth also provides a chat-template-fixed GGUF distribution for llama.cpp, using its Dynamic 2.0 quantization path.

The model’s practical strengths are deeper visual perception, multimodal reasoning, spatial and grounding tasks, and long-form content analysis. Its distribution materials describe improved text understanding, stronger spatial perception, expanded context handling, and the ability to process books or hours-long video. The Thinking edition is a good fit for applications that can benefit from explicit reasoning and tool interaction, while the Apache 2.0 licensing supports open experimentation and deployment.

Jalapeno CloudQwen3-VL-235B-A22B-Thinkingqwen

Quick Info

Powered by
Provider
Jalapeno Cloud
Model key
Qwen3-VL-235B-A22B-Thinking
Release date
Sep 23, 2025
Last updated
Sep 23, 2025
Knowledge cutoff
2025-03-31
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.98
Output token cost
$3.95

Limits

Output tokens
32,768 tokens
Context window
131,072 tokens

Transparent token rates

Compare Qwen3 VL 235B A22B Thinking pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 VL 235B A22B Thinking

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 VL 235B A22B Thinking

More models around Qwen3 VL 235B A22B Thinking