Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Jalapeno Cloud logo

Model details

Qwen3 VL 235B A22B Instruct

Qwen3 VL 235B A22B Instruct stands as the most capable vision-language release in the Qwen family to date, blending a large 235B-parameter mixture-of-experts backbone with 22B active parameters per token. This sparse MoE design lets the model handle multimodal chat, single-image understanding, multi-image reasoning, OCR-style visual question answering, and long-context generation within a single architecture. The combination of high total capacity with selective activation aims to deliver strong multimodal reasoning while keeping inference costs lower than a comparably sized dense model would require.

Distributed as an open-weight artifact under Apache License 2.0, the model is available through community runtimes such as Ollama, where the qwen3-vl:235b-a22b-instruct tag ships around 143 GB of Q4_K_M quantized weights, and through serving stacks like vLLM Ascend for production-grade deployments. Practical strengths include breadth across vision and text tasks, support for long-context workflows, and flexibility for self-hosting thanks to its open licensing. The trade-off is substantial hardware demand: running the full 235B-parameter MoE requires significant memory and compute, making it best suited to teams with dedicated GPU capacity who need top-tier multimodal reasoning rather than lightweight edge deployments.

Jalapeno CloudQwen3-VL-235B-A22B-Instructqwen

Quick Info

Powered by
Provider
Jalapeno Cloud
Model key
Qwen3-VL-235B-A22B-Instruct
Release date
Sep 23, 2025
Last updated
Sep 23, 2025
Knowledge cutoff
2025-03-31
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$1.50

Limits

Output tokens
32,768 tokens
Context window
129,024 tokens

Transparent token rates

Compare Qwen3 VL 235B A22B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 VL 235B A22B Instruct

Jalapeno Cloud

CoverageBenchmark

Roboflow's Playground page independently corroborates the identity and provenance of the underlying Qwen3-VL-235B-A22B-Instruct model that Jalapeno is hosting. It describes the model as an open-weight multimodal vision-language model from Qwen/Alibaba, built on a mixture-of-experts architecture with approximately 22B a The page additionally reports Roboflow-side Vision Evals results, with an overall score of 65.8 (ranked 25 of 34, evals updated August 27, 2026), and 47 inferences logged in the prior 30 days with an average latency of 10.44 seconds. These benchmark and performance figures reflect Roboflow's own deployment and evaluati

Videos about Qwen3 VL 235B A22B Instruct

More models around Qwen3 VL 235B A22B Instruct