Currently listed through these providers:
Model details
Qwen: Qwen3 VL 8B Thinking (retires Oct 9)
Qwen3 VL 8B Thinking is the reasoning-enhanced edition of Alibaba Cloud's Qwen3 vision-language family, built to process images and text together while thinking through complex problems step by step. Unlike standard vision-language models, this variant introduces deliberate reasoning pathways that let it break down visual questions into logical chains, making it notably stronger at multi-step visual reasoning, STEM problem-solving, and causal analysis over image or video inputs. The architecture combines vision encoding with language modeling at 8 billion parameters, supporting native 256K context that can expand to 1M tokens for processing entire books or hours of video with full recall. It handles temporal sequences through Interleaved-MRoPE and timestamp-aware embeddings, enabling second-level indexing in video understanding. The model also brings practical upgrades like operating PC and mobile GUIs, recognizing a wide range of visual content from landmarks to products, and supporting OCR across 32 languages including difficult conditions like low light and blur.
The Thinking variant builds on the base Qwen3-VL architecture by adding deeper visual-language fusion and extended thinking capabilities compared to the Instruct edition, improving performance specifically on long-chain logic tasks and scientific visual analysis. It achieves text-generation quality on par with large text-only language models while maintaining multimodal understanding, so it can discuss images with the same fluency as a dedicated language model discusses text. The combination of extended reasoning, strong spatial perception for 2D and 3D grounding, and lossless text-vision fusion positions this model for workflows that require both visual comprehension and structured analytical thinking, from research document analysis to embodied AI tasks that need to understand and act on visual environments.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- qwen/qwen3-vl-8b-thinking
- Release date
- Oct 14, 2025
- Last updated
- Oct 14, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.18
- Output token cost
- $2.10
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen: Qwen3 VL 8B Thinking (retires Oct 9) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen: Qwen3 VL 8B Thinking (retires Oct 9)
No articles yet. Fetch the latest news to show it here.