Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
EmpirioLabs AI logo

Model details

Qwen3.5 9B

Designed as a mid-size open-weights release in the Qwen family, Qwen3.5 9B blends a hybrid architecture of Gated Delta Networks with a sparse Mixture-of-Experts, a combination aimed at high-throughput inference without heavy latency overhead. Training follows an early-fusion approach over multimodal tokens, which the model card credits with closing the gap to the text-only Qwen3 generation and surpassing prior Qwen3-VL checkpoints on reasoning, coding, agent, and visual understanding evaluations. The result is a single weights package that handles text, image, and video inputs while emitting text, exposing reasoning, tool calling, and structured-output behaviors so it can plug into agent pipelines as well as conventional chat workloads. Published under Apache License 2.0 with post-trained weights available in the Hugging Face Transformers format, it serves developers who want a balance between capability and on-device footprint, with community distributions such as Ollama packaging the 9.65B-parameter model in a 6.6 GB Q4_K_M quantization for local use.

In practice, Qwen3.5 9B fits teams that need a multimodal reasoning model small enough to run on modest hardware yet rich enough to coordinate tools and produce structured responses. The hybrid DeltaNet plus MoE design targets workloads that benefit from long contextual reasoning paired with image or video grounding, making it a practical choice for document understanding, agent orchestration, and multilingual applications, supported linguistically by a notably wide language and dialect coverage. Because the weights are openly licensed, it can be self-hosted, fine-tuned, or integrated into retrieval and tool-augmented stacks, giving organizations flexibility to control cost, latency, and data residency while still benefiting from the family-wide advances in scalable agent reinforcement learning and next-generation multimodal training infrastructure.

EmpirioLabs AIqwen3-5-9bqwen

Quick Info

Powered by
Provider
EmpirioLabs AI
Model key
qwen3-5-9b
Release date
Feb 23, 2026
Last updated
Feb 23, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.09
Output token cost
$0.13

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.5 9B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5 9B

EmpirioLabs AI

Coverage

The Ollama library page for qwen3.5:9b confirms the model is part of the Qwen 3.5 open-source multimodal family, distributed as a 6.6GB Q4_K_M quantization with 9.65B parameters under Apache License 2.0. The page lists companion tags spanning 0.8B through 122B sizes, positioning the 9B within the small-series tier. The According to the page's documentation, Qwen3.5 features a unified Vision-Language foundation with early-fusion training on multimodal tokens, an efficient hybrid architecture combining Gated Delta Networks with sparse Mixture-of-Experts, and scalable reinforcement learning across million-agent environments. It supports

Videos about Qwen3.5 9B

More models around Qwen3.5 9B