Currently listed through these providers:
Model details
Qwen3.8 Flash
Qwen3.8 Flash is positioned as an experimental preview of the architecture that Qwen plans to carry forward into Qwen4, with the model's Transformers configuration already identifying it under that next-generation lineage. It is built as a multimodal mixture-of-experts design that activates only a small fraction of its total parameters per token, paired with a large n-gram embedding memory and a lightweight speculative decoding head. The result is a system that aims to keep active compute low while still serving images, video frames, and text in a single workflow, which makes it attractive for developers who want richer input handling than a plain dense LLM without paying the full inference cost on every call.
In practice the model is shaped for reasoning-heavy and tool-using applications. Inference guides highlight native support for tool calling and reasoning parsing, so Qwen3.8 Flash can plug into agent pipelines that need structured actions and chain-of-thought traces alongside its multimodal understanding. The mixture-of-experts routing and n-gram memory component give it a recognizable bridge toward Qwen4, so teams adopting it now can get a feel for the next-generation behavior on long-context, multi-input workloads. For users comparing gateway-hosted models, it is a sensible choice when multimodal input, reasoning, and tool integration matter more than raw dense-model throughput, and a less obvious fit for tasks that only need simple text completion at minimum cost.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- qwen3.8-flash
- Release date
- Aug 26, 2026
- Last updated
- Aug 26, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.15
- Output token cost
- $0.47
Limits
- Output tokens
- 131,072 tokens
- Context window
- 1,000,000 tokens
Latest news about Qwen3.8 Flash
Videos about Qwen3.8 Flash
More models around Qwen3.8 Flash
This exact model name is also listed by 14 other providers.