Sulat.com
AI models
Requesty logo

Model details

Qwen3.8 Flash Next

Designed as an architectural preview of the upcoming Qwen4 family, this open-weight release blends text, image, and video inputs into a single multimodal mixture-of-experts model from Alibaba's Qwen team. Independent coverage describes it as a substantially redesigned MoE that combines roughly 125B main parameters with an additional pool of N-gram embeddings and only about 6B parameters activated per token, with further upgrades across attention, residual, embedding, and optimization components. The same reporting also notes meaningful reductions in training and inference cost relative to prior Qwen generations, framing the model as an early look at how the next Qwen generation will balance capability and efficiency.

For practitioners exploring cutting-edge open-weight systems, the practical appeal is having a Qwen4-flavored MoE accessible through a familiar API surface rather than self-hosting the full stack. Requesty exposes it under its own catalog id with a 262,144-token context window, making it well suited to long-form document reasoning, multimodal attachments, and tool-calling workflows that benefit from a sizable working window. The single routing path means failover is limited compared with multi-provider listings, and developers should treat the Requesty-specific price point as distinct from the related QwenCloud listing, since the two SKUs carry different rates despite the shared Qwen lineage.

Requestyqwen3.8-flash-nextqwen

Quick Info

Powered by
Provider
Requesty
Model key
qwen3.8-flash-next
Release date
Aug 27, 2026
Last updated
Aug 27, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.20
Output token cost
$0.50

Limits

Output tokens
262,144 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.8 Flash Next pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.8 Flash Next

Requesty

CoverageBenchmark

DataCamp's technical write-up, dated August 27, 2026, characterizes Qwen3.8-Flash-Next as an open-weight, multimodal mixture-of-experts model from Alibaba's Qwen team, released August 26, 2026, and positioned as an early preview of the architecture behind the upcoming Qwen4 family. Reported specs include a 125B-paramet The article quotes QwenCloud pricing of $0.16 input and $0.47 output per million tokens under the qwen3.8-flash listing id, which differs from Requesty's $0.20/$0.50 listing for the qwen3.8-flash-next SKU, so the two price points should not be conflated when reasoning about Requesty's rate specifically. Vendor-reported

Requesty

CoveragePreview

An AI-driven lab explainer on note.com, dated August 26, 2026, frames Qwen3.8-Flash-Next as an open-weight multimodal MoE model intended to preview the Qwen4 architecture, scheduled for release on ModelScope at 23:00 China Standard Time on August 26, 2026. The piece highlights architectural elements including a GDN Hyb A notable contribution is the author's explicit skepticism about the widely circulated "125B-A6B" parameter figures, noting that technically informed readers find the primary-source evidence surprisingly thin for those exact numbers. The article is presented as an auto-translated AI rendering of a Japanese original, so

Requesty

Official sourceBenchmark

Requesty's live model catalog directly lists qwen3.8-flash-next under Alibaba (Qwen) with a single provider route, a 262,144-token context window, and per-million-token pricing of $0.20 input and $0.50 output. The same catalog also surfaces the sibling SKU qwen3.8-flash at $0.16/$0.47 with a 1.0M context, confirming th For developers, the entry implies the model is reachable through https://router.requesty.ai/v1 with the qwen3.8-flash-next model id, making it directly callable alongside the rest of the catalog without bespoke integration work. Because only one provider is listed for this SKU, failover options appear limited relative

Videos about Qwen3.8 Flash Next

More models around Qwen3.8 Flash Next