Sulat.com
AI models
Ofox logo

Model details

Qwen3.8 Flash

Qwen3.8 Flash was introduced by Alibaba's Qwen team as a multimodal model designed to handle text along with visual inputs while delivering stronger coding and office-task performance. According to the launch announcement, the model was trained at roughly one-ninth the cost of the previous Qwen3.7-Plus, signaling an efficiency-focused step in the Qwen lineup rather than a simple scale increase. Its developer-facing positioning centers on serving real productivity workflows such as software development, document handling, and complex research tasks where balanced reasoning and tool use matter more than raw scale.

A defining practical strength is the model's very long working memory: it ships with a default context window that can be extended to handle large files, lengthy conversations, and extensive research materials, which makes it well suited to agent-style and retrieval-heavy applications. Within the broader release, Qwen3.8-Flash-Next was published as an open-weight sibling previewing the next-generation Qwen4 architecture, giving the developer community an early look at the direction of the family. Together, these releases suggest Qwen is iterating toward more cost-efficient multimodal systems that can still manage demanding, long-context workloads.

Ofoxqwen/qwen3.8-flashqwen

Quick Info

Powered by
Provider
Ofox
Model key
qwen/qwen3.8-flash
Release date
Aug 26, 2026
Last updated
Aug 26, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.11
Output token cost
$0.39

Limits

Output tokens
131,072 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen3.8 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.8 Flash

CrossModel

Coverage

Reuters reports that Alibaba's Qwen team released the multimodal Qwen3.8-Flash model on August 26, 2026, positioning it as a cost-efficient upgrade over its predecessor. According to Qwen's statement published on social media, the model supports a default context window of 262,144 tokens, expandable to 1 million tokens The article notes that Qwen would charge 1 yuan (approximately $0.1488) per million input tokens and 3 yuan per million output tokens for access through its application programming interface. In addition, Qwen released open-source weights for a separate Qwen3.8-Flash-Next variant, allowing the developer community to ev

Videos about Qwen3.8 Flash

More models around Qwen3.8 Flash