Alibaba
Alibaba's Qwen team has released Qwen3.8-Omni-Flash, its first omni-modal model built around agentic capabilities. It accepts text, images, audio, and video as inputs and returns text, combining audio-video understanding, reasoning, and tool use in a single model. The model is positioned for workflows that understand content, plan tasks, execute tools, and deliver results. The model is built on the Qwen3.8-Flash-Next architecture, which shipped with open weights in August 2026. It offers a 1M-token context window, with QwenCloud listing 991K max input and 131K max output, plus a 262K max reasoning length. Thinking is enabled by default with reasoning effort set to xhigh. It is available as a hosted API on QwenCloud, Alibaba Cloud Model Studio, and Qwen Studio, with no open weights announced at launch.