Alibaba (China)
Alibaba's Qwen team has released Qwen3.8-Omni-Flash, its first omni-modal model built around agentic capabilities. It accepts text, images, audio, and video inputs and returns text only, unifying audio-video understanding, reasoning, and tool use in one model. The architecture builds on the Qwen3.8-Flash-Next base that shipped with open weights in August 2026. The model offers a 1M-token context window, with QwenCloud listing 991K max input and 131K max output, plus 262K max reasoning length. Thinking is on by default with reasoning effort set to xhigh. It is available as a hosted API on QwenCloud, Alibaba Cloud Model Studio, and Qwen Studio, though no open weights were announced at launch.