CrossModel
The Decoder reports that Qwen3.8-Omni-Flash is Qwen's first multimodal model built for AI agents, processing audio and video together to draw conclusions and edit vlogs, translate short videos, or summarize movies on its own. The context window spans one million tokens, and Qwen positions the model as coming close to m The article describes an agentic video-editing demo in which Qwen3.8-Omni-Flash watches a vlog, plans an edit, and uses tools to deliver a finished short, alongside the open-source Qwen-MM-Plugins that add video editing, speaker recognition, and PDF video notes to agents like Claude Code, Gemini CLI, and Qwen Code, wit