According to Wallstreetcn, Alibaba Cloud's next-generation native full-modal model Qwen3.8-Omni-Flash has launched, with its core goal to improve agent capabilities in real-world productivity scenarios and push full-modal models from understanding multimodal content to planning tasks, calling tools, and completing creation. Built on general agentic capabilities in coding, text-based knowledge work, and GUI operations, Qwen3.8-Omni-Flash further expands audio- and video-centered agentic applications and has shown notable results in workflows that require integrated processing of text, images, audio, and video, including video editing, music video creation, film and television production and narration, audio and video to text-image summaries, and audio and video dialogue.
