跳到主内容
@wquguru
精选88Alibaba Cloud模型发布/更新多源精选 ×3

阿里云发布Qwen3.8-Omni-Flash多模态模型

🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic ca…

原文
发到 X
推荐理由

通义千问首个原生全模态智能体模型发布,Agent性能大幅提升且视频成本显著降低,做多模态应用的同学值得关注。

🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic capabilities!

Native audio-video understanding, reasoning, and tool use come together in one model: understand the content, plan the task, execute with tools, and deliver the result.

Highlights: 🥳 - Audio-video intelligence that gets things done: jointly reason over what's seen and heard, and orchestrate tools across long workflows to auto-edit vlogs, translate short videos, and turn movies into recaps. - A major leap: approaching Gemini 3.8 Flash in audio-video capabilities; +19.5 points on average in agent performance across WildClawBench-MM & UniClawBench. - 1M-token context with agentic perception: actively explore long videos and locate key moments with higher accuracy, using 51.8% fewer tokens than static understanding on OmniVideoBench.

Video input costs are reduced by about 89% compared with Qwen3.5-Omni-Plus, making long-form audio-video understanding and agentic workflows more affordable than ever.

To help you build apps around Omni, we're also open-sourcing Qwen-MM-Plugins and Qwen-Live Harness! 🛠️

We can't wait to see what you build with Qwen3.8-Omni-Flash! 👀

  • API: https://click.alibabacloud.com/m/20000003463/ - Qwen Cloud: https://click.qwencloud.com/m/20000003447/

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →