2026-09-18 14:05
AI123
Qwen发布首款全模态大模型Qwen3.8-Omni-Flash
阿里通义千问团队(Qwen)于2026年9月18日发布其首款围绕智能体能力构建的全模态模型Qwen3.8-Omni-Flash。该模型原生支持音视频理解、推理与工具调用,能够理解内容、规划任务并调用工具完成执行,可用于自动剪辑Vlog、翻译短视频、生成电影解说等场景。
官方介绍显示,Qwen3.8-Omni-Flash在音视频能力上接近Gemini 3.8 Flash,在WildClawBench-MM和UniClawBench基准上的智能体性能平均提升19.5分。该模型拥有100万token的上下文,支持主动探索长视频并准确定位关键时刻,在OmniVideoBench上相比静态理解方式减少51.8%的token使用。
此外,Qwen3.8-Omni-Flash的视频输入成本相比Qwen3.5-Omni-Plus降低约89%,使长视频理解和智能体工作流更加经济。为帮助开发者围绕该模型构建应用,Qwen还同步开源了Qwen-MM-Plugins和Qwen-Live Harness。
来源
- 2026-09-18 11:16 | X:🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic capabilities! Native audio-video understanding, reasoning, and tool u...阅读原文
🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic capabilities! Native audio-video understanding, reasoning, and tool use come together in one model: understand the content, plan the task, execute with tools, and deliver the result. Highlights: 🥳 - Audio-video intellige...