PulseAugur
中
实时 08:03:25
English(EN) Qwen3.5-Omni: Scaling Up, Toward Native Omni-Modal AGI

阿里巴巴的 Qwen3.5-Omni 为多模态大语言模型增加了音频和视频能力

阿里巴巴的 Qwen 团队发布了新一代全模态大语言模型 Qwen3.5-Omni,能够处理文本、图像、音频和视听内容。该系列模型包括 Plus、Flash 和 Light 版本,均支持 256k 上下文窗口,并能处理超过 10 小时的音频。其架构在推理和生成组件中均采用了混合注意力专家混合(MoE)方法。 AI

影响 将大语言模型的能力扩展到原生的音频和视频处理,可能催生更复杂的 AI 代理和应用。

排序理由 前沿实验室模型发布,附带系统卡。[lever_c_demoted from frontier_release: ic=1 ai=1.0]

在 Qwen tech blog 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

阿里巴巴的 Qwen3.5-Omni 为多模态大语言模型增加了音频和视频能力

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
前沿实验室模型发布,附带系统卡。[lever_c_demoted from frontier_release: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
183 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. Qwen tech blog TIER_1 English(EN) · QwenTeam ·

    Qwen3.5-Omni:规模化,迈向原生全模态AGI

    Qwen3.5-Omni is Qwen’s latest generation of fully omnimodal LLM, supporting the understanding of text, images, audio, and audio-visual content. Both the Thinker and Talker in Qwen3.5-Omni adopt the Hybrid-Attention MoE. Qwen3.5-Omni series includes Instruct versions in three size…