PulseAugur
实时 14:16:10
中文(ZH) DeepSeek V4 多模态开源,我们把它的视觉链路拆了一遍

DeepSeek V4 多模态模型权重已发布供检查

DeepSeek 发布了其 V4 多模态模型的权重和参考代码,允许研究人员检查其视觉处理能力。与简单的图像到文本的附加功能不同,V4 将视觉 token 直接集成到其现有架构中,使其能够参与注意力机制、专家混合(MoE)路由和代理推理。该模型采用 Vision Transformer (ViT) 进行初始编码,然后使用 Aligner 将视觉特征压缩并映射到语言模型的空间中。V4 内的专用机制处理视觉 token,包括修改后的注意力规则和不同的 MoE 路由偏差,以更好地保留长上下文框架内图像的空间关系和计算需求。 AI

影响 能够对多模态集成和代理能力进行更深入的研究。

排序理由 前沿实验室模型发布,包含权重和参考代码。[lever_c_从 frontier_release 降级:ic=1 ai=1.0]

在 雷峰网 (Leiphone) 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

DeepSeek V4 多模态模型权重已发布供检查

本文如何被排名

Signal score
28 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
前沿实验室模型发布,包含权重和参考代码。[lever_c_从 frontier_release 降级:ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. 雷峰网 (Leiphone) TIER_1 中文(ZH) ·

    DeepSeek V4 多模态开源,我们拆解了它的视觉链

    <section style="text-align: center; margin: 0px 16px; line-height: 1.75em; display: block;"><img class="rich_pages wxw-img" src="https://static.leiphone.com/uploads/new/images/20260901/6a96aaf3030b8.jpg?imageMogr2/quality/90" style="width: 100%; display: inline-block; text-align:…