PulseAugur
实时 19:10:23
English(EN) How is deepseek behind? Ds v4 pro 0813 is worse than glm 5.3 flash , qwen next and other frontier open models in benchmarks?

DeepSeek V4 Pro 在基准测试中落后于竞争对手,用户质疑其性能

据报道,开源 AI 模型 DeepSeek V4 Pro 在近期基准测试中表现不如 GLM 5.3 FlashQwen Next 等其他前沿模型。用户推测该模型可能未完全训练,或其性能受到 KV 缓存压缩和高效混合注意力等技术因素的影响。此外,还有消息称 DeepSeek 的人才流失到小米等竞争对手,引发了对其实验室未来竞争力的质疑。 AI

影响 引发了对开源前沿模型竞争格局以及影响其性能因素的讨论。

排序理由 用户对 AI 模型性能的讨论和猜测,而非模型创建者直接发布公告或基准测试结果。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

DeepSeek V4 Pro 在基准测试中落后于竞争对手,用户质疑其性能

本文如何被排名

Signal score
5 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Commentary
用户对 AI 模型性能的讨论和猜测,而非模型创建者直接发布公告或基准测试结果。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/power97992 ·

    deepseek 究竟落后多少?Ds v4 pro 0813 在基准测试中不如 glm 5.3 flash、qwen next 等前沿开源模型?

    <!-- SC_OFF --><div class="md"><p>When will they catdh up? They were one of the top labs when ds v3.2 and v3 came out , but now ds v4 pro is worse than qwen 3,8 next in benchmarks. It seems like ds v4 pro is not trained to its full potential , but v4 flash is pretty good. Maybe t…