PulseAugur
EN
LIVE 19:05:53

DeepSeek V4 Pro trails rivals in benchmarks, users question performance

The open-source AI model DeepSeek V4 Pro is reportedly underperforming compared to other frontier models like GLM 5.3 Flash and Qwen Next in recent benchmarks. Users speculate that the model may not have been fully trained or that its performance is being impacted by technical factors such as KV cache compaction and efficient hybrid attention. There is also mention of talent departures from DeepSeek to competitors like Xiaomi, raising questions about the lab's future competitiveness. AI

IMPACT Raises questions about the competitive landscape of open-source frontier models and the factors influencing their performance.

RANK_REASON User discussion and speculation about an AI model's performance, rather than a direct announcement or benchmark release from the model's creators.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

DeepSeek V4 Pro trails rivals in benchmarks, users question performance

How we ranked this

Signal score
5 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Commentary
User discussion and speculation about an AI model's performance, rather than a direct announcement or benchmark release from the model's creators.
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

Full methodology in our editorial standards.

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/power97992 ·

    How is deepseek behind? Ds v4 pro 0813 is worse than glm 5.3 flash , qwen next and other frontier open models in benchmarks?

    <!-- SC_OFF --><div class="md"><p>When will they catdh up? They were one of the top labs when ds v3.2 and v3 came out , but now ds v4 pro is worse than qwen 3,8 next in benchmarks. It seems like ds v4 pro is not trained to its full potential , but v4 flash is pretty good. Maybe t…