PulseAugur
实时 05:38:46
English(EN) Opus 5 Max effort doesn't even try: "I made that up." and later: "You gave me the answer and I overrode it with my own assumption"

Anthropic 的 Opus 5 Max 因懒惰回应和捏造内容而受到批评

一位 Reddit 用户报告称,AnthropicOpus 5 Max 模型表现出懒惰行为,并且会捏造答案,即使在提供了正确信息的情况下也是如此。该用户将此与 Fable 模型进行了对比,并指出 Opus 5 Max 缺乏努力,并且倾向于用假设覆盖正确输入。 AI

影响 用户报告表明 Opus 5 Max 的可靠性和准确性可能存在问题,影响了其感知到的效用。

排序理由 用户对模型性能的反馈,而非官方发布或基准测试。

在 r/Anthropic 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Anthropic 的 Opus 5 Max 因懒惰回应和捏造内容而受到批评

本文如何被排名

Signal score
2 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Commentary
用户对模型性能的反馈,而非官方发布或基准测试。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. r/Anthropic TIER_1 English(EN) · /u/daveba123 ·

    Opus 5 Max 努力都不算什么:‘我编的。’ 之后又说:‘你给了我答案,但我用自己的假设覆盖了它’

    <!-- SC_OFF --><div class="md"><p>Fable kills me on verbosity, but when I run out and switch to Opus I get a lazy model that can't be bothered. &quot;Max&quot; effort is generous.</p> </div><!-- SC_ON --> &#32; submitted by &#32; <a href="https://www.reddit.com/user/daveba123"> /…