PulseAugur
实时 23:50:10
English(EN) What do you think no-guardrails Mythos 2 (or whatever Anthropic's most advanced internal model is by now) would be getting on these benchmarks if, for the sake of the argument, they released whatever it is at full strength, right now?

用户推测 Anthropic 的内部 Mythos 2 模型可能远超公开的 AI

Reddit 上的一场讨论推测了 Anthropic 内部未发布的模型(特别是 Mythos 2)的潜在性能。用户假设这类模型在基准测试上的得分可能远高于 Fable 5.1 等当前公开的模型,表明 Anthropic 在内部可能领先数月。对话还触及了先进的“智能护栏”可能在未来安全的前提下发布更强大的公开模型的可能性。 AI

影响 推测表明,在先进的安全护栏到位后,未来可能出现功能更强大的公开 AI 模型。

排序理由 用户对内部模型能力的推测,并非官方发布或基准测试。

在 r/Anthropic 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

用户推测 Anthropic 的内部 Mythos 2 模型可能远超公开的 AI

本文如何被排名

Signal score
2 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Commentary
用户对内部模型能力的推测,并非官方发布或基准测试。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, opinion
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. r/Anthropic TIER_1 English(EN) · /u/DeepOrangeSky ·

    没有护栏的Mythos 2(或者到目前为止Anthropic最先进的内部模型是什么)如果在此时此刻,为了论证的目的,以全部实力发布,你认为它会在这些基准测试中取得什么成绩?

    <!-- SC_OFF --><div class="md"><p>Like, 20% higher on most of the major benchmarks? (other than the saturated ones that are already in the 80-90% range I mean)</p> <p>It's gotta be waaaay beyond what this Fable 5.1 model is scoring, by this point.</p> <p>Seems like they are proba…