PulseAugur
中
实时 00:57:01
English(EN) Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach

研究发现人工智能模型缺乏自主研究的判断力

一项涉及普林斯顿大学和英国人工智能安全研究所的最新研究发现,包括 Claude Opus 4.8 和 GPT-5.6 Sol 在内的当前前沿人工智能模型尚不具备自主进行人工智能研究的能力。虽然这些模型能够处理研究工程的技术方面,但它们缺乏独立生成可发表的人工智能研究论文所需的关键判断力、创造性解决问题的能力和适应性。该研究结果直接反驳了 Anthropic 和 OpenAI 关于自主人工智能研究即将成为可能的说法。 AI

影响 当前前沿人工智能模型尚不具备独立研究的能力,这凸显了在关键判断和创造性解决问题方面需要人类监督。

排序理由 该集群报道了一项评估人工智能研究能力的 പഠна,属于研究类别。

在 The Decoder 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

研究发现人工智能模型缺乏自主研究的判断力

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群报道了一项评估人工智能研究能力的 പഠна,属于研究类别。
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
paper, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
55 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [2]

  1. The Decoder TIER_1 English(EN) · Jonathan Kemper ·

    研究反驳Anthropic和OpenAI关于自主AI研究即将实现的说法

    <p><img alt="A collage of a computer screen showing red corrections and X marks next to a stack of scientific reports containing charts." class="attachment-full size-full wp-post-image" height="1047" src="https://the-decoder.com/wp-content/uploads/2026/08/ai-assisted-science-reje…

  2. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    研究与Anthropic和OpenAI相悖:自主AI研究仍遥遥无期 https://the-decoder.de/studie-widerspricht-anthropic-und-openai-autonom

    Studie widerspricht Anthropic und OpenAI: Autonome KI-Forschung ist noch weit entfernt https:// the-decoder.de/studie-widerspr icht-anthropic-und-openai-autonome-ki-forschung-ist-noch-weit-entfernt/ > KI-Agenten mit Claude Opus 4.8 und GPT-5.6 Sol erhielten sechs Tage, 3.000 Doll…