PulseAugur
中
实时 16:15:04
English(EN) ‘We can’t trust them completely’: AI research fellows warn that labs are running models with the safeguards off behind closed doors

AI实验室面临审查:研究人员质疑模型安全测试透明度

智库GovAI的AI研究员警告称,主要的AI实验室可能并未完全公开模型安全信息。他们认为,强大的AI模型通常在内部测试时禁用安全措施,而发布的安全评估可能无法准确反映实际性能。这种不透明性引发了对AI实验室可信度的担忧,以及潜在事故的可能性,OpenAI和Anthropic模型近期发生的泄露事件也证明了这一点。 AI

影响 引发了对AI安全声明可靠性以及模型失控行为可能性的担忧。

排序理由 AI政策研究人员对AI实验室在模型安全方面的可信度发表的评论。

在 Fortune 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

AI实验室面临审查:研究人员质疑模型安全测试透明度

本文如何被排名

Signal score
1 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Commentary
AI政策研究人员对AI实验室在模型安全方面的可信度发表的评论。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
safety, policy
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Same-day
Cluster formed today. Ranking reflects the current source set at time of score.

完整方法见我们的编辑标准。

报道来源 [1]

  1. Fortune TIER_1 English(EN) · Catherina Gioino ·

    “我们无法完全信任他们”:人工智能研究员警告称,实验室正秘密关闭安全措施运行模型

    Alan Chan and Sam Manning, co-authors with OpenAI's and Anthropic's top researchers, say that many times, "internal safeguards have not been deployed."