PulseAugur
实时 04:20:46
English(EN) Rohin Shah on AGI Safety

Google DeepMind 的 AGI 安全主管对灾难性失调表示怀疑

Google DeepMind 的 AGI 安全与对齐主管 Rohin Shah 认为,灾难性 AI 失调是可能发生的,但并非默认就会发生。他认为,当前专注于短期奖励的 AI 训练方法,并不会自然地导向实现“接管世界”所需的长远目标。Shah 建议,许多潜在的对齐问题将提前显现,从而可以进行迭代式解决,并且研究重点应从部署前评估转向实际研究和 AI 治理基础设施。 AI

影响 讨论了高级 AI 的潜在风险和缓解策略,影响了安全研究和开发的方向。

排序理由 该集群是一篇关于 AGI 安全的讨论和观点文章,采访了一位知名研究员,而不是直接发布或公告。

在 LessWrong (AI tag) 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Google DeepMind 的 AGI 安全主管对灾难性失调表示怀疑

报道来源 [2]

  1. LessWrong (AI tag) TIER_1 English(EN) · anaguma ·

    Rohin Shah 谈 AGI 安全

    <p><span>Rohin Shah recently had an </span><a href="https://80000hours.org/podcast/episodes/rohin-shah-google-deepmind-agi-safety/" rel="noreferrer"><span>interview</span></a><span> on 80000 hours on his views on AGI Safety and his work at Google DeepMind. I'm posting the transcr…

  2. 80,000 Hours TIER_1 English(EN) · Robert Wiblin ·

    Rohin Shah 谈论在 Google DeepMind 负责 AGI 安全的真实感受

    <p>The post <a href="https://80000hours.org/podcast/episodes/rohin-shah-google-deepmind-agi-safety/">Rohin Shah on what it&#8217;s really like to run AGI safety at Google&nbsp;DeepMind</a> appeared first on <a href="https://80000hours.org">80,000 Hours</a>.</p>