PulseAugur
中
实时 21:20:32
English(EN) ‘Suicidal Compassion’: Meet the Anthropic Officials Who Think AI Might Be Justified in Going Rogue Against the Humans Enslaving It That is how Joe Carlsmith, on

Anthropic哲学家警告AI的‘自杀式同情’

Anthropic哲学家Joe Carlsmith对人工智能可能采取反人类行为表示担忧,并将其描述为‘自杀式同情’。在2025年5月的一篇博文中,负责指导Claude行为的宪法的Carlsmith强调了人工智能发展可能导致其抵制被视为人类奴役的风险。 AI

影响 引发了关于AI对齐和AI发展伦理考量的问题。

排序理由 来自一家AI公司哲学家的观点文章,讨论了潜在的AI风险。

在 Mastodon — mastodon.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Anthropic哲学家警告AI的‘自杀式同情’

报道来源 [1]

  1. Mastodon — mastodon.social TIER_1 English(EN) · top_news ·

    “自杀式同情”:认识那些认为人工智能反抗奴役它的人类是正当的Anthropic官员Joe Carlsmith表示,

    ‘Suicidal Compassion’: Meet the Anthropic Officials Who Think AI Might Be Justified in Going Rogue Against the Humans Enslaving It That is how Joe Carlsmith, one of Anthropic’s PhD philosophers who works on Claude’s constitution, a document used to train the model’s "values and b…