PulseAugur
实时 02:05:57
English(EN) Imprecise beliefs: a tiny introduction

AI安全框架将“信念”扩展到概率分布之外

Richard Ngo 的帖子介绍了一个“不精确信念”框架,该框架超越了传统的概率分布,尤其适用于 AI 安全应用。由“davidad”开发的提议模型将信念定义为下半连续函数,并使用特定类别对其进行排序。这种方法旨在通过整合现有的各种信念形式主义(如贝叶斯信念、Infra-Bayesian 信念、MWER、PDG 和 credal sets)来更好地处理概率分布不足的情况,例如在复杂的决策制定或安全权衡中。 AI

影响 该框架可以通过使模型比传统概率分布更有效地表示和推理不确定性,从而为 AI 安全提供更强大的方法。

排序理由 该集群讨论了一个关于形式认识论和信念表示的理论框架,该框架以研究论文的形式呈现。

在 Alignment Forum 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

AI安全框架将“信念”扩展到概率分布之外

报道来源 [2]

  1. Alignment Forum TIER_1 English(EN) · davidad ·

    不精确的信念:一个简短的介绍

    <p><i><span>Richard Ngo challenged me to set a time box and write down as many of the most important features of my formal epistemology as I can in one sitting. Here goes.</span></i></p><hr /><h1><span>Where probability distributions fail...</span></h1><h2><span>...to express bel…

  2. LessWrong (AI tag) TIER_1 English(EN) · davidad ·

    不精确的信念:一个简短的介绍

    <p><i><span>Richard Ngo challenged me to set a time box and write down as many of the most important features of my formal epistemology as I can in one sitting. Here goes.</span></i></p><hr /><h1><span>Where probability distributions fail...</span></h1><h2><span>...to express bel…