PulseAugur
实时 07:07:07
English(EN) Effects of Theory of Mind and Prosocial Beliefs on Steering Human-Aligned Behaviors of LLMs in Ultimatum Games

Theory of Mind 增强了 LLM 在最后通牒博弈中的对齐能力

一篇新的研究论文探讨了心智理论(ToM)和亲社会信念如何影响大型语言模型(LLMs)在最后通牒博弈中的行为。该研究使用了 2,700 次模拟,LLM 代理被初始化为具有“贪婪”、“公平”或“无私”的信念,并具有不同程度的 ToM 推理能力。结果表明,ToM 显著增强了与人类规范的对齐、决策一致性和谈判结果,特别是对于推理模型。研究还发现,不同的博弈角色从不同顺序的 ToM 中受益,并且 Llama 3.3 70B 表现出的推理与其行为和信念最一致。 AI

影响 这项研究表明,将心智理论纳入 LLM 可能导致在复杂社交互动中产生更可预测且与人类对齐的 AI 行为。

排序理由 该集群包含一篇发表在 arXiv 上的研究论文,详细介绍了 LLM 行为的实验结果。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.AI 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Theory of Mind 增强了 LLM 在最后通牒博弈中的对齐能力

本文如何被排名

Signal score
2 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群包含一篇发表在 arXiv 上的研究论文,详细介绍了 LLM 行为的实验结果。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, safety
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
1 days old
Coverage has settled into its steady-state source set.

完整方法见我们的编辑标准

报道来源 [1]

  1. arXiv cs.AI TIER_1 English(EN) · Neemesh Yadav, Yihuai Lan, Shan Dong, Mai Hieu Hien, Palakorn Achananuparp, Jing Jiang, Ee-Peng Lim ·

    心智理论和亲社会信念对引导大型语言模型在最后通牒博弈中实现人类对齐行为的影响

    arXiv:2505.24255v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown potential in simulating human behaviors and performing theory-of-mind (ToM) reasoning, crucial for complex social interactions. We investigate ToM reasoning's role in aligning agenti…