PulseAugur
实时 07:08:16
English(EN) CaM-Wolf: Causal-Aware Multimodal Agents for Social Deduction Games

AI智能体应对社交推理游戏中的欺骗与推理 · 追踪2个来源

研究人员开发了能够玩复杂社交推理游戏的新型AI智能体,这些游戏需要欺骗和推理等细微技能。其中一个名为CaM-Wolf的智能体集成了多模态感知,处理视频输入并使用具因果意识的推理器来理解隐藏的角色,从而增强了人机交互。另一项研究引入了基于“Secret Hitler”的ParliamentBench框架,用于评估LLM在欺骗和推理方面的能力,发现顶级模型表现良好,但在维持一致的欺骗性角色方面存在困难。 AI

影响 社交推理游戏AI智能体的进步可能带来更复杂的人机交互和更完善的AI安全评估。

排序理由 两篇arXiv论文介绍了用于社交推理游戏的新型AI智能体和基准测试。

在 arXiv cs.AI 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

AI智能体应对社交推理游戏中的欺骗与推理 · 追踪2个来源

报道来源 [2]

  1. arXiv cs.AI TIER_1 English(EN) · Zheng Zhang, Nanjie Yao, Jiarui He, Deheng Ye, Peilin Zhao, Hao Wang ·

    CaM-Wolf:用于社交推理游戏的因果感知多模态智能体

    arXiv:2607.26393v1 Announce Type: new Abstract: Social deduction games (SDGs) such as Werewolf have become challenging testbeds for AI agents. These games require complex social skills such as reasoning, deception, and collaboration. While recent advances in large language models…

  2. arXiv cs.CL TIER_1 English(EN) · Niklas Bauer, Lars Benedikt Kaesberg, Akiko Aizawa, Jan Philip Wahle, Bela Gipp, Terry Ruas ·

    智能体能欺骗吗?在ParliamentBench中评估推理与欺骗,使用社交推理游戏

    arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabilities is fundamental to safety. Controlled social deduction games provide a repr…