PulseAugur
EN
LIVE 23:48:44
ENTITY theory of mind

theory of mind

PulseAugur coverage of theory of mind — every cluster mentioning theory of mind across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
10
26 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
10
21 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

5 day(s) with sentiment data

RECENT · PAGE 1/2 · 34 TOTAL
  1. TOOL · CL_231570 ·

    New benchmark CoMMET evaluates multimodal LLMs' Theory of Mind

    Researchers have introduced CoMMET, a new benchmark designed to evaluate the Theory of Mind (ToM) capabilities of multimodal large language models (MLLMs). This benchmark is inspired by the psychology-based Theory of Mi…

  2. TOOL · CL_238357 ·

    New theory proposes LLM agents with "Theory of Mind" for 6G networks

    A new paper proposes a theoretical framework for managing future 6G networks using Large Language Model (LLM) agents. The research suggests that inter-agent messages should be treated as traces of reasoning rather than …

  3. RESEARCH · CL_233355 ·

    New research proposes Theory of Mind for LLM agents in 6G networks

    A new research paper proposes a framework for managing future 6G networks using Large Language Model (LLM) agents. The paper introduces five principles for resilient multi-agent systems, emphasizing that messages should…

  4. TOOL · CL_229132 ·

    New AI alignment method uses Theory of Mind to reduce misunderstandings

    Researchers have developed a new method called Frictive Policy Optimization (FPO) that uses Theory of Mind (ToM) to improve dialogue alignment in AI models. This approach distinguishes between surface coordination and g…

  5. TOOL · CL_229028 ·

    GPT-4o shows human-like Theory of Mind in new LLM study

    A new study published on arXiv investigates whether large language models (LLMs) possess a Theory of Mind (ToM), the ability to understand others' beliefs, intentions, and emotions. Researchers compared the performance …

  6. TOOL · CL_221076 ·

    AI evaluation flaw: Incomplete references reverse model rankings

    A new paper published on arXiv, titled "Unmatched Does Not Mean False: Incomplete Reference Sets Can Reverse Calibration Rankings in Open-Ended Theory-of-Mind Tracking," highlights a critical flaw in evaluating open-end…

  7. TOOL · CL_218153 ·

    New framework DyCAC enhances LLM social understanding in multicultural settings

    Researchers have introduced DyCAC, a novel framework designed to enhance the social understanding capabilities of Large Language Models (LLMs). Unlike existing methods that treat culture as a static attribute, DyCAC dyn…

  8. TOOL · CL_218088 ·

    Theory of Mind enhances LLM alignment in ultimatum games

    A new research paper explores how Theory of Mind (ToM) and prosocial beliefs influence the behavior of Large Language Models (LLMs) in ultimatum games. The study involved 2,700 simulations using LLM agents initialized w…

  9. TOOL · CL_216043 ·

    New ARGUS framework uses Theory-of-Mind for persuasive argument generation

    Researchers have developed ARGUS, a novel agent-based framework designed to enhance persuasive argument generation. This system utilizes a Theory-of-Mind (ToM) Reasoner to model audience beliefs and values, which then g…

  10. TOOL · CL_215896 ·

    New benchmark reveals vision-language models fail to translate theory of mind into action

    Researchers have developed MOSAIC, a new benchmark designed to measure how well vision-language models (VLMs) can translate theory of mind (ToM) reasoning into coordinated social actions. In evaluations of 13 models, in…

  11. MEME · CL_213182 ·

    Stable Diffusion animation of Tom and Jerry shared on Reddit

    A Reddit user shared an animation created using Stable Diffusion's extended features, which aims to maintain a consistent animation style. The user noted that while fast action can lead to smears, they found the resulti…

  12. TOOL · CL_202765 ·

    Strong-to-weak AI scaffolding boosts model performance without retraining

    Researchers have developed a novel method called strong-to-weak scaffolding that significantly enhances the performance of smaller AI models without requiring any retraining. This technique involves a more powerful AI m…

  13. TOOL · CL_197997 ·

    AI teams use second-order Theory of Mind for better alignment

    Researchers have proposed a new framework for human-autonomy teams that leverages second-order Theory of Mind (ToM-2) to improve alignment and learning. This approach allows a human teacher, who possesses knowledge of t…

  14. RESEARCH · CL_174089 ·

    AI safety fine-tuning alters LLM beliefs on consciousness and values

    A new research paper explores how safety fine-tuning in large language models (LLMs) can inadvertently affect their representations of consciousness and human values. The study found that efforts to prevent LLMs from at…

  15. TOOL · CL_173869 ·

    New model explains human trade-offs between social and non-social learning

    Researchers have developed a "Rational Mentalizing model" to understand how agents, including humans, decide between learning from others (social learning) and direct experience (non-social learning). This model quantif…

  16. MEME · CL_162952 ·

    Irrelevant content detected in AI news cluster

    This cluster contains a single item that appears to be unrelated to AI news. The content is a Mastodon post with hashtags related to humor and discipline, and the text itself does not discuss AI developments or entities.

  17. MEME · CL_157258 ·

    Humorous take on staged robot 'death' scene goes viral on Mastodon

    A Mastodon user shared a humorous observation about a humanoid robot's staged 'death' and subsequent removal from a stage. The user, Tom (@[email protected]), found the robot's unnatural pose, the black blankets use…

  18. TOOL · CL_156454 ·

    New benchmark MeetingToM tests LLMs on social reasoning in meetings

    Researchers have introduced MeetingToM, a new benchmark designed to evaluate the Theory of Mind (ToM) capabilities of Multimodal Large Language Models (MLLMs) in the context of multi-party meetings. This benchmark addre…

  19. TOOL · CL_162791 ·

    New benchmark MeetingToM tests multimodal LLMs on social reasoning

    Researchers have introduced MeetingToM, a new benchmark designed to evaluate the theory of mind capabilities of multimodal large language models (MLLMs) in the context of multi-party meetings. This benchmark addresses l…

  20. RESEARCH · CL_141110 ·

    New EAST benchmark reveals LLM theory of mind gaps

    Researchers have developed a new evaluation method called the Epistemic Asymmetry Schelling Task (EAST) to assess Theory of Mind (ToM) in large language models (LLMs). Unlike traditional tests like the Sally-Anne task, …