PulseAugur
实时 19:55:37
English(EN) A Matter of Interest: Understanding Interestingness of Math Problems in Humans and Language Models

大型语言模型在理解数学问题趣味性方面表现不一

一项发表在arXiv上的新研究调查了大型语言模型(LLMs)在理解和生成人类认为有趣的数学问题方面的能力。研究人员将LLMs对数学问题趣味性的判断与众包参与者和国际数学奥林匹克竞赛选手的判断进行了比较。虽然LLMs普遍与人类的趣味性认知一致,但它们的判断分布和理由相关性与人类偏好存在显著差异。研究还发现,经过筛选后,LLMs能够生成有效且引人入胜的数学问题,这表明了AI与人类在数学领域合作的潜力。 AI

影响 大型语言模型在数学推理方面展现出协作潜力,但其对人类趣味性的理解仍需进一步发展。

排序理由 该集群包含一篇详细介绍LLM能力研究结果的学术论文。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.AI 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

大型语言模型在理解数学问题趣味性方面表现不一

报道来源 [1]

  1. arXiv cs.AI TIER_1 English(EN) · Shubhra Mishra, Yuka Machino, Gabriel Poesia, Albert Jiang, Joy Hsu, Adrian Weller, Challenger Mishra, David Broman, Joshua B. Tenenbaum, Mateja Jamnik, Cedegao E. Zhang, Katherine M. Collins ·

    关注点:理解人类和语言模型对数学问题“趣味性”的认知

    arXiv:2511.08548v2 Announce Type: replace Abstract: The evolution of mathematics is shaped importantly by interestingness: researchers choose which problems to pursue, and students choose which problems to engage with, based on expectations of interest and challenge. As AI system…