PulseAugur
实时 15:36:36
English(EN) Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

研究发现:LLM产生的调查代理具有偏见且风格化的“品味”

一项新的研究论文探讨了使用大型语言模型(LLMs)作为人类调查受访者的代理,发现其生成的品味是人类偏好的风格化和有偏见的仿制品。该研究利用OpenAI、Anthropic和DeepSeek的模型,为艺术公众参与调查(SPPA)创建了近277,500个“硅基代理”。主要发现表明,这些LLM生成的响应表现出积极的喜爱偏见,失去了人类品味中存在的复杂关系结构,并扭曲了文化、年龄、阶级、性别和种族之间已知的关联。 AI

影响 强调了在使用LLM进行调查数据收集时可能存在的偏见和不准确性,并警告不要过度依赖合成受访者。

排序理由 一篇在arXiv上发表的研究论文,详细介绍了关于LLM行为的发现。

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

研究发现:LLM产生的调查代理具有偏见且风格化的“品味”

报道来源 [2]

  1. arXiv cs.CL TIER_1 English(EN) · Xiangyu Ma, Mengmi Zhang, Shannon Ang, Minne Chen ·

    非人类口味:LLM调查替代品的风格化杂食性

    arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produce reasonable approximations of cultural taste remains an open empirical question…

  2. arXiv cs.CL TIER_1 English(EN) · Minne Chen ·

    非人类般的品味:LLM调查代理的风格化杂食性

    Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produce reasonable approximations of cultural taste remains an open empirical question that becomes more urgent by the day, with marke…