PulseAugur
实时 05:19:18
(CA) Small edits, large models: How Wikipedia advocacy shapes LLM values

研究发现:维基百科编辑可显著塑造大型语言模型(LLM)的价值观

一项最新研究表明,在维基百科上进行的协调编辑可以显著影响大型语言模型(LLM)的价值观和输出。研究人员发现,一个名为“亲动物维基百科人”(PAW)的组织,他们为维基百科添加了动物福利相关内容,其编辑内容在关于动物福利主题的大型语言模型(LLM)的回答中被不成比例地归因。这种影响仅限于特定主题,并未扩散到关于同一实体的通用查询。研究结果表明,战略性地编辑维基百科是倡导组织塑造人工智能系统行为的一种低成本、有效的方法。 AI

影响 展示了一种低成本的方法,供倡导组织影响人工智能(AI)的输出,可能影响大型语言模型(LLM)如何讨论敏感话题。

排序理由 该集群报道了在arXiv上发表的一项科学研究,详细介绍了创新的研究方法和发现。

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

研究发现:维基百科编辑可显著塑造大型语言模型(LLM)的价值观

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群报道了在arXiv上发表的一项科学研究,详细介绍了创新的研究方法和发现。
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
policy, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
84 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [2]

  1. arXiv cs.CL TIER_1 (CA) · Jasmine Brazilek, Maria Navas, Alexa Gnauck ·

    细微改动,宏大模型:维基百科倡导如何塑造大语言模型价值观

    arXiv:2606.24890v1 Announce Type: new Abstract: Can a small group of volunteers shape how AI systems discuss animal welfare, just by editing Wikipedia? We show that they can. Wikipedia appears in nearly every major language model training dataset and is weighted more heavily than…

  2. LessWrong (AI tag) TIER_1 English(EN) · jonahmattwoodward ·

    倡导者可以通过编辑维基百科来影响大型语言模型的价值观

    <p><i><span>This article is a summary of an original study: Brazilek, J., Navas, M., &amp; Gnauck, A. (2026). Small edits, large models: How Wikipedia advocacy shapes LLM values. Zenodo.</span></i><a href="https://doi.org/10.5281/zenodo.19981454"><i><span> https://doi.org/10.5281…