PulseAugur
实时 11:01:24
English(EN) Beyond the pale: Assessing prevalence and contents of extremist speech in LLM training data

在开放LLM训练数据Dolma中发现极端言论

一项新的研究论文在Dolma数据集中发现了大量极端言论。Dolma是一个用于OLMo系列模型的大型开放训练语料库。研究人员开发了一个结合自动化处理和专家验证的流程来估计此类内容的普遍性。他们的发现表明,Dolma可能包含数十万份含有极端主义材料的文件,包括直接煽动暴力的言论,这引发了对数据策展和大型语言模型预训练的担忧。 AI

影响 强调了LLM训练数据中潜在的风险,促使改进数据策展和安全措施。

排序理由 研究论文分析LLM训练数据中的有害内容。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

在开放LLM训练数据Dolma中发现极端言论

报道来源 [1]

  1. arXiv cs.CL TIER_1 English(EN) · Dmitry Nikolaev, Ashley A. Mattheis ·

    超越界限:评估大型语言模型训练数据中极端言论的普遍性和内容

    arXiv:2608.14813v1 Announce Type: new Abstract: Despite a strong interest on the part of the research community in the topic of trustworthy and safe AI, the composition of the text corpora that large language models (LLMs) encounter in pre- and post-training has not yet drawn muc…