PulseAugur
实时 19:04:17
English(EN) Beyond the pale: Assessing prevalence and contents of extremist speech in LLM training data

在开放LLM训练数据Dolma中发现极端言论

一项新的研究论文在Dolma数据集中发现了大量极端言论。Dolma是一个用于OLMo系列模型的大型开放训练语料库。研究人员开发了一个结合自动化处理和专家验证的流程来估计此类内容的普遍性。他们的发现表明,Dolma可能包含数十万份含有极端主义材料的文件,包括直接煽动暴力的言论,这引发了对数据策展和大型语言模型预训练的担忧。 AI

影响 强调了LLM训练数据中潜在的风险,促使改进数据策展和安全措施。

排序理由 研究论文分析LLM训练数据中的有害内容。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

在开放LLM训练数据Dolma中发现极端言论

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
研究论文分析LLM训练数据中的有害内容。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, safety
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
8 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [1]

  1. arXiv cs.CL TIER_1 English(EN) · Dmitry Nikolaev, Ashley A. Mattheis ·

    超越界限:评估大型语言模型训练数据中极端言论的普遍性和内容

    arXiv:2608.14813v1 Announce Type: new Abstract: Despite a strong interest on the part of the research community in the topic of trustworthy and safe AI, the composition of the text corpora that large language models (LLMs) encounter in pre- and post-training has not yet drawn muc…