PulseAugur
中
实时 11:34:41
English(EN) Detoxify: A framework for abusive text transformation using LLMs

大型语言模型框架Detoxify在保留意图的同时转化辱骂性文本

研究人员开发了一个名为Detoxify的框架,该框架利用大型语言模型(LLMs)将辱骂性文本转化为非辱骂性版本,同时保留原始意图。该研究评估了Gemini、GPT-4o、DeepSeek和Groq四种大型语言模型在识别和转化推文和评论中的仇恨言论和脏话方面的性能。结果表明,Groq产生的输出与其他模型相比有显著差异,经常改变上下文,而GPT-4o和DeepSeek在转化方面表现出相似性。 AI

影响 这项研究可能有助于改进内容审核工具和更安全的在线交流平台。

排序理由 该集群基于一篇在arXiv上发表的学术论文,详细介绍了一个新框架及其评估。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.AI 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

大型语言模型框架Detoxify在保留意图的同时转化辱骂性文本

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群基于一篇在arXiv上发表的学术论文,详细介绍了一个新框架及其评估。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, safety
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
84 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. arXiv cs.AI TIER_1 English(EN) · Rohitash Chandra, Jiyong Choi, Jayesh Sonawane ·

    Detoxify:使用LLM进行辱骂性文本转换的框架

    arXiv:2507.10177v2 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) have demonstrated significant advancements in natural language processing tasks, their effectiveness in the classification and transformation of abusive text into non-abusive versions …