PulseAugur
实时 07:02:36
实体 Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM

PulseAugur coverage of Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM — every cluster mentioning Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM across labs, papers, and developer communities, ranked by signal.

Show in brief
总计 · 30天
1
90 天内 1
发布 · 30天
0
90 天内 0
论文 · 30天
1
90 天内 1
层级分布 · 90 天
主题
情绪 · 30 天

1 天有情绪数据

最近 · 第 1/1 页 · 共 1 条
  1. TOOL · CL_169812 ·

    调查详细介绍了LLM在生成和缓解有害内容方面的双重作用

    一篇题为“Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM”的最新调查论文系统地回顾了大型语言模型(LLM)的双重性质。该论文探讨了LLM如何既能生成有毒或带有偏见语言等有害内容,又能通过检测和预防作为安全工具。它提出了LLM危害和防御的分类法,分析了对抗性攻击,并评估了如RLHF和提示工程等缓解策略。