Hate Speech Detection
PulseAugur coverage of Hate Speech Detection — every cluster mentioning Hate Speech Detection across labs, papers, and developer communities, ranked by signal.
-
PEFT boosts LLM hate speech detection in Roman Urdu to over 93% F1
A new research paper explores the effectiveness of Parameter-Efficient Fine-Tuning (PEFT) methods, specifically Low-Rank Adaptation (LoRA), for hate speech detection in Roman Urdu. The study found that while zero-shot i…
-
New research explores privacy vs. hate speech detection trade-off
A new research paper introduces the concept of a "Privacy-HSD Trade-off," highlighting the potential for hate speech detection (HSD) systems to inadvertently compromise user privacy by encoding authorship. The study exp…
-
New framework rethinks hate speech model evaluation with human rationales
Researchers have developed a new framework to evaluate hate speech detection models, focusing on the variation in human explanations (rationales) beyond simple majority votes. The study proposes organizing classificatio…
-
New clustering method models annotator perspectives in NLP tasks
Researchers have developed a new agreement-based clustering technique to better model annotator perspectives in subjective Natural Language Processing tasks. This method aims to capture the nuances of disagreement among…