PulseAugur
EN
LIVE 09:03:47
ENTITY HateXplain

HateXplain

PulseAugur coverage of HateXplain — every cluster mentioning HateXplain across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
3 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 3 TOTAL
  1. TOOL · CL_171830 ·

    Span-guided vs. unguided AI detoxification: a nuanced comparison

    A new research paper explores the effectiveness of span-guided detoxification in AI language models, comparing it to unguided methods. The study found that while span-guided rewriting is preferred when it preserves mean…

  2. TOOL · CL_117726 ·

    New methods probe generative models for bias and improve performance

    Researchers have developed new methods, Attribution Graphs (AGs) and Causal Probing, to analyze the internal workings of generative models. These techniques aim to identify and correct issues like spurious correlations,…

  3. TOOL · CL_117577 ·

    Hate speech annotation pipeline flaw silences minority values

    A new research paper highlights a critical flaw in how hate speech datasets are annotated, specifically concerning the boundary between offensive and hateful content. The study reveals that annotator disagreement is not…