ENTITY
RealToxicityPrompts
RealToxicityPrompts
PulseAugur coverage of RealToxicityPrompts — every cluster mentioning RealToxicityPrompts across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
1 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
1 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
New methods enhance LLM steering for behavior control · 2 sources tracked
Two new research papers introduce novel methods for steering large language models to suppress undesired behaviors. GAPS (Gated Activation steering via Posterior and Separability) employs dimension-level gates to select…
-
Open-source safety guard models evaluated; smaller Qwen Guard leads in recall
A new research paper evaluates 14 open-source safety guard models using a benchmark of over 79,000 samples across eight safety categories. The study found that model size does not correlate with safety detection perform…