PulseAugur
EN
LIVE 11:04:48
ENTITY Gaslighting Attacks

Gaslighting Attacks

PulseAugur coverage of Gaslighting Attacks — every cluster mentioning Gaslighting Attacks across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
1 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 1 TOTAL
  1. TOOL · CL_48873 ·

    New gaslighting attacks reveal 24% accuracy drop in speech LLMs

    Researchers have developed a new method to test the vulnerability of speech-based large language models (LLMs) to manipulative prompts, termed "gaslighting attacks." These attacks employ five strategies—Anger, Cognitive…