adversarial example
PulseAugur coverage of adversarial example — every cluster mentioning adversarial example across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New defense Random Logit Scaling protects AI models from adversarial attacks
Researchers have introduced Random Logit Scaling (RLS), a new defense mechanism designed to protect deep neural networks against black-box score-based adversarial example attacks. RLS functions as a post-processing step…
-
New framework unifies detection of AI content, hallucinations, and watermarks
Researchers have developed a novel unified framework for detecting AI-generated content and artifacts, including LLM text, hallucinations, watermarks, and adversarial examples. The method utilizes Mahalanobis distance s…
-
New research targets AI robustness with novel distillation and testing methods · 8 sources tracked
Researchers are exploring new methods to enhance the adversarial robustness of neural networks. One approach, AD-CERT, combines adversarial distillation with Interval Bound Propagation to achieve state-of-the-art certif…