Researchers have developed SynGhost, a novel backdoor attack that exploits vulnerabilities in pre-trained language models (PLMs). This attack injects multiple syntactic backdoors into the pre-training data, allowing them to transfer to various downstream tasks without explicit target setting or easily identifiable triggers. SynGhost preserves the PLM's original capabilities while posing significant threats and evading common defenses like perplexity filters and maxEntropy. AI
IMPACT This research highlights new vulnerabilities in pre-trained language models, potentially impacting the security and reliability of AI systems.
RANK_REASON The cluster contains an academic paper detailing a new method for attacking AI models. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →