Researchers have developed RAGSieve, a novel framework designed to detect knowledge poisoning in retrieval-augmented generation (RAG) systems. Unlike existing methods that rely on trusted corpora or specific attack artifacts, RAGSieve uses self-referenced contrastive learning. Its components, RAGSieve-Query (RSQ) and RAGSieve-Graph (RSG), analyze query-local and corpus-local data respectively to identify injected documents that promote false claims. Experiments show RAGSieve significantly outperforms previous methods in detecting poisoned data while minimizing the removal of legitimate documents, offering practical protection at both ingestion and query stages. AI
IMPACT Enhances the security and reliability of AI systems by providing a robust method for detecting malicious data injection.
RANK_REASON Academic paper detailing a new method for detecting knowledge poisoning in AI. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- CatalyzeX
- CleanBase
- DagsHub
- gMTP
- Gotit.pub
- Hugging Face
- RAGSieve
- RAGSieve-Graph
- RAGSieve-Query
- Retrieval-Augmented Generation
- ScienceCast
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →