A new research paper explores retrieval-based in-context learning (RetICL) strategies for detecting defamatory content online, specifically focusing on German criminal law. While few-shot prompting shows improvement over zero-shot, retrieval-based methods offer only marginal gains and can even underperform a static set of demonstrations. The study found that model choice is more critical than other system choices, and current models tend to over-predict criminal relevance while still missing a significant portion of actual defamatory posts, making them suitable for triage rather than autonomous moderation. AI
IMPACT Current AI models show limitations in accurately identifying defamatory content, suggesting a need for further development before autonomous moderation can be reliably implemented.
RANK_REASON The item is an academic paper detailing research findings on AI model performance for a specific task. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →