Researchers have introduced the Medical Prompt Injection Benchmark (MPIB), a new dataset and evaluation suite designed to assess the clinical safety of Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) systems. MPIB focuses on identifying risks associated with prompt injection attacks, both direct and indirect, within clinical contexts. The benchmark utilizes the Clinical Harm Event Rate (CHER) to measure severe clinical harm and distinguishes between moderate and high-severity outcomes, revealing significant divergences between different LLMs and defense strategies. AI
IMPACT This benchmark will enable better evaluation of LLM safety in clinical settings, potentially leading to more secure AI integration in healthcare.
RANK_REASON The cluster describes a new academic benchmark and dataset for evaluating LLM safety, published on arXiv. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →