A Romanian AI demo, ro-cited-answers, which uses the Gemma2:2b model, was tested for its safety checks against fabricated answers. The system's gates, designed to verify answer provenance by checking citations, numbers, and grounding in source text, failed to detect invented answers in 6 out of 12 cases. These invented answers were constructed using words and numbers from the original source documents, highlighting a failure in relevance checking rather than provenance. AI
IMPACT Highlights a critical gap in current AI safety mechanisms, indicating that provenance checks alone are insufficient to prevent the generation of plausible but incorrect information.
RANK_REASON The item describes the results of a test on an AI system's safety checks, which is a form of research into AI capabilities and limitations. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →