A new diagnostic framework called V-DEAL has been developed to identify safety vulnerabilities in Video Large Language Models (LLMs). Researchers found that harmful videos paired with benign queries are more susceptible to attacks than when paired with harmful queries. V-DEAL analyzes model behavior, understanding, and internal representations to pinpoint this issue, revealing that while models accurately recognize harmful content, their refusal tendency is weaker when visual understanding is involved compared to textual understanding. An intervention method using prompt injection was also introduced, significantly reducing attack success rates. AI
IMPACT Introduces a novel method for improving the safety and reliability of Video LLMs, potentially impacting their deployment in sensitive applications.
RANK_REASON Academic paper detailing a new diagnostic framework for Video LLMs. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- CatalyzeX
- DagsHub
- Gotit.pub
- Hugging Face
- ScienceCast
- V-DEAL
- Video Large Language Models
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →