Researchers have developed the Multi-modal Knowledge Preserving Adapter (MKP-Adapter), a novel approach for ensuring backward compatibility in multi-modal large language models (MLLMs) without needing to update the core backbone model. This adapter-only method addresses the challenge of preserving new embedding knowledge while maintaining compatibility, utilizing a multi-level preservation loss and a focal re-weighting strategy. Experiments show MKP-Adapter achieves strong backward compatibility across various multi-modal tasks, including image, text, visual document, and video retrieval, with minimal added latency. AI
IMPACT This method could significantly reduce the cost and complexity of updating embedding models in multi-modal AI systems.
RANK_REASON The cluster contains an academic paper detailing a new technical method for LLMs. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →