Researchers have introduced AVBench, a new automated benchmark designed to evaluate audio-video generative models, particularly those focused on human-centric scenarios. The benchmark incorporates fine-grained metrics across visual quality, audio quality, and cross-modal consistency, aiming to capture details often missed by existing evaluations. AVBench utilizes specialized evaluators trained through preference learning on a large dataset, deriving continuous scores from binary decisions to better align with human judgment and serve as a reward signal for RLHF. AI
影响 Provides a more accurate and automated method for assessing the capabilities of audio-video generative models.
排序理由 The cluster contains a research paper introducing a new benchmark for evaluating AI models. [lever_c_demoted from research: ic=1 ai=1.0]
AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →