PulseAugur
实时 08:18:57
English(EN) Beauty is in the AI of the beholder: MLLMs systematically overrate facial attractiveness

研究发现AI模型一贯高估面部吸引力

发表在arXiv上的一项新研究显示,与人类判断相比,多模态大语言模型(MLLM)系统性地高估了面部吸引力。研究人员将2,513名人类参与者的评分与四个商业AI模型(ClaudeGemini、GPT和Grok)的评分进行了比较。虽然AI模型在吸引力的人类排名顺序上表现出很强的相关性,但它们对人脸的评分普遍比人类更积极,且范围更窄。只有面部年龄是人类和MLLM之间吸引力的一致预测因素,而种族和性别等其他因素在模型之间显示出不一致的模式。特别是Grok,其与人类评分的一致性最低。 AI

影响 表明当前MLLM在美容评估等主观任务上可能不可靠,突显了AI与人类感知之间的差异。

排序理由 发表在arXiv上的研究论文,详细介绍了AI模型能力的研究结果。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.CV 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

研究发现AI模型一贯高估面部吸引力

本文如何被排名

Signal score
17 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
发表在arXiv上的研究论文,详细介绍了AI模型能力的研究结果。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. arXiv cs.CV TIER_1 English(EN) · Santiago Grandas, Juan Sebastian Cely-Acosta, Mohit Mendiratta, Shafee Hassan, Macken Murphy ·

    美丑由AI决定:MLLM系统性地高估面部吸引力

    arXiv:2609.02512v1 Announce Type: new Abstract: Beauty assessments from Multimodal Large Language Models (MLLMs) are increasingly popular amongst users, companies, and aestheticians. This raises the question of whether these AI models can accurately reflect human judgments of att…