PulseAugur
实时 18:08:51
English(EN) An interesting evolution from Mike Thelwall. A year ago, he was cautiously exploring whether LLMs could approximate expert research assessment. Today, he report

ChatGPT-5 mini 以 0.905 的相关性近似专家研究评估

Mike Thelwall 报告称,ChatGPT-5 mini 在专家研究质量评估方面取得了高达 0.905 的相关性。这项评估使用超过 107,000 篇英国研究论文的数据集,跨越多个学科,将大型语言模型的表现与 REF2021 质量评估进行了比较。 AI

影响 展示了大型语言模型在大型研究质量评估中提供协助的潜力,可能简化学术评估流程。

排序理由 关于大型语言模型在专家评估方面表现的研究论文。[lever_c_demoted from research: ic=1 ai=1.0]

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

ChatGPT-5 mini 以 0.905 的相关性近似专家研究评估

报道来源 [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    An interesting evolution from Mike Thelwall. A year ago, he was cautiously exploring whether LLMs could approximate expert research assessment. Today, he report

    An interesting evolution from Mike Thelwall. A year ago, he was cautiously exploring whether LLMs could approximate expert research assessment. Today, he reports that ChatGPT-5 mini reached correlations of up to 0.905 with expert REF2021 quality assessments across some discipline…