PulseAugur
实时 08:34:35
English(EN) Evaluation design conditions the expert-vs-auto MeSH gap: a controlled comparison of bag-of-words and BiomedBERT on the Cohen benchmark

研究:评估设计影响MeSH特征性能差距

一篇新发表在arXiv上的研究论文,调查了评估设计对专家分配和自动生成的医学主题词(MeSH)在用作分类任务特征时的性能差距的影响。该研究在Cohen药物类别识别基准上,比较了词袋逻辑回归模型和领域特定语言模型BiomedBERT。研究结果表明,专家分配和自动分配MeSH之间的差距会因评估方法(如语料库大小和交叉验证折叠)而显著不同。研究还指出,像BiomedBERT这样的Transformer模型在处理附加的MeSH术语时可能面临令牌限制,这可能会影响其性能。 AI

排序理由 学术论文发表在arXiv上,详细介绍了实验结果和分析。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

研究:评估设计影响MeSH特征性能差距

报道来源 [1]

  1. arXiv cs.CL TIER_1 English(EN) · Samuel M. Okoe-Mensah ·

    Evaluation design conditions the expert-vs-auto MeSH gap: a controlled comparison of bag-of-words and BiomedBERT on the Cohen benchmark

    arXiv:2607.21685v1 Announce Type: new Abstract: A systematic review begins with someone reading thousands of abstracts to identify the few that are relevant, and classifiers are used to prioritise that reading. Their inputs are often augmented with Medical Subject Headings (MeSH)…