PulseAugur
实时 09:17:28
English(EN) Language Equality has a Price: A Systematic Investigation of Multi-turn LLM Performance for EU-24+

研究发现:商业 LLM 在 30 种语言上的表现优于开源模型

一篇新发表在 arXiv 上的研究调查了大型语言模型 (LLM) 在包括 24 种欧盟官方语言在内的 30 种语言上的性能。研究发现,即使在网络文本量明显较少的语言中,商业 LLM 的表现也始终优于开源模型。研究还显示,尽管商业系统在全球范围内保持优势,但非英语语言的运行成本更高,平均得分也低于英语。 AI

影响 凸显了商业 LLM 和开源 LLM 在多语言能力方面存在的巨大差距,表明当前的开源模型不足以实现全球语言的平等。

排序理由 该集群包含一篇评估 LLM 在特定基准上性能的学术论文。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

研究发现:商业 LLM 在 30 种语言上的表现优于开源模型

报道来源 [1]

  1. arXiv cs.CL TIER_1 English(EN) · Sherzod Hakimov, Karl Osswald, Jelle Psurek, Eszter Bukovszky, A. Altar L\"user, David Schlangen ·

    Language Equality has a Price: A Systematic Investigation of Multi-turn LLM Performance for EU-24+

    arXiv:2608.01395v1 Announce Type: new Abstract: We evaluate large language models (LLMs) as language agents playing goal-directed dialogue games in self-play across 30 languages: the 24 official EU languages plus six others. Unlike static or preference-based evaluation, this para…