PulseAugur
实时 10:00:06
English(EN) IndicQE-APE: A Benchmark for Quality Estimation and Automatic Post-Editing for Indic Languages

发布了用于印度语言质量估计和后编辑的新基准

研究人员推出了 IndicQE-APE,这是一个旨在整合和评估印度语言质量估计和自动后编辑的新基准。该基准结合了 WMT 共享任务的数据和一个扩展的英语-马拉雅拉姆语资源,创建了一个包含九种语言对的超过 126,000 个实例的数据集。该研究在此新数据集上对几个大型语言模型和 COMET 指标进行了基准测试,揭示了它们性能和跨语言评估挑战的见解。 AI

影响 该基准旨在改进印度语言的 AI 模型的评估和开发,可能为这些地区带来更好的机器翻译和语言处理工具。

排序理由 该集群描述了在 arXiv 上发布的新学术基准和相关论文。

在 Hugging Face Daily Papers 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

发布了用于印度语言质量估计和后编辑的新基准

报道来源 [2]

  1. arXiv cs.CL TIER_1 English(EN) · Diptesh Kanojia, Archchana Sindhujan, Sourabh Deoghare, Daria Sokova, Shenbin Qian, Girish Koushik, Tharindu Ranasinghe, Constantin Or\u{a}san, Chrysoula Zerva, Ricardo Rei, Fr\'ed\'eric Blain, Andr\'e F. T. Martins, Marco Turchi, Matteo Negri, Rajen Cha… ·

    IndicQE-APE:面向印度语言的质量评估和自动后编辑基准

    arXiv:2608.16344v1 Announce Type: new Abstract: Indic quality estimation (QE) and automatic post-editing (APE) data is spread across separate releases, so no single resource supports training and evaluation across tasks and language pairs on one footing. We consolidate the WMT 20…

  2. Hugging Face Daily Papers TIER_1 English(EN) ·

    IndicQE-APE:面向印度语言的质量评估和自动后编辑基准

    Indic quality estimation (QE) and automatic post-editing (APE) data is spread across separate releases, so no single resource supports training and evaluation across tasks and language pairs on one footing. We consolidate the WMT 2020--2024 shared-task lineage with an extended En…