PulseAugur
实时 20:11:08

Riazi-8B: 乌尔都语大语言模型增强低资源语言的数学推理能力

研究人员开发了Riazi-8B,一个专门为乌尔都语数学推理设计的新型大语言模型。该模型解决了现有以英语为中心的大语言模型的局限性,这些模型在乌尔都语等低资源语言上的表现不佳。Riazi-8B通过两步过程创建:首先在乌尔都语维基百科上进行预训练,然后使用从GSM8K派生的乌尔都语思维链数据进行微调。在MGSM-Urdu基准测试上的评估表明,与其他的乌尔都语指令微调模型相比,Riazi-8B在答案正确性、推理质量和乌尔都语生成方面有了显著提升。 AI

影响 将大语言模型的数学推理能力扩展到低资源语言,可能使乌尔都语用户受益,并为未来的多语言人工智能开发提供信息。

排序理由 该集群描述了一篇详细介绍为低资源语言创建和评估专用大语言模型的新研究论文。

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Riazi-8B: 乌尔都语大语言模型增强低资源语言的数学推理能力

报道来源 [2]

  1. arXiv cs.CL TIER_1 English(EN) · Azher Ali, Ibtsam Haider, Raja Khurram Shahzad, Seemab Latif, Mehwish Fatima ·

    Riazi-8B:一个用于数学推理的乌尔都语大型语言模型

    arXiv:2606.25568v1 Announce Type: new Abstract: Recent LLMs demonstrate strong mathematical reasoning capabilities, but existing gains rely heavily on English-centric training resources and benchmarks. As a result, reasoning performance degrades substantially in low-resource lang…

  2. arXiv cs.CL TIER_1 English(EN) · Mehwish Fatima ·

    Riazi-8B:一个用于数学推理的乌尔都语大型语言模型

    Recent LLMs demonstrate strong mathematical reasoning capabilities, but existing gains rely heavily on English-centric training resources and benchmarks. As a result, reasoning performance degrades substantially in low-resource languages such as Urdu, where reasoning-oriented dat…