PulseAugur
实时 04:40:44
English(EN) Large Language Models for Code Generation from Multilingual Prompts: A Curated Benchmark and a Study on Code Quality

研究发现大语言模型在代码生成中存在语言偏见 · 跟踪 3 个来源

一篇新发表在 arXiv 上的研究论文探讨了提示语言对不同大语言模型(LLMs)代码生成质量的影响。研究人员发现,用于提示 GPT-4o miniDeepSeekClaude 等模型的语言会显著影响生成代码的功能正确性和结构质量。该研究使用了包含 460 个 PythonJava 编码任务的基准,并将提示翻译成中文、印地语、西班牙语和意大利语,结果显示英文提示并不总是能产生最佳结果,并且生成的代码有时会混合语言。 AI

影响 这项研究强调了进行多语言提示工程的必要性,以优化不同大语言模型和编程语言的代码生成质量。

排序理由 该集群包含一篇详细介绍大语言模型代码生成基准和研究的论文。

在 arXiv cs.AI 阅读 →

AI 生成摘要 · Google Gemini · 来自 3 个来源。 我们如何撰写摘要 →

研究发现大语言模型在代码生成中存在语言偏见 · 跟踪 3 个来源

报道来源 [3]

  1. arXiv cs.AI TIER_1 English(EN) · Saima Afrin, Alessandro Midolo, Camilo Escobar-Vel\'asquez, Mario Linares-V\'asquez, Weiyuan Ding, Bowen Xu, Massimiliano Di Penta, Antonio Mastropaolo ·

    用于多语言提示代码生成的语言大模型:一个精选基准和代码质量研究

    arXiv:2607.14816v1 Announce Type: cross Abstract: Large Language Models (LLMs) perform differently on identical programming tasks when prompted in different natural languages, a phenomenon known as language bias. While this behavior has been widely studied for general text genera…

  2. arXiv cs.AI TIER_1 English(EN) · Antonio Mastropaolo ·

    用于多语言提示代码生成的语言大模型:一个精选基准和代码质量研究

    Large Language Models (LLMs) perform differently on identical programming tasks when prompted in different natural languages, a phenomenon known as language bias. While this behavior has been widely studied for general text generation, its impact on code generation quality and pr…

  3. Hugging Face Daily Papers TIER_1 English(EN) ·

    用于多语言提示代码生成的语言大模型:一个精选基准和代码质量研究

    Large Language Models (LLMs) perform differently on identical programming tasks when prompted in different natural languages, a phenomenon known as language bias. While this behavior has been widely studied for general text generation, its impact on code generation quality and pr…