PulseAugur
中
实时 00:26:40
English(EN) The Right Answer, the Wrong Direction: Why Transformers Fail at Counting and How to Fix It

研究人员发现 Transformer 知道计数但难以输出

一篇新论文指出了 Transformer 模型中一个特定的瓶颈,阻碍了它们执行计数任务的能力。研究人员发现,虽然 Pythia、Qwen3 和 Mistral 等模型在内部准确地存储计数信息,但它们难以将这些信息转化为正确的输出 token。对注意力权重进行有针对性的干预,显著提高了模型在自回归任务中生成正确计数的 ist, 表明输出路径存在几何错位。 AI

影响 识别出 Transformer 在计数任务中的特定读出瓶颈,可能指导未来的模型架构。

排序理由 该集群包含一篇学术论文,详细介绍了关于 Transformer 模型局限性的新发现。

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

研究人员发现 Transformer 知道计数但难以输出

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群包含一篇学术论文,详细介绍了关于 Transformer 模型局限性的新发现。
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
paper, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
156 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [2]

  1. arXiv cs.LG TIER_1 English(EN) · Gabriel Garcia ·

    正确的答案,错误的方向:为什么 Transformer 在计数方面会失败以及如何修复它

    arXiv:2605.03258v1 Announce Type: new Abstract: Large language models often fail at simple counting tasks, even when the items to count are explicitly present in the prompt. We investigate whether this failure occurs because transformers do not represent counts internally, or bec…

  2. arXiv cs.CL TIER_1 English(EN) · Gabriel Garcia ·

    正确的答案,错误的方向:为什么 Transformer 在计数方面会失败以及如何修复它

    Large language models often fail at simple counting tasks, even when the items to count are explicitly present in the prompt. We investigate whether this failure occurs because transformers do not represent counts internally, or because they cannot convert those representations i…