PulseAugur
实时 21:29:50
English(EN) 🤖 We got 100% on ARC-AGI-3 ft09 with zero model calls. The failures are more interesting. I've been building an experimental reasoning system at Orivael and tes

推理系统在 ARC-AGI-3 上实现 100% 准确率,无需 LLM

Orivael 开发的一个实验性推理系统在 ARC-AGI-3 ft09 基准测试中取得了完美的 100% 分数,并且没有使用任何大型语言模型。该系统的开发者强调,测试中遇到的失败比成功更有启发性。这一成就表明了一种不依赖当前 LLM 架构的、用于实现通用人工智能推理的新颖方法。 AI

影响 展示了复杂推理任务的潜在 LLM 替代方案,促使对非 LLM AGI 方法进行进一步研究。

排序理由 该条目描述了一个基准测试上的新颖研究发现,而非产品发布或重要的行业事件。[lever_c_demoted from research: ic=1 ai=1.0]

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

推理系统在 ARC-AGI-3 上实现 100% 准确率,无需 LLM

报道来源 [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🤖 我们在 ARC-AGI-3 ft09 上获得了 100% 的分数,没有调用任何模型。失败之处更有趣。我一直在 Orivael 构建一个实验性推理系统,并进行了测试

    🤖 We got 100% on ARC-AGI-3 ft09 with zero model calls. The failures are more interesting. I've been building an experimental reasoning system at Orivael and testing it against ARC-AGI-3. One of the runs just scored 100% on ft09. The unusual part: There is no LLM in the loop. Not …