PulseAugur
实时 21:38:46
English(EN) GPT-6 Astra scores highest on GoBench

GPT-6 Astra 在新的基于围棋的推理基准测试中领先

一项名为 GoBench 的新基准测试已推出,旨在通过围棋游戏衡量通用推理能力。该基准测试显示,GPT-6 Astra 的 Elo 评分为 2568,高于 Sol (1929 Elo) 和 Opus 5 high (2076 Elo) 等其他模型。此外,当与编码能力集成时,GPT-6 Astra 的 Codex 变体达到了 3563 的 Elo,显著优于 Sol 的 Codex (2656 Elo)。 AI

影响 建立了一种新的 AI 推理评估方法,可能推动向更通用智能的发展。

排序理由 新基准测试的推出和 AI 模型的性能结果。[lever_c_demoted from research: ic=1 ai=1.0]

在 r/OpenAI 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

GPT-6 Astra 在新的基于围棋的推理基准测试中领先

本文如何被排名

Signal score
8 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
新基准测试的推出和 AI 模型的性能结果。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. r/OpenAI TIER_2 English(EN) · /u/Roland31415 ·

    GPT-6 Astra 在 GoBench 上得分最高

    <table> <tr><td> <a href="https://www.reddit.com/r/OpenAI/comments/1wi75t2/gpt6_astra_scores_highest_on_gobench/"> <img alt="GPT-6 Astra scores highest on GoBench" src="https://preview.redd.it/umg4j6oaoxph1.png?width=140&amp;height=69&amp;auto=webp&amp;s=2853b7a1e05b66877ffaa1029…