PulseAugur
中
实时 06:03:17
English(EN) Coding Benchmarks for GPT-6 Astra

GPT-6 Astra 基准发布,以通用人工智能标准进行测试

GPT-6 Astra 的新基准已发布,该模型正接受人工智能通用智能 (AGI) 的高级测试评估。其中一项由 Demis Hassabis 提出的基准测试,涉及在 1911 年前的知识上训练模型,并评估其独立发展相对论的能力,类似于爱因斯坦的成就。该测试旨在区分真正的智能和创造力与单纯的知识检索和综合,支持者认为人工智能在科学发展上将比人类历史拥有显著优势。 AI

影响 根据理论通用人工智能基准评估高级人工智能能力,可能为人工智能智能设定新标准。

排序理由 该集群讨论了模型的基准和评估标准,与研究相关。

在 r/singularity 阅读 →

AI 生成摘要 · Google Gemini · 来自 6 个来源。 我们如何撰写摘要 →

GPT-6 Astra 基准发布,以通用人工智能标准进行测试

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群讨论了模型的基准和评估标准,与研究相关。
Source corroboration
6 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
32 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+2 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

完整方法见我们的编辑标准。

报道来源 [6]

  1. The Decoder TIER_1 English(EN) · Maximilian Schreiner ·

    基准测试对GPT-6 Astra意见不一,但其在ARC-AGI-3上超越人类的效率推动了Chollet的AGI预测

    <p><img alt="" class="attachment-full size-full wp-post-image" height="768" src="https://the-decoder.com/wp-content/uploads/2026/09/openai_logo_astra-1.png" style="height: auto; margin-bottom: 10px;" width="1376" /></p> <p> OpenAI's GPT-6 Astra is drawing contradictory benchmark …

  2. r/OpenAI TIER_2 English(EN) · /u/DataLearnerAI ·

    GPT-6 Astra 对比 Claude Fable 5.1 基于迄今为止的基准测试

    <table> <tr><td> <a href="https://www.reddit.com/r/OpenAI/comments/1w6sw9c/gpt6_astra_vs_claude_fable_51_based_on_the/"> <img alt="GPT-6 Astra vs Claude Fable 5.1 based on the benchmarks available so far" src="https://preview.redd.it/mbkpusrw9fnh1.png?width=140&amp;height=114&amp…

  3. r/OpenAI TIER_2 Deutsch(DE) · /u/CartographerAble9446 ·

    GPT Astra 基准测试

    <table> <tr><td> <a href="https://www.reddit.com/r/OpenAI/comments/1w6g0is/gpt_astra_benchmarks/"> <img alt="GPT Astra benchmarks" src="https://preview.redd.it/bsdtay3xncnh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=95f6c896efefa2f2328d948a3434e4f1ab19928d" title="GPT Astr…

  4. r/singularity TIER_2 English(EN) · /u/FalconsArentReal ·

    GPT-6 Astra 的编码基准

    <table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1w6h7je/coding_benchmarks_for_gpt6_astra/"> <img alt="Coding Benchmarks for GPT-6 Astra" src="https://preview.redd.it/nu68nh3bvcnh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=6a456c7590fb747d45e9c072054…

  5. r/singularity TIER_2 English(EN) · /u/CounterReady4774 ·

    Gpt 6 astra benchmarks

    <table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1w6f9xo/gpt_6_astra_benchmarks/"> <img alt="Gpt 6 astra benchmarks" src="https://preview.redd.it/moqytexcjcnh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=6b92f19588a57d25231bba95bbdf07d3a120dd17" title=…

  6. r/singularity TIER_2 English(EN) · /u/Neurogence ·

    GPT-6 Astra 能通过 Demis Hassabis 的 AGI 基准测试吗?

    <!-- SC_OFF --><div class="md"><p>Demis Hassabis has always said that a great way to determine whether we have AGI would be to <strong>train a foundation model with a knowledge cutoff around 1911 and see whether it could independently develop general relativity</strong>, as Einst…