PulseAugur
EN
LIVE 09:57:26

GPT-6 Astra benchmarks released, tested against AGI criteria

New benchmarks have been released for GPT-6 Astra, a model that is being evaluated against advanced tests for artificial general intelligence (AGI). One proposed benchmark, attributed to Demis Hassabis, involves training the model on knowledge up to 1911 and assessing its ability to independently develop general relativity, similar to Einstein's achievement. This test aims to distinguish true intelligence and creativity from mere knowledge retrieval and synthesis, with proponents arguing that an AI would have significant advantages over historical human scientific development. AI

IMPACT Evaluates advanced AI capabilities against theoretical AGI benchmarks, potentially setting new standards for AI intelligence.

RANK_REASON The cluster discusses benchmarks and evaluation criteria for a model, aligning with research.

Read on r/singularity →

AI-generated summary · Google Gemini · from 6 sources. How we write summaries →

GPT-6 Astra benchmarks released, tested against AGI criteria

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
The cluster discusses benchmarks and evaluation criteria for a model, aligning with research.
Source corroboration
6 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
5 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+2 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

Full methodology in our editorial standards.

COVERAGE [6]

  1. The Decoder TIER_1 English(EN) · Maximilian Schreiner ·

    Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward

    <p><img alt="" class="attachment-full size-full wp-post-image" height="768" src="https://the-decoder.com/wp-content/uploads/2026/09/openai_logo_astra-1.png" style="height: auto; margin-bottom: 10px;" width="1376" /></p> <p> OpenAI's GPT-6 Astra is drawing contradictory benchmark …

  2. r/OpenAI TIER_2 English(EN) · /u/DataLearnerAI ·

    GPT-6 Astra vs Claude Fable 5.1 based on the benchmarks available so far

    <table> <tr><td> <a href="https://www.reddit.com/r/OpenAI/comments/1w6sw9c/gpt6_astra_vs_claude_fable_51_based_on_the/"> <img alt="GPT-6 Astra vs Claude Fable 5.1 based on the benchmarks available so far" src="https://preview.redd.it/mbkpusrw9fnh1.png?width=140&amp;height=114&amp…

  3. r/OpenAI TIER_2 Deutsch(DE) · /u/CartographerAble9446 ·

    GPT Astra benchmarks

    <table> <tr><td> <a href="https://www.reddit.com/r/OpenAI/comments/1w6g0is/gpt_astra_benchmarks/"> <img alt="GPT Astra benchmarks" src="https://preview.redd.it/bsdtay3xncnh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=95f6c896efefa2f2328d948a3434e4f1ab19928d" title="GPT Astr…

  4. r/singularity TIER_2 English(EN) · /u/FalconsArentReal ·

    Coding Benchmarks for GPT-6 Astra

    <table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1w6h7je/coding_benchmarks_for_gpt6_astra/"> <img alt="Coding Benchmarks for GPT-6 Astra" src="https://preview.redd.it/nu68nh3bvcnh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=6a456c7590fb747d45e9c072054…

  5. r/singularity TIER_2 English(EN) · /u/CounterReady4774 ·

    Gpt 6 astra benchmarks

    <table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1w6f9xo/gpt_6_astra_benchmarks/"> <img alt="Gpt 6 astra benchmarks" src="https://preview.redd.it/moqytexcjcnh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=6b92f19588a57d25231bba95bbdf07d3a120dd17" title=…

  6. r/singularity TIER_2 English(EN) · /u/Neurogence ·

    Can GPT-6 Astra Pass The Demis Hassabis Benchmark For AGI?

    <!-- SC_OFF --><div class="md"><p>Demis Hassabis has always said that a great way to determine whether we have AGI would be to <strong>train a foundation model with a knowledge cutoff around 1911 and see whether it could independently develop general relativity</strong>, as Einst…