PulseAugur
实时 04:02:12
English(EN) GPT6 Astra ranks 5th on Artificial Analysis, tied with its predecessor. Also scores 99.9% on ARC AGI 3, where the average human scores 48%. Same model, two diff

GPT-6 Astra 的性能在更新的 AI 基准测试中受到审视

Artificial Analysis 已将其智能指数更新至 4.2 版本,解决了此前对其 GPT-6 Astra 评分的批评。虽然 GPT-6 Astra 在更新后的指数中得分高于其前代,但仍落后于 Anthropic 的 Claude Fable 5.1。另外,GPT-6 Astra 在 ARC AGI 3 基准测试中取得了近乎完美的得分,显著优于人类平均得分,并在编码任务中展现了先进的能力。 AI

影响 更新的基准测试和性能指标为理解领先 AI 模型之间的比较能力提供了见解。

排序理由 该集群讨论了基准测试结果和 AI 指数的更新,属于研究和产品评估范畴。

在 Mastodon — mastodon.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 4 个来源。 我们如何撰写摘要 →

GPT-6 Astra 的性能在更新的 AI 基准测试中受到审视

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群讨论了基准测试结果和 AI 指数的更新,属于研究和产品评估范畴。
Source corroboration
4 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
17 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [4]

  1. The Decoder TIER_1 English(EN) · Matthias Bastian ·

    Artificial Analysis 在 GPT-6 Astra 评分引发质疑后对其智能指数进行大修

    <p><img alt="" class="attachment-full size-full wp-post-image" height="768" src="https://the-decoder.com/wp-content/uploads/2026/09/openai_dark_gpt6_stars.png" style="height: auto; margin-bottom: 10px;" width="1376" /></p> <p> Artificial Analysis has released version 4.2 of its I…

  2. Mastodon — mastodon.social TIER_1 English(EN) · sidj79 ·

    GPT6 Astra 在 Artificial Analysis 中排名第五,与其前代持平。在 ARC AGI 3 测试中得分 99.9%,而人类平均得分仅为 48%。同一模型,两种不同

    GPT6 Astra ranks 5th on Artificial Analysis, tied with its predecessor. Also scores 99.9% on ARC AGI 3, where the average human scores 48%. Same model, two different stories. The gap: one prompt got it to open Blender, model and rig a character, then build a playable Unreal Engin…

  3. Mastodon — mastodon.social TIER_1 English(EN) · CuratedHackerNews ·

    GPT-6 Astra 在人工智能分析编码代理指数中取得重大进展 https:// artificialanalysis.ai/articles /benchmarking-gpt-6-astra # ai

    GPT-6 Astra makes major gains in the Artificial Analysis Coding Agent Index https:// artificialanalysis.ai/articles /benchmarking-gpt-6-astra # ai

  4. r/OpenAI TIER_2 English(EN) · /u/Blackham ·

    Gpt6 Astra 在人工智能分析指数上不敌 fable

    <table> <tr><td> <a href="https://www.reddit.com/r/OpenAI/comments/1w6jds7/gpt6_astra_does_not_beat_fable_on_artificial/"> <img alt="Gpt6 Astra does NOT beat fable on Artificial Analysis Index" src="https://external-preview.redd.it/bRHCg8t9KcpzuwcklnJ9_sS5HHMOqyIauV1XjEb1rl8.png?…