PulseAugur
实时 09:19:22
English(EN) GPT-6 Astra Runs 30.9 Minute Tasks. Codex Cannot Yet

OpenAI 的 GPT-6 Astra 显示任务范围延长 8.6 倍,但访问受限

根据英国人工智能安全研究所的数据,OpenAI 的新 GPT-6 Astra 模型展示了显著更长的自主任务范围,达到 30.9 分钟,而 GPT-5.6 Sol 为 3.6 分钟。这种扩展的能力允许更复杂、无人值守的任务,例如在不到 19 小时内从屏幕截图中完成游戏,而 GPT-5.6 Sol 完成此任务花费了超过 96 小时。尽管取得了这些进展,GPT-6 Astra 的访问仍然受到限制,许多用户,特别是那些依赖具有 ChatGPT 账户的 Codex 的用户,目前仍无法运行该模型。 AI

影响 在自主任务范围方面设定了新的 SOTA,可能支持更长的无人值守的代理工作流。

排序理由 前沿实验室模型发布,附带系统卡和独立基准测试。[lever_c 从 frontier_release 降级:ic=1 ai=1.0]

在 dev.to — Claude Code tag 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

OpenAI 的 GPT-6 Astra 显示任务范围延长 8.6 倍,但访问受限

本文如何被排名

Signal score
77 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
前沿实验室模型发布,附带系统卡和独立基准测试。[lever_c 从 frontier_release 降级:ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. dev.to — Claude Code tag TIER_1 English(EN) · RAXXO Studios ·

    GPT-6 Astra 可运行 30.9 分钟任务。Codex 尚不能

    <ul> <li><p>Codex rejects gpt-6.0-astra on ChatGPT accounts, so most solo devs cannot run it yet</p></li> <li><p>UK AISI measured autonomous task horizon at 30.9 minutes against 3.6 minutes for GPT-5.6 Sol</p></li> <li><p>Artificial Analysis puts cost per task between 0.46 USD at…