Artificial Analysis has updated its Intelligence Index to version 4.2, addressing criticisms regarding its previous scoring of GPT-6 Astra. While GPT-6 Astra now scores higher than its predecessor, it still falls behind Anthropic's Claude Fable 5.1 in the updated index. Separately, GPT-6 Astra achieved a near-perfect score on the ARC AGI 3 benchmark, significantly outperforming the average human score, and demonstrated advanced capabilities in coding tasks. AI
IMPACT Updated benchmarks and performance metrics provide insights into the comparative capabilities of leading AI models.
RANK_REASON The cluster discusses benchmark results and updates to an AI index, which falls under research and product evaluation.
Read on Mastodon — mastodon.social →
- Artificial Analysis
- Artificial Analysis Coding Agent Index
- GPT-6 Astra
- fable
- ARC AGI 3
- Claude Fable 5.1
AI-generated summary · Google Gemini · from 4 sources. How we write summaries →