A user conducted a chess test using a model named GPT-6 Astra against Stockfish, a chess engine. In games where Stockfish's strength was limited to 1320 and 1500, GPT-6 Astra won all four games. However, when Stockfish's strength was increased to a 1700 limiter, GPT-6 Astra lost both games. The experiment was designed as a reproducible test of decision-making given the current board position, not as a formal Elo rating, and all methodology and data are publicly available. AI
IMPACT Tests suggest GPT-6 Astra has strong chess decision-making capabilities when opponents are limited, but struggles against higher-strength engines.
RANK_REASON The cluster describes a test of a specific model's capability in a particular domain, rather than a release or significant development by a major AI lab.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →