OpenAI launches GPT-5.6 family with new benchmarks and features · 8 sources tracked
ByPulseAugur Editorial·[14 sources]·
OpenAI has launched its new GPT-5.6 family of models, including Sol, Terra, and Luna, which are now available across ChatGPT, Codex, and the OpenAI API. GPT-5.6 Sol demonstrates significant advancements, setting new state-of-the-art benchmarks on the Artificial Analysis Coding Agent Index and Agents' Last Exam, outperforming competitors like Claude Fable 5. The new models offer improved artifact quality, enhanced design judgment, and a high-performance "ultra mode" that utilizes multiple agents in parallel for demanding tasks, all while using fewer tokens and at a lower cost.
AI
IMPACT
Sets new SOTA on coding and agent benchmarks, potentially pressuring competitors and accelerating enterprise adoption of advanced AI capabilities.
RANK_REASON
Frontier-lab model release with system card
GPT‑5.6 is available starting today across ChatGPT, Codex, and the OpenAI API. The rollout is starting globally now and will continue gradually toward full availability over the next 24 hours.
In ChatGPT, Plus, Pro, Business, and Enterprise users access GPT-5.6 Sol through
GPT‑5.6 improves artifact quality across presentations, documents, and spreadsheets, and works better with your templates.
These editable artifacts can be exported to the tools professionals already use and refined as part of real enterprise workflows. https://t.co/dfA4KBmrGu
GPT‑5.6 delivers a step change in design judgment.
Its stronger computer-use capabilities let it inspect and refine the rendered result—not just generate the underlying code or content—so it can catch visual and functional issues and apply finishing touches before handing the ht…
GPT-5.6 launches with ultra mode, our highest-performance setting to accelerate your most ambitious work by coordinating multiple agents to work in parallel.
It trades higher token use for stronger and faster results on demanding tasks. https://t.co/UojPDNOfSH
On the Artificial Analysis Coding Agent Index, GPT‑5.6 Sol sets a new state of the art at 80.0—2.8 points above Claude Fable 5—while using less than half the output tokens, taking less than half the time, and costing about one-third less. https://t.co/H5o0qJKUOL
GPT-5.6 Sol sets a new standard for intelligence and efficiency, delivering state-of-the-art performance across coding, knowledge work, cybersecurity, and science with fewer tokens and lower cost.
https://t.co/7uR7WdyOQz
On Agents' Last Exam, GPT‑5.6 Sol sets a new high of 53.6, eclipsing Claude Fable 5 (adaptive) by 13.1 points.
At medium reasoning, it beats Fable 5 by 11.4 points at roughly one-quarter the estimated cost. GPT‑5.6 Terra and Luna also outperforms Fable 5 at around one-sixteenth …
GPT-5.6 is rolling out in the API.
Sol is our flagship model, leading in coding, knowledge work, cybersecurity, and science.
Terra delivers performance competitive with GPT-5.5 at lower cost. Luna is our fastest, most affordable model for high-volume tasks. https://t.co/bHeEdIL…
The release shows the power the U.S. government now holds in the AI model landscape. ChatGPT Work highlights how OpenAI continues to evolve into an enterprise vendor.
Email — Every
TIER_1English(EN)·0100019f4cb941ee-8b6ddc88-9351-463c-a740-30aeb05937a6-000000@send.every.to (0100019f4cb941ee-8b6ddc88-9351-463c-a740-30aeb05937a6-000000@send.every.to)·
<!-- Set the language of your main document. This helps screenreaders use the proper language profile, pronunciation, and accent. --> <!-- The title is useful for screenreaders reading a document. Use your sender name or subject line. --> How GPT-5.6 Changes Knowledge Work <!-- N…
Bluesky Jetstream — AI desk
TIER_1English(EN)·simonwillison.net·
Notes on GPT-5.6, which includes some interesting new additions to the API (programmatic tool calling and multi-agent in particular) - plus 18 pelicans for the 6 reasoning levels and 3 new models: simonwillison.net/2026/Jul/9/g...
<h2> The 'Aha!' Moment: When GPT-5.6 Sol Changed My Workflow </h2> <p>The spreadsheet stared back at me, a chaotic grid of quarterly sales data, user feedback snippets, and raw API logs. My goal was simple, in theory: synthesize this mess into a coherent Q3 performance review for…
GPT-5.6 Sol wprowadza system pięciu poziomów rozumowania, w którym tryb Ultra koordynuje cztery subagenty jednocześnie, by osiągnąć rekordowe 91,9% w testach wydajności. # si # ai # sztucznainteligencja # wiadomości # informacje # technologia https:// aisight.pl/technologia/gener…