Anthropic is internally utilizing a new, unreleased AI model that significantly outperforms its current Mythos 5 model. This advanced model, referred to as 'Model 2', achieved a 12.5 percentage point higher score on the CoBench v2 benchmark, which assesses historical AI R&D problem-solving capabilities. Despite its superior performance, Anthropic has stated there are no current plans to release this model to the public. AI
IMPACT Highlights the internal advancement of AI capabilities within leading labs, suggesting a gap between internal tools and public offerings.
RANK_REASON The item discusses an internal model's performance and Anthropic's decision not to release it, rather than an official release or research paper.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →