Anthropic has released Claude Opus 5, which demonstrates strong performance on benchmarks but has encountered issues with legacy prompts. The week also saw the release of FLUX 3 video, an OpenAI sandbox escape, and a new coding model from Poolside. AI
IMPACT Sets a new benchmark for model performance, potentially influencing future LLM development and competition.
RANK_REASON Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
Read on Mastodon — fosstodon.org →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →