Anthropic has released Claude Sonnet 4.5, featuring a 200K token context window and a new "extended thinking mode." This mode allows the AI to interleave reasoning with action, pausing to reflect on intermediate results and adjust its approach mid-task. This capability is particularly beneficial for AI agents, especially in coding tasks, enabling them to reason over entire codebases, plan complex refactors across multiple files, and verify their work more reliably. While benchmarks show Sonnet 4.5 leading in multi-file refactoring, it slightly trails in single-file bug fixes compared to models like GPT-4.1 and Gemini 2.5 Pro. AI
IMPACT Enhances AI agent capabilities for complex reasoning and coding tasks, potentially accelerating development in agentic workflows.
RANK_REASON New model release from a frontier lab (Anthropic) with significant new capabilities. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
- 200K token context window
- AI agents
- Anthropic
- Claude API
- Claude Sonnet 4.5
- extended thinking mode
- Gemini 2.5 Pro
- GitHub
- GPT-4.1
- SWE-bench Verified
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →