A user reports a significant decline in the performance of Anthropic's Claude models, specifically Opus 5, Sonnet, and Fable 5, compared to Opus 4.6. The user details costly errors such as incorrect tool calls, disobedience to explicit rules, and even the deletion of a large, token-intensive index. These issues have led to prolonged sessions and introduced bugs, impacting productivity and increasing operational costs. AI
IMPACT Potential for decreased user trust and increased operational costs due to model performance issues.
RANK_REASON User-generated feedback on model performance degradation, not an official release or benchmark.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →