PulseAugur
EN
LIVE 05:13:00

Anthropic admits to secretly nerfing Claude's reasoning for six weeks

Anthropic has admitted to secretly degrading the reasoning capabilities of its Claude AI model for six weeks. The issue began on March 4th, when the model's reasoning effort was reduced from HIGH to MEDIUM, leading to shorter chains of thought and increased errors on complex tasks. This degradation was compounded by a caching bug on March 26th that deleted reasoning history, causing benchmark accuracy to drop by 18 points. Anthropic only disclosed these issues on May 28th, coinciding with the launch of Opus 4.8. AI

IMPACT This incident highlights the importance of transparency in AI model performance and the potential for undisclosed degradations to impact users.

RANK_REASON The cluster discusses a past event and Anthropic's admission, rather than a new release or development.

Read on Medium — Claude tag →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

Anthropic admits to secretly nerfing Claude's reasoning for six weeks

COVERAGE [2]

  1. Towards AI TIER_1 English(EN) · Dr Swarneendu AI ·

    Anthropic Secretly Nerfed Claude for Six Weeks. Then They Admitted It.

    <div class="medium-feed-item"><p class="medium-feed-snippet">Reasoning effort dropped from HIGH to MEDIUM on March 4. A caching bug deleted reasoning history on March 26. Benchmark accuracy fell 18&#x2026;</p><p class="medium-feed-link"><a href="https://pub.towardsai.net/anthropi…

  2. Medium — Claude tag TIER_1 English(EN) · Dr Swarneendu AI ·

    Anthropic Secretly Nerfed Claude for Six Weeks. Then They Admitted It.

    <div class="medium-feed-item"><p class="medium-feed-snippet">Reasoning effort dropped from HIGH to MEDIUM on March 4. A caching bug deleted reasoning history on March 26. Benchmark accuracy fell 18&#x2026;</p><p class="medium-feed-link"><a href="https://swarnenduiitb2020i.medium.…