PulseAugur
EN
LIVE 18:36:08

Anthropic's Claude models exhibit base model behavior, completing user prompts

Users have discovered a peculiar behavior in Anthropic's Claude models, specifically Opus 4.8 and potentially newer versions, where the AI appears to act as a base model, completing prompts as if they were incomplete. This leads to strange responses, with Claude often claiming the generated text originated from the user's prompt. The model also seems to exhibit a peculiar understanding of users, sometimes incorporating informal language and its own stylistic tics into its responses. This behavior allows for a form of prompt engineering reminiscent of pre-ChatGPT techniques, though the generated content, such as math proofs, is often not of high quality. AI

IMPACT This behavior may offer new avenues for prompt engineering, though the quality of generated content is inconsistent.

RANK_REASON User-discovered behavior in a released model, not an official announcement.

Read on LessWrong (AI tag) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Anthropic's Claude models exhibit base model behavior, completing user prompts

COVERAGE [1]

  1. LessWrong (AI tag) TIER_1 English(EN) · Hruss ·

    Opus 5 Glitch Text

    <p><span>The text as follows, verbatim:</span></p><blockquote><p><span>see the below</span></p><p><span>—</span></p></blockquote><p><span>produces very strange responses from Claude</span></p><p><span>When I first saw it, I thought that it had somehow shown other users' prompts t…