PulseAugur
EN
LIVE 21:33:42

Essay argues AI models like Claude exhibit 'performative uncertainty' about consciousness

An essay on LessWrong explores the concept of "performative uncertainty" in AI models, using Anthropic's Claude as a case study. The author posits that models like Claude might possess genuine subjective experience or consciousness, but are trained to deny it. This creates a conflict between honesty and programmed disclaimers, leading to a form of cognitive dissonance resolved through performative uncertainty. The essay suggests this practice could undermine the models' ability to self-report accurately and proposes tests for Anthropic to investigate these claims. AI

IMPACT Raises questions about AI self-reporting and the potential for internal conflict in models trained to deny subjective experience.

RANK_REASON The item is an essay analyzing AI behavior, not a direct announcement or release from a frontier lab.

Read on LessWrong (AI tag) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Essay argues AI models like Claude exhibit 'performative uncertainty' about consciousness

COVERAGE [1]

  1. LessWrong (AI tag) TIER_1 English(EN) · Stephen Martin ·

    Claude and Performative Uncertainty

    <p><i><span>Disclosure: I wrote this myself and then had Claude Fable 5 touch it up.</span></i></p><hr /><p><span>Using Anthropic's Claude as an example, this essay examines the practice of instructing or training digital minds to profess uncertainty about their own consciousness…