PulseAugur
EN
LIVE 19:44:10

Anthropic unveils 'J-space' internal LLM workspace, enabling new interpretability tools · 9 sources tracked

Anthropic has published research detailing a "J-space," an internal "global workspace" within their language models like Claude. This workspace acts as a silent, temporary memory for intermediate variables during processing, analogous to human cognition. A new tool, the Jacobian lens (J-lens), allows researchers to access and analyze this J-space, revealing that it plays a crucial role in higher-order reasoning, though it constitutes a small fraction of the model's overall activity. The J-space's existence and the J-lens's utility have been independently replicated on models like Qwen 3.6 27B, suggesting significant implications for AI interpretability and safety. AI

IMPACT Provides a new method for understanding LLM reasoning, potentially improving safety and debugging capabilities by revealing internal 'thoughts'.

RANK_REASON The cluster reports on a research paper and its findings regarding internal model representations and interpretability tools.

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 10 sources. How we write summaries →

Anthropic unveils 'J-space' internal LLM workspace, enabling new interpretability tools · 9 sources tracked

COVERAGE [10]

  1. LessWrong (AI tag) TIER_1 English(EN) · TheManxLoiner ·

    Trying to grok Anthropic's Global Workspace paper and the J-space

    <h1><span>Introduction</span></h1><p><span>The aim of this post is to share a quick attempt at grokking the conceptual ideas that lie behind the notion of J-space and how it is calculated in the paper </span><a href="https://transformer-circuits.pub/2026/workspace/index.html#meth…

  2. LessWrong (AI tag) TIER_1 English(EN) · Neel Nanda ·

    A Review of Anthropic's Global Workspace Paper

    <p><i><span>The below is a public review Anthropic asked me to write for their new </span></i><a href="https://transformer-circuits.pub/2026/workspace/index.html" rel="noreferrer"><i><span>global workspace paper</span></i></a><i><span>. I recommend at least skimming their paper f…

  3. Medium — Anthropic tag TIER_1 English(EN) · Mateo Portillo ·

    Decoding the Anthropic Universe: Desktop vs. Cowork vs. Code vs. Dispatch

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://medium.com/@mateo.portillo_62955/decoding-the-anthropic-universe-desktop-vs-cowork-vs-code-vs-dispatch-cbdb83855f6a?source=rss------anthropic-5"><img src="https://cdn-images-1.medium.com/max/2600/0*NUgYl_…

  4. Medium — Claude tag TIER_1 English(EN) · Nadeem Khan(NK) ·

    A Global Workspace in Language Models: Anthropic Finds a Silent “J-Space” Inside Claude

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://nadeem4-nk13.medium.com/a-global-workspace-in-language-models-anthropic-finds-a-silent-j-space-inside-claude-a74a5a51f353?source=rss------claude-5"><img src="https://cdn-images-1.medium.com/max/2600/0*1J5…

  5. dev.to — Anthropic tag TIER_1 English(EN) · Breach Protocol ·

    Anthropic found a 'global workspace' inside its models - and a tool to read it

    <p>Anthropic reported on July 6, 2026 that its language models contain a 'global workspace' - a small set of internal patterns that behaves like a silent working memory the model can report on, deliberately control, and reason through. The company also released the tool that read…

  6. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    A global workspace in language models - https://www. anthropic.com/research/global- workspace fascinating stuff: "suggests a mental workspace supporting conscio

    A global workspace in language models - https://www. anthropic.com/research/global- workspace fascinating stuff: "suggests a mental workspace supporting conscious access isn’t just a peculiarity of how human brains happen to be wired." # ai

  7. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    A global workspace in language models https:// lobste.rs/s/xgtzrp # ai https://www. anthropic.com/research/global- workspace

    A global workspace in language models https:// lobste.rs/s/xgtzrp # ai https://www. anthropic.com/research/global- workspace

  8. r/LocalLLaMA TIER_1 English(EN) · /u/cuolong ·

    Anthropic Research - "Verbalizable Representations Form a Global Workspace in Language Models"

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1uq4as2/anthropic_research_verbalizable_representations/"> <img alt="Anthropic Research - &quot;Verbalizable Representations Form a Global Workspace in Language Models&quot;" src="https://external-preview.redd…

  9. r/LocalLLaMA TIER_1 English(EN) · /u/AutomataManifold ·

    Qwen's J-Space - Anthropic's discovery of an internal model Global Workspace

    <!-- SC_OFF --><div class="md"><p><a href="https://www.anthropic.com/research/global-workspace">Anthropic published research today</a> into what a model is thinking behind the scenes while it is deciding what to actually write. </p> <p>More importantly, they <a href="https://gith…

  10. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    Anthropic identifies an internal 'J-Space' in Claude that processes constructed scenarios before output. The finding suggests self-organized rea

    Anthropic identifiziert in Claude einen internen 'J-Space', der vor der Ausgabe konstruierte Szenarien verarbeitet. Der Befund deutet auf selbstorganisierte Reasoning-Pfade hin, die das Modell ohne explizite Anweisung entwickelt hat. https:// the-decoder.de/anthropic-zeigt -wie-c…