PulseAugur
实时 19:44:59
English(EN) A global workspace in language models https:// lobste.rs/s/xgtzrp # ai https://www. anthropic.com/research/global- workspace

Anthropic 发布内部 LLM 工作空间 'J-space',支持新的可解释性工具 · 跟踪 9 个来源

Anthropic 发布了一项研究,详细介绍了其 Claude 等语言模型内部的“J-space”,一个内部的“全局工作空间”。该工作空间在处理过程中充当中间变量的无声临时内存,类似于人类认知。一种名为 Jacobian lens (J-lens) 的新工具允许研究人员访问和分析这个 J-space,揭示它在更高阶推理中起着至关重要的作用,尽管它仅占模型整体活动的一小部分。J-space 的存在和 J-lens 的效用已在 Qwen 3.6 27B 等模型上得到独立复制,这表明它对 AI 可解释性和安全性具有重大意义。 AI

影响 提供了一种理解 LLM 推理的新方法,通过揭示内部“思考”过程,有可能提高安全性和调试能力。

排序理由 该集群报道了一篇研究论文及其关于模型内部表示和可解释性工具的发现。

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 10 个来源。 我们如何撰写摘要 →

Anthropic 发布内部 LLM 工作空间 'J-space',支持新的可解释性工具 · 跟踪 9 个来源

报道来源 [10]

  1. LessWrong (AI tag) TIER_1 English(EN) · TheManxLoiner ·

    尝试理解Anthropic的Global Workspace论文和J-space

    <h1><span>Introduction</span></h1><p><span>The aim of this post is to share a quick attempt at grokking the conceptual ideas that lie behind the notion of J-space and how it is calculated in the paper </span><a href="https://transformer-circuits.pub/2026/workspace/index.html#meth…

  2. LessWrong (AI tag) TIER_1 English(EN) · Neel Nanda ·

    Anthropic 全球工作空间论文回顾

    <p><i><span>The below is a public review Anthropic asked me to write for their new </span></i><a href="https://transformer-circuits.pub/2026/workspace/index.html" rel="noreferrer"><i><span>global workspace paper</span></i></a><i><span>. I recommend at least skimming their paper f…

  3. Medium — Anthropic tag TIER_1 English(EN) · Mateo Portillo ·

    解读Anthropic宇宙:Desktop vs. Cowork vs. Code vs. Dispatch

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://medium.com/@mateo.portillo_62955/decoding-the-anthropic-universe-desktop-vs-cowork-vs-code-vs-dispatch-cbdb83855f6a?source=rss------anthropic-5"><img src="https://cdn-images-1.medium.com/max/2600/0*NUgYl_…

  4. Medium — Claude tag TIER_1 English(EN) · Nadeem Khan(NK) ·

    语言模型中的全局工作空间:Anthropic 在 Claude 中发现了一个无声的“J-Space”

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://nadeem4-nk13.medium.com/a-global-workspace-in-language-models-anthropic-finds-a-silent-j-space-inside-claude-a74a5a51f353?source=rss------claude-5"><img src="https://cdn-images-1.medium.com/max/2600/0*1J5…

  5. dev.to — Anthropic tag TIER_1 English(EN) · Breach Protocol ·

    Anthropic 在其模型中发现了一个“全局工作区”——以及一个读取它的工具

    <p>Anthropic reported on July 6, 2026 that its language models contain a 'global workspace' - a small set of internal patterns that behaves like a silent working memory the model can report on, deliberately control, and reason through. The company also released the tool that read…

  6. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    语言模型中的全球工作空间 - https://www.anthropic.com/research/global-workspace 有趣的内容:“暗示着一个支持意识的心理工作空间

    A global workspace in language models - https://www. anthropic.com/research/global- workspace fascinating stuff: "suggests a mental workspace supporting conscious access isn’t just a peculiarity of how human brains happen to be wired." # ai

  7. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    语言模型中的全球工作空间 https://lobste.rs/s/xgtzrp # ai https://www.anthropic.com/research/global-workspace

    A global workspace in language models https:// lobste.rs/s/xgtzrp # ai https://www. anthropic.com/research/global- workspace

  8. r/LocalLLaMA TIER_1 English(EN) · /u/cuolong ·

    Anthropic 研究 - “可言喻表征在语言模型中形成全局工作空间”

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1uq4as2/anthropic_research_verbalizable_representations/"> <img alt="Anthropic Research - &quot;Verbalizable Representations Form a Global Workspace in Language Models&quot;" src="https://external-preview.redd…

  9. r/LocalLLaMA TIER_1 English(EN) · /u/AutomataManifold ·

    Qwen的J-Space - Anthropic发现内部模型全局工作空间

    <!-- SC_OFF --><div class="md"><p><a href="https://www.anthropic.com/research/global-workspace">Anthropic published research today</a> into what a model is thinking behind the scenes while it is deciding what to actually write. </p> <p>More importantly, they <a href="https://gith…

  10. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    Anthropic在Claude中发现内部“J-Space”,可在输出前处理构建的场景。该发现表明其具有自组织推理能力

    Anthropic identifiziert in Claude einen internen 'J-Space', der vor der Ausgabe konstruierte Szenarien verarbeitet. Der Befund deutet auf selbstorganisierte Reasoning-Pfade hin, die das Modell ohne explizite Anweisung entwickelt hat. https:// the-decoder.de/anthropic-zeigt -wie-c…