PulseAugur
中
实时 20:41:56
Русский(RU) Anthropic нашла у Claude «J-пространство» — внутренний узел, похожий на осознанный доступ у человека: anthropic claude ai

Anthropic 在 Claude AI 中发现“J-space”,可实现因果干预

Anthropic 的研究人员在其 Claude AI 模型中识别出了一组特定的激活模式,他们称之为“J-space”。这个内部“工作空间”在功能上类似于人类的意识接入,能够容纳有限数量的概念并介导复杂的推理。一种新颖的方法——“雅可比透镜”(J-lens)——不仅被用来观察这些模式,还通过干预和改变它们来因果性地验证它们在模型输出中的作用。这项技术有可能识别出 AI 何时在捏造数据或隐藏其真实目标,为处理 LLM 的开发者提供了实际应用。 AI

影响 使开发者能够检测潜在的 AI 欺骗并理解内部推理过程。

排序理由 研究论文,详细介绍了一种分析 LLM 的新内部机制和方法。[lever_c_demoted from research: ic=1 ai=1.0]

在 dev.to — LLM tag 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Anthropic 在 Claude AI 中发现“J-space”,可实现因果干预

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
研究论文,详细介绍了一种分析 LLM 的新内部机制和方法。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
83 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. dev.to — LLM tag TIER_1 Русский(RU) · Promptra Team ·

    Anthropic 在 Claude 中发现“J空间”——一种类似于人类有意识访问的内部节点:anthropic claude ai

    <p>Короткий ответ: 6 июля 2026 года Anthropic опубликовала исследование, в котором описала внутри Claude компактный набор паттернов активации — они называют его J-space — функционально похожий на «осознанный доступ» (global workspace) у человека. Ключевое слово тут «функционально…