Neuronpedia
PulseAugur coverage of Neuronpedia — every cluster mentioning Neuronpedia across labs, papers, and developer communities, ranked by signal.
-
Anthropic unveils J-lens for debugging LLM internal states
Anthropic has developed a new interpretability technique called the Jacobian Lens (J-lens) to better understand the internal workings of large language models. This tool provides insights into intermediate concepts and …
-
Anthropic unveils J-Lens to visualize LLM internal thought processes
Anthropic has introduced a new interpretability technique called the Jacobian Lens (J-Lens) to visualize the internal thought processes of its large language models, specifically Claude. This J-Lens reveals a hidden "J-…
-
Anthropic unveils 'J-space' for Claude AI's internal reasoning
Anthropic has introduced a new internal mechanism for its Claude models called "J-space," which allows the AI to process and store information internally without generating external output. This J-space is described as …
-
Anthropic's Jacobian Lens tool available on Neuronpedia
The Jacobian Lens, a tool developed by Anthropic's Jacobian Lens library, is now available for specific models through Neuronpedia. This pre-fitted lens allows for local visualization and exploration of model behavior, …
-
Anthropic's NLA tech translates LLM 'thoughts' into human language
Anthropic has introduced Natural Language Autoencoders (NLAs), a new method that translates the internal numerical 'thoughts' (activations) of large language models into human-readable text. This technique allows resear…