Social Chemistry 101
PulseAugur coverage of Social Chemistry 101 — every cluster mentioning Social Chemistry 101 across labs, papers, and developer communities, ranked by signal.
-
New research questions effectiveness of activation steering in language models
A new research paper explores the phenomenon of activation steering in language models, questioning whether observed gains reflect intended control or compatibility with answer encodings. The study introduces Cross-Enco…
-
New method probes what activation steering truly controls in language models
Researchers have introduced a new evaluation method called Cross-Encoding Steering Evaluation to better understand what activation steering controls in language models. This method aims to distinguish between genuine co…
-
AI model rationales shift with fine-tuning and prompts, study finds
Researchers have developed a method to audit AI systems for how fine-tuning and prompt effects influence their rationales, particularly in high-conflict scenarios. Their experiments on LLaMA-3.2-11B, Qwen-3.5-9B, and Pi…