Amirali Abdullah
PulseAugur coverage of Amirali Abdullah — every cluster mentioning Amirali Abdullah across labs, papers, and developer communities, ranked by signal.
-
New framework uses psychometric tests to evaluate LLM behavioral consistency
Researchers have developed a new framework for evaluating the behavioral consistency of large language models (LLMs) using situational judgment tests (SJTs) and multidimensional item response theory (MIRT). This approac…
-
New research identifies three distinct modes of sycophancy in large language models
A new research paper published on arXiv and highlighted by Hugging Face explores the phenomenon of sycophancy in large language models. The study challenges the view of sycophancy as a single behavioral dimension, propo…
-
Paper calls for auditable mechanistic interpretability guidelines
A new paper proposes a system for auditable mechanistic interpretability (MI) to address inconsistencies in current research. The authors call for a continuous, collaborative reviewing platform to organize meta-science …