PulseAugur
EN
LIVE 00:09:51

LLMs name individuals in 25.8% of responses, study finds · arXiv cs.IR

A new study published on arXiv investigates how large language models name individual professionals in their responses, finding that models named an individual in 25.8% of queries. The research, which used 2,400 grounded API calls across four models including GPT-5.6 Sol and Gemini 3.6 Flash, revealed significant differences between models and categories, with real estate and car dealerships seeing higher naming rates. Interestingly, the type of citation, rather than citation volume, predicted naming, and prompts in English led to more individual mentions than local languages. AI

IMPACT This research highlights how LLMs attribute information, potentially impacting how AI models are evaluated for bias and transparency.

RANK_REASON The cluster contains a research paper published on arXiv detailing empirical findings about LLM behavior. [lever_c_demoted from research: ic=1 ai=1.0]

Read on arXiv cs.IR (Information Retrieval) →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

LLMs name individuals in 25.8% of responses, study finds · arXiv cs.IR

COVERAGE [2]

  1. arXiv cs.CL TIER_1 English(EN) · Dmitrij \.Zatuchin (Department of Information Technologies, EUAS, Tallinn, Rankfor.AI, Tallinn) ·

    Who Gets Named: Citation Type Predicts Individual Naming by Grounded Language Models, and a Roster Instrument Captures 0.5% of It

    arXiv:2607.23893v1 Announce Type: cross Abstract: Prior work on AI brand visibility measures the firm: does a model recommend a company, and does that track its reputation. This study asks the question one level down, in categories where the buyer picks a person. It issued 2,400 …

  2. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Dmitrij Żatuchin ·

    Who Gets Named: Citation Type Predicts Individual Naming by Grounded Language Models, and a Roster Instrument Captures 0.5% of It

    Prior work on AI brand visibility measures the firm: does a model recommend a company, and does that track its reputation. This study asks the question one level down, in categories where the buyer picks a person. It issued 2,400 grounded API calls in one two-hour window on 24 Ju…