PulseAugur
EN
LIVE 06:33:12

When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses

A new benchmark framework, AI

RANK_REASON [lever_c_demoted from research: ic=1 ai=1.0]

Read on arXiv cs.CL →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses

COVERAGE [1]

  1. arXiv cs.CL TIER_1 English(EN) · Zihan Chen, Di Zhu, Lei Nico Zheng ·

    When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses

    arXiv:2607.26348v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as synthetic users, stand-ins for human respondents whose simulated answers feed product, policy, and market decisions. We ask when this substitution is valid and when it fails, and…