A new benchmark called JuryBench has been developed to study how LLM-simulated jurors are influenced by defendant statements in U.S. criminal law cases. The research analyzed over 432,000 decisions from 20 frontier LLMs, finding that emotional persuasion can negatively impact perceived guilt, and that background affinity between defendants and jurors is a significant factor in sentencing. Ideology also plays a strong role in shaping severity judgments, highlighting both the potential and risks of using LLMs to model jury reasoning. AI
IMPACT Highlights potential biases in LLM-simulated legal decision-making, informing the development of fairer AI systems.
RANK_REASON Academic paper introducing a new benchmark and analysis of LLM behavior. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →