Researchers have developed a new AI control mechanism called runtime action interference (RAI) that regulates action pacing and filters specific behaviors after an AI's initial inference. This system was implemented in a replication of AlphaStar for StarCraft II and tested in a human participant study. The study found that disclosing the AI's capabilities led to lower perceived fairness and higher perceived toxicity, while trust varied across different expertise levels. AI
IMPACT This research highlights the importance of separating AI capability disclosure from its control mechanisms to accurately assess human perception of fairness and toxicity.
RANK_REASON Academic paper detailing a new AI control mechanism and its evaluation. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →