D-study
PulseAugur coverage of D-study — every cluster mentioning D-study across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New protocol standardizes auditing of LLM brand recommendations
Researchers have introduced the Dice Roll Method, a standardized protocol for auditing the brand recommendations of large language models (LLMs). This method addresses the lack of standardization in current auditing pra…
-
New framework evaluates dependability in AI-powered essay scoring
Researchers have introduced a new conditional generalizability framework to evaluate the dependability of automated essay scoring systems. This framework treats encoder architectures and scoring-head families as a unive…
-
New framework evaluates AI scoring dependability across diverse conditions
A new conditional generalizability framework has been introduced to evaluate the dependability of automated scoring systems, particularly in contexts like automated essay scoring. This framework treats different encoder…