ENTITY
Hardt
Hardt
PulseAugur coverage of Hardt — every cluster mentioning Hardt across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
LLM evaluation bias: Winner's curse inflates performance metrics
A common practice in LLM evaluation, where multiple prompt variations are tested against a fixed dataset and the best-performing one is selected, can lead to inflated performance metrics. This is due to the 'winner's cu…
-
New paper suggests more AI models create pricing arbitrage opportunities
A recent paper by Olmedo, Schölkopf, and Hardt suggests that an increasing number of AI models creates opportunities for computational arbitrage in pricing. The research indicates that the proliferation of models can le…