PulseAugur
实时 09:20:13
English(EN) UpliftBench: Revealing Outcome-Regime and Objective Mismatch in Uplift Evaluation

新的UpliftBench基准揭示了提升建模中的指标不匹配

引入了一个名为UpliftBench的新基准来评估提升建模,这是一种用于个性化定位的技术。该基准揭示了提升估计器性能排名中的分歧主要是由于不同的指标而非模型本身造成的。UpliftBench在各种数据集上采用了多目标协议,并强调像Qini和AUUC这样的指标与实际效果准确性显示出不同程度的一致性,其中一些指标在结构上不足以支持某些政策决策。 AI

影响 强调了评估提升模型中的关键问题,可能导致更准确的个性化定位系统。

排序理由 该集群包含一篇介绍用于评估机器学习模型的新基准的研究论文。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.LG 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新的UpliftBench基准揭示了提升建模中的指标不匹配

报道来源 [1]

  1. arXiv cs.LG TIER_1 English(EN) · Binshuang Li ·

    UpliftBench: Revealing Outcome-Regime and Objective Mismatch in Uplift Evaluation

    arXiv:2608.00915v1 Announce Type: new Abstract: Uplift modeling (conditional-average-treatment-effect estimation) drives personalized targeting, yet published uplift benchmarks frequently disagree on which estimator performs best; we show the disagreement is substantially about m…