PulseAugur
EN
LIVE 10:12:50

New research tackles deep learning bias, training dynamics, and reliability

Researchers are exploring new theoretical frameworks and practical methods to improve deep learning models. One paper introduces DISCO, a technique for mitigating dataset bias by estimating conditional distance correlation, outperforming existing methods across diverse datasets. Another study frames neural network training as a Hamilton-Jacobi problem, linking it to tropical algebra and PDEs, and offering insights into generalization and robustness. Additionally, new research challenges the assumption that calibration alone improves early-exit neural networks, proposing an alternative approach that considers prediction correctness and computation cost. Finally, studies are investigating how deep networks retain or forget their initial biases during training, with implications for understanding inductive bias and generalization. AI

IMPACT These papers introduce novel theoretical frameworks and practical methods for bias mitigation, understanding training dynamics, and improving model reliability, potentially leading to more robust and trustworthy AI systems.

RANK_REASON Multiple arXiv papers presenting novel research and methodologies in deep learning.

Read on arXiv cs.AI →

AI-generated summary · Google Gemini · from 11 sources. How we write summaries →

New research tackles deep learning bias, training dynamics, and reliability

COVERAGE [11]

  1. arXiv cs.AI TIER_1 English(EN) · Emre Kavak, Tom Nuno Wolf, Christian Wachinger ·

    DISCO: Mitigating Bias in Deep Learning with Conditional Distance Correlation

    arXiv:2506.11653v3 Announce Type: replace-cross Abstract: Dataset bias often leads deep learning models to exploit spurious correlations instead of task-relevant signals. We introduce the Standard Anti-Causal Model (SAM), a unifying causal framework that characterizes bias mechan…

  2. arXiv cs.AI TIER_1 English(EN) · Jose Marie Antonio Mi\~noza, Erika Fille T. Legara, Christopher P. Monterola ·

    The Hamilton-Jacobi Theory of Deep Learning

    arXiv:2605.28983v1 Announce Type: cross Abstract: In this paper, training a neural network is identified, exactly, as a search through Hamilton--Jacobi initial-value problems: each gradient step selects the initial data of a viscous Hamilton--Jacobi equation whose Hopf--Cole prop…

  3. arXiv cs.LG TIER_1 English(EN) · Piotr Kubaty, Filip Szatkowski, Grzegorz Choczy\'nski, Eric Nalisnick, Bartosz W\'ojcik ·

    Rethinking Calibration for Early-Exit Neural Networks

    arXiv:2508.21495v3 Announce Type: replace Abstract: Early-exit neural networks (EENNs) accelerate inference by allowing intermediate classifiers to stop computation once predictions are confident enough. Most methods rely on confidence thresholds for exiting, and consequently, im…

  4. arXiv cs.AI TIER_1 English(EN) · Ramya Hebbalaguppe, Ajay Shastry, Soumya Suvra Ghosal, Chetan Arora ·

    Enhancing Deep Neural Network Reliability with Refinement and Calibration

    arXiv:2605.23249v1 Announce Type: cross Abstract: Although deep neural networks (DNNs) achieve high predictive accuracy, their confidence estimates are often unreliable, potentially compromising user trust in their decisions. This has motivated research on calibrated models, wher…

  5. arXiv cs.AI TIER_1 English(EN) · Chetan Arora ·

    Enhancing Deep Neural Network Reliability with Refinement and Calibration

    Although deep neural networks (DNNs) achieve high predictive accuracy, their confidence estimates are often unreliable, potentially compromising user trust in their decisions. This has motivated research on calibrated models, where calibration measures how well a model's predicte…

  6. arXiv stat.ML TIER_1 English(EN) · Mohua Das, Pierfrancesco Beneventano, Shibshankar Dey, Gareth H. McKinkey, Tomaso Poggio ·

    Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias

    arXiv:2605.29152v1 Announce Type: cross Abstract: Randomly initialized neural networks induce a prior over functions, but the predictor used in practice is produced only after training. We ask how much of this initial bias survives the training pipeline. To make the question meas…

  7. arXiv stat.ML TIER_1 English(EN) · Minhao Yao, Ruoyu Wang, Xihong Lin, Lin Liu, Zhonghua Liu ·

    Deep Neural Network Training as Random Effects: An Optimization-Inference Duality

    arXiv:2605.27991v1 Announce Type: new Abstract: Deep neural networks (DNNs) have achieved remarkable empirical success, yet their training dynamics remain understood mainly from optimization rather than statistical principles. Here we develop a statistical framework for DNN train…

  8. arXiv stat.ML TIER_1 English(EN) · Tomaso Poggio ·

    Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias

    Randomly initialized neural networks induce a prior over functions, but the predictor used in practice is produced only after training. We ask how much of this initial bias survives the training pipeline. To make the question measurable, we introduce initialization memory: the de…

  9. arXiv stat.ML TIER_1 English(EN) · Zhonghua Liu ·

    Deep Neural Network Training as Random Effects: An Optimization-Inference Duality

    Deep neural networks (DNNs) have achieved remarkable empirical success, yet their training dynamics remain understood mainly from optimization rather than statistical principles. Here we develop a statistical framework for DNN training in the over-parameterized regime by showing …

  10. Medium — fine-tuning tag TIER_1 English(EN) · Louis Develle ·

    Stop Guessing: A Systematic Methodology for Tuning Deep Learning Models

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://medium.com/heuritech/stop-guessing-a-systematic-methodology-for-tuning-deep-learning-models-b3ea18e7e7c6?source=rss------fine_tuning-5"><img src="https://cdn-images-1.medium.com/max/2600/0*4S4N5fVa46wTjrt…

  11. Medium — fine-tuning tag TIER_1 English(EN) · Louis Develle ·

    Stop Guessing: A Systematic Methodology for Tuning Deep Learning Models

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://medium.com/@develle.louis/stop-guessing-a-systematic-methodology-for-tuning-deep-learning-models-b3ea18e7e7c6?source=rss------fine_tuning-5"><img src="https://cdn-images-1.medium.com/max/1024/1*Y17rzUfUhW…