PulseAugur
EN
LIVE 08:19:09

New arXiv papers explore risk-based MDPs and Bellman equation dualities

Two new research papers published on arXiv explore advanced concepts in sequential decision-making for artificial intelligence. The first paper introduces ERQDP, a novel method for finite-horizon Markov decision process planning under risk-based objectives, which offers an enumeration-free approach and provides certified solutions or explicit residual gaps. The second paper delves into the theoretical underpinnings of the Bellman equation, demonstrating how its recursive properties stem from three fundamental conditions related to state dynamics, return decomposition, and uncertainty aggregation, unifying concepts across reinforcement learning, control, and decision theory. AI

IMPACT These papers advance theoretical frameworks for AI decision-making, potentially improving the robustness and efficiency of AI agents in complex environments.

RANK_REASON Two academic papers published on arXiv detailing theoretical advancements in AI decision-making.

Read on arXiv cs.AI →

AI-generated summary · Google Gemini · from 3 sources. How we write summaries →

New arXiv papers explore risk-based MDPs and Bellman equation dualities

COVERAGE [3]

  1. arXiv cs.LG TIER_1 English(EN) · Kavya Ravichandran ·

    Algorithmic Approaches to Sequential Decision-Making and Social Epistemology

    arXiv:2607.20636v1 Announce Type: cross Abstract: As humans, we face many decisions that require us to choose between sticking to something and giving up. This thesis uses algorithmic tools to derive insights about such decision-making problems in theoretical models, studying bot…

  2. arXiv cs.AI TIER_1 English(EN) · Irmaan (Mohammad), Mirzanejad, Nadjet Bourdache, Abdel-Illah Mouaddib ·

    Long-Term Sequential Decision Making under Risk

    arXiv:2607.19914v1 Announce Type: new Abstract: We study finite-horizon MDP planning under \emph{root-based} (resolute) risk objectives that apply a rank-dependent functional to the distribution of total returns. Such objectives are non-linear in the return distribution and gener…

  3. arXiv cs.AI TIER_1 English(EN) · Fernando E. Rosas, David Hyland, Daniel Polani ·

    Generalised Bellman recurrence and three dualities in sequential decision-making

    arXiv:2607.18077v1 Announce Type: cross Abstract: What gives the Bellman equation its form? We show that the recursive properties of optimal value functions follow from three conditions: that the dynamics decomposes through sufficient statistics, that the return decomposes recurs…