PulseAugur
实时 11:13:25
实体 Monotone Neural Policy Iteration

Monotone Neural Policy Iteration

PulseAugur coverage of Monotone Neural Policy Iteration — every cluster mentioning Monotone Neural Policy Iteration across labs, papers, and developer communities, ranked by signal.

Show in brief
总计 · 30天
1
90 天内 1
发布 · 30天
0
90 天内 0
论文 · 30天
1
90 天内 1
层级分布 · 90 天
主题
情绪 · 30 天

1 天有情绪数据

最近 · 第 1/1 页 · 共 1 条
  1. TOOL · CL_247853 ·

    新方法解决高维Hamilton-Jacobi-Bellman方程

    研究人员开发了一种新颖的神经半离散方法,用于求解高维一阶Hamilton-Jacobi-Bellman (HJB) 方程。该方法利用中心差分和人工粘性创建单调算子,然后使用移位网络查询进行评估。该方法允许策略迭代来求解Bellman方程,而无需张量网格,从而提高了适定性和数值依赖域的显式界限。