PulseAugur
实时 06:45:05
实体 Regularized Latent Dynamics Prediction

Regularized Latent Dynamics Prediction

PulseAugur coverage of Regularized Latent Dynamics Prediction — every cluster mentioning Regularized Latent Dynamics Prediction across labs, papers, and developer communities, ranked by signal.

Show in brief
总计 · 30天
1
90 天内 1
发布 · 30天
0
90 天内 0
论文 · 30天
1
90 天内 1
层级分布 · 90 天
主题
情绪 · 30 天

1 天有情绪数据

最近 · 第 1/1 页 · 共 1 条
  1. TOOL · CL_221274 ·

    新的RLDP方法改进了行为基础模型的零样本强化学习

    研究人员推出了一种新的行为基础模型(BFMs)状态特征学习方法——正则化潜在动力学预测(RLDP)。BFMs旨在使智能体能够适应未知的奖励和任务,但其有效性受状态特征选择的限制。RLDP通过在潜在空间中添加一个自监督的下一状态预测目标的正交正则化来解决这个问题,这有助于保持特征多样性。该方法在零样本强化学习中已显示出可媲美甚至超越现有复杂表征学习技术的性能,特别是在其他方法表现不佳的、数据集覆盖有限的场景中表现出色。