一篇新论文全面概述了博弈中的学习,探讨了单智能体决策过程和多智能体交互。它引入了一系列正则化学习策略,旨在平衡探索与利用。该工作提出了对抗性赌博机的遗憾界限以及零和博弈的均衡收敛结果,将战略稳定性与动态学习吸引子联系起来。 AI
排序理由 该条目是提交到arXiv的研究论文。[lever_c_demoted from research: ic=1 ai=0.7]
- alphaXiv
- arXiv
- CatalyzeX Code Finder for Papers
- computer science
- DagsHub
- game theory
- Gotit.pub
- Hugging Face
- Influence Flower
- Panayotis Mertikopoulos
- Regret, equilibrium, and learning in games: A guided tour
- ScienceCast
AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →