This paper explores learning dynamics in linear-quadratic stochastic games where players have limited information about their opponents. Researchers developed an $\epsilon$-greedy iterated least-squares algorithm that converges to the Nash equilibrium even without full system parameter knowledge. The study applied this to a dynamic Cournot competition, finding that limited information and high price stickiness reduce firm profits and market welfare, though revealing aggregate output can accelerate convergence and mitigate these losses. AI
IMPACT This research could inform the development of AI agents capable of strategic decision-making in complex, competitive environments.
RANK_REASON Academic paper detailing a new learning algorithm for stochastic games. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- CatalyzeX
- Connected Papers
- Cournot
- DagsHub
- Gotit.pub
- Hugging Face
- Influence Flower
- Litmaps
- Nash
- ScienceCast
- scite Smart Citations
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →