Decentralized learning in two-player zero-sum games: A LR-I lagging anchor algorithm

Xiaosong Lu, Howard M. Schwartz · 2011

This paper presents a LR-Ilagging anchor algorithm that combines a lagging anchor method to the LR-Ilearning algorithm. We prove that this decentralized learning algorithm converges in strategies to a Nash equilibrium in two-player, zero-sum, two-action matrix games, while only needing knowledge of their own action and reward.

Read the paper · More papers on PaperTik