Modelling the dynamics of multiagent Q-learning with ε-greedy exploration
Eduardo Rodrigues Gomes, Ryszard Kowalczyk · 2009
We present a framework to model the dynamics of Multiagent Q-learning with ε-greedy exploration. The applicability of the framework is tested through experiments in typical games selected from the literature.