Utility based Q-learning to facilitate cooperation in Prisoner's Dilemma games

Koichi Moriyama · Web Intelligence and Agent Systems An International Journal · 2009

This work deals with Q-learning in a multiagent environment. There are many multiagent Q-learning methods, and most of them aim to converge to a Nash equilibrium, which is not desirable in games like the Prisoner's Dilemma (PD). However, normal Q-lea

Read the paper · More papers on PaperTik