Reinforcement using supervised learning for policy generalization
Julien Laumônier · 2007
Applying reinforcement learning in large Markov Decision Process (MDP) is an important issue for solving very large problems. Since the exact resolution is often intractable, many approaches have been proposed to approximate the