Proto-value Functions: A Laplacian Framework for Learning Representation and Control in Markov Decision Processes

MahadevanSridhar, MaggioniMauro · Journal of Machine Learning Research · 2007

This paper introduces a novel spectral framework for solving Markov decision processes (MDPs) by jointly learning representations and optimal policies. The major components of the framework describ...

Read the paper · More papers on PaperTik