Online inverse reinforcement learning for nonlinear systems
Ryan Self, Michael Harlan, Rushikesh Kamalapurkar · 2019 IEEE Conference on Control Technology and Applications (CCTA) · 2019
This paper focuses on the development of an online inverse reinforcement learning (IRL) technique for a class of nonlinear systems. The developed approach utilizes observed state and input trajectories, and determines the unknown reward function and the unknown value function online. A parameter estimation technique is utilized to facilitate estimation of the reward function in the presence of unknown dynamics. Theoretical guarantees for convergence of the reward function estimates are established, and simulation results are presented to demonstrate the performance of the developed technique.