Trial and Error Experience Replay Based Deep Reinforcement Learning

Cheng Zhang, Liang Ma · 2019

The environment with sparse rewards in reinforcement learning is a common problem and the agent learns inefficiently using general methods. A new solution called trialand-error experience replay is proposed. In this method, the general hindsight experience replay is combined with a curiositydriven model, by which the sample-efficiency will be improved although extrinsic rewards are sparse. It is demonstrated as an algorithm to control a virtual robotic arm to reach a mobile goal. Through analysis the robotic arm can explore and learn based on failure trajectories which shows that the agent mimics a human who failed repeatedly but still tries to learn something from the unexpected outcomes.

Read the paper · More papers on PaperTik