Reinforcement Learning with Long Short-Term Memory

Bram Bakker · 2001

This paper presents reinforcement learning with a Long ShortTerm Memory recurrent neural network: RL-LSTM. Model-free RL-LSTM using Advantage### learning and directed exploration can solve non-Markovian tasks with long-term dependencies between relevantevents. This is demonstrated in a T-maze task, as well as in a di#cult variation of the pole balancing task. 1

Read the paper · More papers on PaperTik