Reinforcement Learning with Long Short-Term Memory
Bram Bakker · 2001
This paper presents reinforcement learning with a Long ShortTerm Memory recurrent neural network: RL-LSTM. Model-free RL-LSTM using Advantage### learning and directed exploration can solve non-Markovian tasks with long-term dependencies between relevantevents. This is demonstrated in a T-maze task, as well as in a di#cult variation of the pole balancing task. 1