NN DynaQ with prioritized sweeping and multiple predecessors
Benoît Girard, Lise Aubin · Figshare · 2018
Figures illustrating the model, the task and the results of an neural network implementation of a DynaQ reinforcement learning algorithm with prioritized sweeping, in a case where a (state,action) couple can have multiple predecessors.