NN DynaQ with prioritized sweeping and multiple predecessors

Benoît Girard, Lise Aubin · Figshare · 2018

Figures illustrating the model, the task and the results of an neural network implementation of a DynaQ reinforcement learning algorithm with prioritized sweeping, in a case where a (state,action) couple can have multiple predecessors.

Read the paper · More papers on PaperTik