Q-Managed: A new algorithm for a multiobjective reinforcement learning

Thiago Henrique Freire de Oliveira, Luiz Paulo de Souza Medeiros, Adrião Duarte Dória Neto, Jorge D. de Melo · Software Impacts · 2021

Multi-objective reinforcement learning involves the use of reinforcement learning techniques to address problems with multiple objectives. To resolve this, we use a hybrid multi-objective optimization method that provides the mathematical guarantee that all policies belonging to the Pareto Front can be found. The hybridization gave rise to Q-Managed , which is given by the ε − constraint method and the Q-Learning algorithm, where the first limits the environment dynamically based on the agent's learning. Thus, when a region no longer provides improvement, it becomes a constraint, preventing the agent from returning. The simplicity and its performance come from a single-policy algorithms.

Read the paper · More papers on PaperTik