Opposition-Based Q(λ) Algorithm

Maryam Shokri, Hamid R. Tizhoosh, Mohamed S. Kamel · The 2006 IEEE International Joint Conference on Neural Network Proceedings · 2006

The problem of delayed reward in reinforcement learning is usually tackled by implementing the mechanism of eligibility traces. In this paper we introduce an extension of eligibility traces to solve one of the challenging problems in reinforcement learning. The concept of opposition traces is proposed in this work to deal with large state space problems in reinforcement learning applications. We combine the idea of opposition and eligibility traces to construct the opposition-based Q(lambda). The results are compared with the conventional Watkins' Q(lambda) and reflect a remarkable performance increase.

Read the paper · More papers on PaperTik