Brainstormers 2D - Team Description 2005
Martin Riedmiller, Thomas Gabel, Johannes F. Knabe, Hauke Strasdat · 2005
Abstract. The main focus of the Brainstormers ’ effort in the RoboCup soccer simulation 2D domain is to develop and to apply machine learning techniques in complex domains. In particular, we are interested in applying reinforcement learning methods, where the training signal is only given in terms of success or failure. Our final goal is a learning system, where we only plug in “win the match ” – and our agents learn to generate the appropriate behavior. Unfortunately, even from very optimistic complexity estimations it becomes obvious, that in the soccer simulation domain, both conventional solution methods and also advanced today’s reinforcement learning techniques come to their limit – there are more than (108 ×50) 23 different states and more than (1000) 300 different policies per agent per half time. This paper outlines the architecture of the Brainstormers team, focuses on the use of reinforcement learning to learn various elements of our agents ’ behavior, and highlights other advanced artificial intelligence methods we are employing. 1