Adversarial Reinforcement Learning in a Cyber Security Simulation

Richard Elderman, Leon J. J. Pater, Albert S. Thie, Mădălina M. Drugan, Marco Wiering · 2017

This paper focuses on cyber-security simulations in networks modeled as a Markov game with incomplete information and stochastic elements. The resulting game is an adversarial sequential decision making problem played with two agents, the attacker and defender. The two agents pit one reinforcement learning technique, like neural networks, Monte Carlo learning and Q-learning, against each other and examine their effectiveness against learning opponents. The results showed that Monte Carlo learning with the Softmax exploration strategy is most effective in performing the defender role and also for learning attacking strategies.

Read the paper · More papers on PaperTik