Concurrent Learning of Control in Multi-agent Sequential Decision Tasks
Bikramjit Banerjee · Aquila Digital Community (University of Southern Mississippi) · 2018
The overall objective of this project was to develop multi-agent reinforcement learning (MARL) approaches for intelligent agents to autonomously learn distributed control policies in decentralized partially observable Markov decision processes (Dec-POMDPs), without prior knowledge of the model parameters.