Concurrent Learning of Control in Multi-agent Sequential Decision Tasks

Bikramjit Banerjee · Aquila Digital Community (University of Southern Mississippi) · 2018

The overall objective of this project was to develop multi-agent reinforcement learning (MARL) approaches for intelligent agents to autonomously learn distributed control policies in decentralized partially observable Markov decision processes (Dec-POMDPs), without prior knowledge of the model parameters.

Read the paper · More papers on PaperTik