Design of Scalable Population of Reinforcement Learning Agents for Autonomous 5G Radio Link Control
John Cosmas, Kareem Ali, Ali Mahbas, Prince Kwaku Boakye, John Miguel, Victor Gabillon, Alexandre Kazmierowski, Lewis Sear · 2024
This research demonstrates how MATLAB's Reinforcement Learning Markov Decision Process (MDP) Example Model can be used to design Radio Link Control MDP Reinforcement Learning (RL) agent. Since the number of agents in MATLAB's RL toolbox is not scalable beyond one agent, then an agent scalability scheme is required to design RL agents in MATLAB's RL toolbox and then realize multiple lightweight simultaneously operable Python instances of it for each of the multiple user equipment UE in a network.