Tambe, MilindKwun, Mujin2023-07-0620232023-06-302023Kwun, Mujin. 2023. Encouraging Cooperation in Multi Agent Reinforcement Learning. Bachelor's thesis, Harvard College.30315543https://nrs.harvard.edu/URN-3:HUL.INSTREPOS:37376428Encouraging cooperation in Multi-agent reinforcement learning (MARL) remains a big area of research. In addition, additional complexity as well as non-stationarity when scaling up from the single-agent setting makes convergence to optimal policies difficult compared to single-agent reinforcement learning. In this thesis, we build on previous work demonstrating the empirical effectiveness of policy-gradient methods in multi-agent settings, specifically Proximal Policy Optimization(PPO). We introduce a novel test-bed for multi-agent reinforcement learning and evaluate the effectiveness of a decentralized PPO framework in this test-bed. Furthermore, motivated by literature that shows the benefits of reward shaping on convergence in the single agent setting, we apply domain specific reward shaping to our PPO method to encourage cooperation and faster convergence to a good joint policy.application/pdfenStatisticsEncouraging Cooperation in Multi Agent Reinforcement LearningThesis or Dissertation2023-07-06