Major-Minor Mean Field Multi-Agent Reinforcement Learning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Cui, Kai, Fabian, Christian, Tahir, Anam, Koeppl, Heinz
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866929337311166464
author Cui, Kai
Fabian, Christian
Tahir, Anam
Koeppl, Heinz
author_facet Cui, Kai
Fabian, Christian
Tahir, Anam
Koeppl, Heinz
contents Multi-agent reinforcement learning (MARL) remains difficult to scale to many agents. Recent MARL using Mean Field Control (MFC) provides a tractable and rigorous approach to otherwise difficult cooperative MARL. However, the strict MFC assumption of many independent, weakly-interacting agents is too inflexible in practice. We generalize MFC to instead simultaneously model many similar and few complex agents -- as Major-Minor Mean Field Control (M3FC). Theoretically, we give approximation results for finite agent control, and verify the sufficiency of stationary policies for optimality together with a dynamic programming principle. Algorithmically, we propose Major-Minor Mean Field MARL (M3FMARL) for finite agent systems instead of the limiting system. The algorithm is shown to approximate the policy gradient of the underlying M3FC MDP. Finally, we demonstrate its capabilities experimentally in various scenarios. We observe a strong performance in comparison to state-of-the-art policy gradient MARL methods.
format Preprint
id arxiv_https___arxiv_org_abs_2303_10665
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Major-Minor Mean Field Multi-Agent Reinforcement Learning
Cui, Kai
Fabian, Christian
Tahir, Anam
Koeppl, Heinz
Machine Learning
Multiagent Systems
Optimization and Control
Multi-agent reinforcement learning (MARL) remains difficult to scale to many agents. Recent MARL using Mean Field Control (MFC) provides a tractable and rigorous approach to otherwise difficult cooperative MARL. However, the strict MFC assumption of many independent, weakly-interacting agents is too inflexible in practice. We generalize MFC to instead simultaneously model many similar and few complex agents -- as Major-Minor Mean Field Control (M3FC). Theoretically, we give approximation results for finite agent control, and verify the sufficiency of stationary policies for optimality together with a dynamic programming principle. Algorithmically, we propose Major-Minor Mean Field MARL (M3FMARL) for finite agent systems instead of the limiting system. The algorithm is shown to approximate the policy gradient of the underlying M3FC MDP. Finally, we demonstrate its capabilities experimentally in various scenarios. We observe a strong performance in comparison to state-of-the-art policy gradient MARL methods.
title Major-Minor Mean Field Multi-Agent Reinforcement Learning
topic Machine Learning
Multiagent Systems
Optimization and Control
url https://arxiv.org/abs/2303.10665