Empirical Study on Robustness and Resilience in Cooperative Multi-Agent Reinforcement Learning

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Li, Simin, Mao, Zihao, Li, Hanxiao, Jing, Zonglei, bian, Zhuohang, Guo, Jun, Wang, Li, Han, Zhuoran, Xu, Ruixiao, Yu, Xin, Ma, Chengdong, Ma, Yuqing, An, Bo, Yang, Yaodong, Lv, Weifeng, Liu, Xianglong
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866914108766420992
author Li, Simin
Mao, Zihao
Li, Hanxiao
Jing, Zonglei
bian, Zhuohang
Guo, Jun
Wang, Li
Han, Zhuoran
Xu, Ruixiao
Yu, Xin
Ma, Chengdong
Ma, Yuqing
An, Bo
Yang, Yaodong
Lv, Weifeng
Liu, Xianglong
author_facet Li, Simin
Mao, Zihao
Li, Hanxiao
Jing, Zonglei
bian, Zhuohang
Guo, Jun
Wang, Li
Han, Zhuoran
Xu, Ruixiao
Yu, Xin
Ma, Chengdong
Ma, Yuqing
An, Bo
Yang, Yaodong
Lv, Weifeng
Liu, Xianglong
contents In cooperative Multi-Agent Reinforcement Learning (MARL), it is a common practice to tune hyperparameters in ideal simulated environments to maximize cooperative performance. However, policies tuned for cooperation often fail to maintain robustness and resilience under real-world uncertainties. Building trustworthy MARL systems requires a deep understanding of robustness, which ensures stability under uncertainties, and resilience, the ability to recover from disruptions--a concept extensively studied in control systems but largely overlooked in MARL. In this paper, we present a large-scale empirical study comprising over 82,620 experiments to evaluate cooperation, robustness, and resilience in MARL across 4 real-world environments, 13 uncertainty types, and 15 hyperparameters. Our key findings are: (1) Under mild uncertainty, optimizing cooperation improves robustness and resilience, but this link weakens as perturbations intensify. Robustness and resilience also varies by algorithm and uncertainty type. (2) Robustness and resilience do not generalize across uncertainty modalities or agent scopes: policies robust to action noise for all agents may fail under observation noise on a single agent. (3) Hyperparameter tuning is critical for trustworthy MARL: surprisingly, standard practices like parameter sharing, GAE, and PopArt can hurt robustness, while early stopping, high critic learning rates, and Leaky ReLU consistently help. By optimizing hyperparameters only, we observe substantial improvement in cooperation, robustness and resilience across all MARL backbones, with the phenomenon also generalizing to robust MARL methods across these backbones. Code and results available at https://github.com/BUAA-TrustworthyMARL/adv_marl_benchmark .
format Preprint
id arxiv_https___arxiv_org_abs_2510_11824
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Empirical Study on Robustness and Resilience in Cooperative Multi-Agent Reinforcement Learning
Li, Simin
Mao, Zihao
Li, Hanxiao
Jing, Zonglei
bian, Zhuohang
Guo, Jun
Wang, Li
Han, Zhuoran
Xu, Ruixiao
Yu, Xin
Ma, Chengdong
Ma, Yuqing
An, Bo
Yang, Yaodong
Lv, Weifeng
Liu, Xianglong
Multiagent Systems
Artificial Intelligence
Machine Learning
In cooperative Multi-Agent Reinforcement Learning (MARL), it is a common practice to tune hyperparameters in ideal simulated environments to maximize cooperative performance. However, policies tuned for cooperation often fail to maintain robustness and resilience under real-world uncertainties. Building trustworthy MARL systems requires a deep understanding of robustness, which ensures stability under uncertainties, and resilience, the ability to recover from disruptions--a concept extensively studied in control systems but largely overlooked in MARL. In this paper, we present a large-scale empirical study comprising over 82,620 experiments to evaluate cooperation, robustness, and resilience in MARL across 4 real-world environments, 13 uncertainty types, and 15 hyperparameters. Our key findings are: (1) Under mild uncertainty, optimizing cooperation improves robustness and resilience, but this link weakens as perturbations intensify. Robustness and resilience also varies by algorithm and uncertainty type. (2) Robustness and resilience do not generalize across uncertainty modalities or agent scopes: policies robust to action noise for all agents may fail under observation noise on a single agent. (3) Hyperparameter tuning is critical for trustworthy MARL: surprisingly, standard practices like parameter sharing, GAE, and PopArt can hurt robustness, while early stopping, high critic learning rates, and Leaky ReLU consistently help. By optimizing hyperparameters only, we observe substantial improvement in cooperation, robustness and resilience across all MARL backbones, with the phenomenon also generalizing to robust MARL methods across these backbones. Code and results available at https://github.com/BUAA-TrustworthyMARL/adv_marl_benchmark .
title Empirical Study on Robustness and Resilience in Cooperative Multi-Agent Reinforcement Learning
topic Multiagent Systems
Artificial Intelligence
Machine Learning
url https://arxiv.org/abs/2510.11824