Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhang, Songyuan, So, Oswin, Black, Mitchell, Serlin, Zachary, Fan, Chuchu
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912340419543040
author Zhang, Songyuan
So, Oswin
Black, Mitchell
Serlin, Zachary
Fan, Chuchu
author_facet Zhang, Songyuan
So, Oswin
Black, Mitchell
Serlin, Zachary
Fan, Chuchu
contents Tasks for multi-robot systems often require the robots to collaborate and complete a team goal while maintaining safety. This problem is usually formalized as a constrained Markov decision process (CMDP), which targets minimizing a global cost and bringing the mean of constraint violation below a user-defined threshold. Inspired by real-world robotic applications, we define safety as zero constraint violation. While many safe multi-agent reinforcement learning (MARL) algorithms have been proposed to solve CMDPs, these algorithms suffer from unstable training in this setting. To tackle this, we use the epigraph form for constrained optimization to improve training stability and prove that the centralized epigraph form problem can be solved in a distributed fashion by each agent. This results in a novel centralized training distributed execution MARL algorithm named Def-MARL. Simulation experiments on 8 different tasks across 2 different simulators show that Def-MARL achieves the best overall performance, satisfies safety constraints, and maintains stable training. Real-world hardware experiments on Crazyflie quadcopters demonstrate the ability of Def-MARL to safely coordinate agents to complete complex collaborative tasks compared to other methods.
format Preprint
id arxiv_https___arxiv_org_abs_2504_15425
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL
Zhang, Songyuan
So, Oswin
Black, Mitchell
Serlin, Zachary
Fan, Chuchu
Robotics
Artificial Intelligence
Machine Learning
Multiagent Systems
Optimization and Control
Tasks for multi-robot systems often require the robots to collaborate and complete a team goal while maintaining safety. This problem is usually formalized as a constrained Markov decision process (CMDP), which targets minimizing a global cost and bringing the mean of constraint violation below a user-defined threshold. Inspired by real-world robotic applications, we define safety as zero constraint violation. While many safe multi-agent reinforcement learning (MARL) algorithms have been proposed to solve CMDPs, these algorithms suffer from unstable training in this setting. To tackle this, we use the epigraph form for constrained optimization to improve training stability and prove that the centralized epigraph form problem can be solved in a distributed fashion by each agent. This results in a novel centralized training distributed execution MARL algorithm named Def-MARL. Simulation experiments on 8 different tasks across 2 different simulators show that Def-MARL achieves the best overall performance, satisfies safety constraints, and maintains stable training. Real-world hardware experiments on Crazyflie quadcopters demonstrate the ability of Def-MARL to safely coordinate agents to complete complex collaborative tasks compared to other methods.
title Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL
topic Robotics
Artificial Intelligence
Machine Learning
Multiagent Systems
Optimization and Control
url https://arxiv.org/abs/2504.15425