Multi-Agent Reinforcement Learning with Selective State-Space Models

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Daniel, Jemma, de Kock, Ruan, Nessir, Louay Ben, Abramowitz, Sasha, Mahjoub, Omayma, Khlifi, Wiem, Formanek, Claude, Pretorius, Arnu
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866914992823992320
author Daniel, Jemma
de Kock, Ruan
Nessir, Louay Ben
Abramowitz, Sasha
Mahjoub, Omayma
Khlifi, Wiem
Formanek, Claude
Pretorius, Arnu
author_facet Daniel, Jemma
de Kock, Ruan
Nessir, Louay Ben
Abramowitz, Sasha
Mahjoub, Omayma
Khlifi, Wiem
Formanek, Claude
Pretorius, Arnu
contents The Transformer model has demonstrated success across a wide range of domains, including in Multi-Agent Reinforcement Learning (MARL) where the Multi-Agent Transformer (MAT) has emerged as a leading algorithm in the field. However, a significant drawback of Transformer models is their quadratic computational complexity relative to input size, making them computationally expensive when scaling to larger inputs. This limitation restricts MAT's scalability in environments with many agents. Recently, State-Space Models (SSMs) have gained attention due to their computational efficiency, but their application in MARL remains unexplored. In this work, we investigate the use of Mamba, a recent SSM, in MARL and assess whether it can match the performance of MAT while providing significant improvements in efficiency. We introduce a modified version of MAT that incorporates standard and bi-directional Mamba blocks, as well as a novel "cross-attention" Mamba block. Extensive testing shows that our Multi-Agent Mamba (MAM) matches the performance of MAT across multiple standard multi-agent environments, while offering superior scalability to larger agent scenarios. This is significant for the MARL community, because it indicates that SSMs could replace Transformers without compromising performance, whilst also supporting more effective scaling to higher numbers of agents. Our project page is available at https://sites.google.com/view/multi-agent-mamba .
format Preprint
id arxiv_https___arxiv_org_abs_2410_19382
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Multi-Agent Reinforcement Learning with Selective State-Space Models
Daniel, Jemma
de Kock, Ruan
Nessir, Louay Ben
Abramowitz, Sasha
Mahjoub, Omayma
Khlifi, Wiem
Formanek, Claude
Pretorius, Arnu
Machine Learning
Artificial Intelligence
Multiagent Systems
I.2.11
The Transformer model has demonstrated success across a wide range of domains, including in Multi-Agent Reinforcement Learning (MARL) where the Multi-Agent Transformer (MAT) has emerged as a leading algorithm in the field. However, a significant drawback of Transformer models is their quadratic computational complexity relative to input size, making them computationally expensive when scaling to larger inputs. This limitation restricts MAT's scalability in environments with many agents. Recently, State-Space Models (SSMs) have gained attention due to their computational efficiency, but their application in MARL remains unexplored. In this work, we investigate the use of Mamba, a recent SSM, in MARL and assess whether it can match the performance of MAT while providing significant improvements in efficiency. We introduce a modified version of MAT that incorporates standard and bi-directional Mamba blocks, as well as a novel "cross-attention" Mamba block. Extensive testing shows that our Multi-Agent Mamba (MAM) matches the performance of MAT across multiple standard multi-agent environments, while offering superior scalability to larger agent scenarios. This is significant for the MARL community, because it indicates that SSMs could replace Transformers without compromising performance, whilst also supporting more effective scaling to higher numbers of agents. Our project page is available at https://sites.google.com/view/multi-agent-mamba .
title Multi-Agent Reinforcement Learning with Selective State-Space Models
topic Machine Learning
Artificial Intelligence
Multiagent Systems
I.2.11
url https://arxiv.org/abs/2410.19382