Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
Fuente:
arXiv
Saved in:
| Main Authors: | Bonnet, Clément, Luo, Daniel, Byrne, Donal, Surana, Shikha, Abramowitz, Sasha, Duckworth, Paul, Coyette, Vincent, Midgley, Laurence I., Tegegn, Elshadai, Kalloniatis, Tristan, Mahjoub, Omayma, Macfarlane, Matthew, Smit, Andries P., Grinsztajn, Nathan, Boige, Raphael, Waters, Cemlyn N., Mimouni, Mohamed A., Sob, Ulrich A. Mbou, de Kock, Ruan, Singh, Siddarth, Furelos-Blanco, Daniel, Le, Victor, Pretorius, Arnu, Laterre, Alexandre |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Should we be going MAD? A Look at Multi-Agent Debate Strategies for LLMs
by: Smit, Andries, et al.
Published: (2023)
by: Smit, Andries, et al.
Published: (2023)
Combinatorial Optimization with Policy Adaptation using Latent Space Search
by: Chalumeau, Felix, et al.
Published: (2023)
by: Chalumeau, Felix, et al.
Published: (2023)
Oryx: a Scalable Sequence Model for Many-Agent Coordination in Offline MARL
by: Formanek, Claude, et al.
Published: (2025)
by: Formanek, Claude, et al.
Published: (2025)
Generative Model for Small Molecules with Latent Space RL Fine-Tuning to Protein Targets
by: Sob, Ulrich A. Mbou, et al.
Published: (2024)
by: Sob, Ulrich A. Mbou, et al.
Published: (2024)
SPO: Sequential Monte Carlo Policy Optimisation
by: Macfarlane, Matthew V, et al.
Published: (2024)
by: Macfarlane, Matthew V, et al.
Published: (2024)
How much can change in a year? Revisiting Evaluation in Multi-Agent Reinforcement Learning
by: Singh, Siddarth, et al.
Published: (2023)
by: Singh, Siddarth, et al.
Published: (2023)
Efficiently Quantifying Individual Agent Importance in Cooperative MARL
by: Mahjoub, Omayma, et al.
Published: (2023)
by: Mahjoub, Omayma, et al.
Published: (2023)
Multi-Agent Reinforcement Learning with Selective State-Space Models
by: Daniel, Jemma, et al.
Published: (2024)
by: Daniel, Jemma, et al.
Published: (2024)
Generalisable Agents for Neural Network Optimisation
by: Tessera, Kale-ab, et al.
Published: (2023)
by: Tessera, Kale-ab, et al.
Published: (2023)
Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation
by: Osman, Asim, et al.
Published: (2026)
by: Osman, Asim, et al.
Published: (2026)
Breaking the Performance Ceiling in Reinforcement Learning requires Inference Strategies
by: Chalumeau, Felix, et al.
Published: (2025)
by: Chalumeau, Felix, et al.
Published: (2025)
Sable: a Performant, Efficient and Scalable Sequence Model for MARL
by: Mahjoub, Omayma, et al.
Published: (2024)
by: Mahjoub, Omayma, et al.
Published: (2024)
Overconfident Oracles: Limitations of In Silico Sequence Design Benchmarking
by: Surana, Shikha, et al.
Published: (2025)
by: Surana, Shikha, et al.
Published: (2025)
Memory-Enhanced Neural Solvers for Routing Problems
by: Chalumeau, Felix, et al.
Published: (2024)
by: Chalumeau, Felix, et al.
Published: (2024)
Sarcoidose: Uma forma rara de apresentação
by: Jessica Cemlyn -Jones
Published: (2009)
by: Jessica Cemlyn -Jones
Published: (2009)
Avaliação da densidade mineral óssea em doentes com fibrose quística
by: Jessica Cemlyn-Jones
Published: (2008)
by: Jessica Cemlyn-Jones
Published: (2008)
Physics informed Transformer-VAE for biophysical parameter estimation: PROSAIL model inversion in Sentinel-2 imagery
by: Mensah, Prince, et al.
Published: (2025)
by: Mensah, Prince, et al.
Published: (2025)
Dispelling the Mirage of Progress in Offline MARL through Standardised Baselines and Evaluation
by: Formanek, Claude, et al.
Published: (2024)
by: Formanek, Claude, et al.
Published: (2024)
A Geospatial Approach to Predicting Desert Locust Breeding Grounds in Africa
by: Yusuf, Ibrahim Salihu, et al.
Published: (2024)
by: Yusuf, Ibrahim Salihu, et al.
Published: (2024)
Challenges and opportunities of clinical pharmacy services in Ethiopia: A qualitative study from healthcare practitioners’ perspective
by: Henok G. Tegegn
Published: (2018)
by: Henok G. Tegegn
Published: (2018)
POLICÍA COSTERA DE VIGO. ESTUDIO PILOTO CUASI-EXPERIMENTAL SOBRE RESCATE Y RCP
by: R. Barcala-Furelos
Published: (2017)
by: R. Barcala-Furelos
Published: (2017)
Selective Reincarnation: Offline-to-Online Multi-Agent Reinforcement Learning
by: Formanek, Claude, et al.
Published: (2023)
by: Formanek, Claude, et al.
Published: (2023)
Coordination Failure in Cooperative Offline MARL
by: Tilbury, Callum Rhys, et al.
Published: (2024)
by: Tilbury, Callum Rhys, et al.
Published: (2024)
Putting Data at the Centre of Offline Multi-Agent Reinforcement Learning
by: Formanek, Claude, et al.
Published: (2024)
by: Formanek, Claude, et al.
Published: (2024)
The Sleeping Beauty Problem: Sleeping Kelly is a Thirder
by: Abramowitz, Ben
Published: (2025)
by: Abramowitz, Ben
Published: (2025)
Capital Games and Growth Equilibria
by: Abramowitz, Ben
Published: (2025)
by: Abramowitz, Ben
Published: (2025)
Adjusting to the new Asia. / Morton Abramowitz, Stephen Bosworth
by: Abramowitz, Morton
by: Abramowitz, Morton
Para que la intervención funcione. Mejorar la capacidad de acción de la Organización de las Naciones Unidas / Morton Abramowitz, Thomas Pickering
by: Abramowitz, Morton
by: Abramowitz, Morton
Multimodal CLIP Inference for Meta-Few-Shot Image Classification
by: Ferragu, Constance, et al.
Published: (2024)
by: Ferragu, Constance, et al.
Published: (2024)
AlphaBeta is not as good as you think: a simple class of synthetic games for a better analysis of deterministic game-solving algorithms
by: Boige, Raphaël, et al.
Published: (2025)
by: Boige, Raphaël, et al.
Published: (2025)
La intervención prehospitalaria urgente en el campo de fútbol
by: Roberto J. Barcala Furelos
Published: (2007)
by: Roberto J. Barcala Furelos
Published: (2007)
Forecasting the 2008 presidential election with the time-for-change model / Alan I. Abramowitz
by: Abramowitz, Alan I
Published: (2008)
by: Abramowitz, Alan I
Published: (2008)
InstaGeo: Compute-Efficient Geospatial Machine Learning from Data to Deployment
by: Yusuf, Ibrahim Salihu, et al.
Published: (2025)
by: Yusuf, Ibrahim Salihu, et al.
Published: (2025)
Urban transport in Asia : n operational agenda for the 1990s / Peter Midgley
by: Midgley, Peter
Published: (1994)
by: Midgley, Peter
Published: (1994)
Spin-down of a pulsar with a yielding crust
by: Sob'yanin, Denis Nikolaevich
Published: (2024)
by: Sob'yanin, Denis Nikolaevich
Published: (2024)
Nondipole interaction between two uniformly magnetized spheres and its relation to superconducting levitation
by: Sob'yanin, Denis Nikolaevich
Published: (2024)
by: Sob'yanin, Denis Nikolaevich
Published: (2024)
Perfect nonradiating electromagnetic source and its self-action
by: Sob'yanin, Denis Nikolaevich
Published: (2023)
by: Sob'yanin, Denis Nikolaevich
Published: (2023)
Axiomatic Choice
by: Abramowitz, Ben, et al.
Published: (2025)
by: Abramowitz, Ben, et al.
Published: (2025)
Identifying and Improving Support for Caregivers of Adults with Dementia
by: Amy Abramowitz, et al.
Published: (2025)
by: Amy Abramowitz, et al.
Published: (2025)
Contribution of the 2021 COVID-19 Vaccination Regime to COVID-19 Transmission and Control in South Africa: A Mathematical Modeling Perspective
by: Tegegn, Tesfalem Abate, et al.
Published: (2023)
by: Tegegn, Tesfalem Abate, et al.
Published: (2023)
Similar Items
-
Should we be going MAD? A Look at Multi-Agent Debate Strategies for LLMs
by: Smit, Andries, et al.
Published: (2023) -
Combinatorial Optimization with Policy Adaptation using Latent Space Search
by: Chalumeau, Felix, et al.
Published: (2023) -
Oryx: a Scalable Sequence Model for Many-Agent Coordination in Offline MARL
by: Formanek, Claude, et al.
Published: (2025) -
Generative Model for Small Molecules with Latent Space RL Fine-Tuning to Protein Targets
by: Sob, Ulrich A. Mbou, et al.
Published: (2024) -
SPO: Sequential Monte Carlo Policy Optimisation
by: Macfarlane, Matthew V, et al.
Published: (2024)