CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Hedman, Marcel, Tessera, Kale-ab Abebe, Formanek, Juan Claude, Sims, Anya, Zamboni, Riccardo, McInroe, Trevor, Torr, John, Fosong, Elliot |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Selective Reincarnation: Offline-to-Online Multi-Agent Reinforcement Learning
by: Formanek, Claude, et al.
Published: (2023)
by: Formanek, Claude, et al.
Published: (2023)
Terra Nova: A Comprehensive Challenge Environment for Intelligent Agents
by: McInroe, Trevor
Published: (2025)
by: McInroe, Trevor
Published: (2025)
Remembering the Markov Property in Cooperative MARL
by: Tessera, Kale-ab Abebe, et al.
Published: (2025)
by: Tessera, Kale-ab Abebe, et al.
Published: (2025)
Efficient Offline Reinforcement Learning: First Imitate, then Improve
by: Jelley, Adam, et al.
Published: (2024)
by: Jelley, Adam, et al.
Published: (2024)
Probing Dec-POMDP Reasoning in Cooperative MARL
by: Tessera, Kale-ab, et al.
Published: (2026)
by: Tessera, Kale-ab, et al.
Published: (2026)
Planning to Go Out-of-Distribution in Offline-to-Online Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2023)
by: McInroe, Trevor, et al.
Published: (2023)
PixelBrax: Learning Continuous Control from Pixels End-to-End on the GPU
by: McInroe, Trevor, et al.
Published: (2025)
by: McInroe, Trevor, et al.
Published: (2025)
HyperMARL: Adaptive Hypernetworks for Multi-Agent RL
by: Tessera, Kale-ab Abebe, et al.
Published: (2024)
by: Tessera, Kale-ab Abebe, et al.
Published: (2024)
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2022)
by: McInroe, Trevor, et al.
Published: (2022)
Assistax: A Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics
by: Hinckeldey, Leonard, et al.
Published: (2025)
by: Hinckeldey, Leonard, et al.
Published: (2025)
Agent-Temporal Credit Assignment for Optimal Policy Preservation in Sparse Multi-Agent Reinforcement Learning
by: Kapoor, Aditya, et al.
Published: (2024)
by: Kapoor, Aditya, et al.
Published: (2024)
Fairness over Equality: Correcting Social Incentives in Asymmetric Sequential Social Dilemmas
by: Demir, Alper, et al.
Published: (2026)
by: Demir, Alper, et al.
Published: (2026)
Redistributing Rewards Across Time and Agents for Multi-Agent Reinforcement Learning
by: Kapoor, Aditya, et al.
Published: (2025)
by: Kapoor, Aditya, et al.
Published: (2025)
Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learning
by: Zhang, Weipu, et al.
Published: (2025)
by: Zhang, Weipu, et al.
Published: (2025)
Enhancing Tactile-based Reinforcement Learning for Robotic Control
by: Miller, Elle, et al.
Published: (2025)
by: Miller, Elle, et al.
Published: (2025)
Scalable Multi-Agent Offline Reinforcement Learning and the Role of Information
by: Zamboni, Riccardo, et al.
Published: (2025)
by: Zamboni, Riccardo, et al.
Published: (2025)
How much can change in a year? Revisiting Evaluation in Multi-Agent Reinforcement Learning
by: Singh, Siddarth, et al.
Published: (2023)
by: Singh, Siddarth, et al.
Published: (2023)
Putting Data at the Centre of Offline Multi-Agent Reinforcement Learning
by: Formanek, Claude, et al.
Published: (2024)
by: Formanek, Claude, et al.
Published: (2024)
LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots
by: Han, Dongge, et al.
Published: (2024)
by: Han, Dongge, et al.
Published: (2024)
Coordination Failure in Cooperative Offline MARL
by: Tilbury, Callum Rhys, et al.
Published: (2024)
by: Tilbury, Callum Rhys, et al.
Published: (2024)
Efficiently Quantifying Individual Agent Importance in Cooperative MARL
by: Mahjoub, Omayma, et al.
Published: (2023)
by: Mahjoub, Omayma, et al.
Published: (2023)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
by: Sims, Anya, et al.
Published: (2024)
by: Sims, Anya, et al.
Published: (2024)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
by: Garcin, Samuel, et al.
Published: (2025)
by: Garcin, Samuel, et al.
Published: (2025)
Generalisable Agents for Neural Network Optimisation
by: Tessera, Kale-ab, et al.
Published: (2023)
by: Tessera, Kale-ab, et al.
Published: (2023)
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
by: Zamboni, Riccardo, et al.
Published: (2025)
by: Zamboni, Riccardo, et al.
Published: (2025)
D-CODA: Diffusion for Coordinated Dual-Arm Data Augmentation
by: Liu, I-Chun Arthur, et al.
Published: (2025)
by: Liu, I-Chun Arthur, et al.
Published: (2025)
Global dynamical structures from infinitesimal data
by: McInroe, Benjamin, et al.
Published: (2024)
by: McInroe, Benjamin, et al.
Published: (2024)
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning
by: Sun, Zeyi, et al.
Published: (2025)
by: Sun, Zeyi, et al.
Published: (2025)
roto 2.0: The Robot Tactile Olympiad
by: Miller, Elle, et al.
Published: (2026)
by: Miller, Elle, et al.
Published: (2026)
Forgetting is Everywhere
by: Sanati, Ben, et al.
Published: (2025)
by: Sanati, Ben, et al.
Published: (2025)
A Structural Model of PBH Evaporation and Gamma-Ray Suppression
by: ab_ab
Published: (2026)
by: ab_ab
Published: (2026)
Thickness Structure Hypothesis Hierarchical Unification
by: ab_ab
Published: (2026)
by: ab_ab
Published: (2026)
Thickness Structure Hypothesis Hierarchical Unification
by: ab_ab
Published: (2026)
by: ab_ab
Published: (2026)
Thickness Structure Hypothesis Hierarchical Unification
by: ab_ab
Published: (2026)
by: ab_ab
Published: (2026)
Oryx: a Scalable Sequence Model for Many-Agent Coordination in Offline MARL
by: Formanek, Claude, et al.
Published: (2025)
by: Formanek, Claude, et al.
Published: (2025)
Learning Complex Teamwork Tasks Using a Given Sub-task Decomposition
by: Fosong, Elliot, et al.
Published: (2023)
by: Fosong, Elliot, et al.
Published: (2023)
Dispelling the Mirage of Progress in Offline MARL through Standardised Baselines and Evaluation
by: Formanek, Claude, et al.
Published: (2024)
by: Formanek, Claude, et al.
Published: (2024)
CODA
by: Milagros Rodríguez Cáceres
Published: (2017)
by: Milagros Rodríguez Cáceres
Published: (2017)
Opportunities of Reinforcement Learning in South Africa's Just Transition
by: Formanek, Claude, et al.
Published: (2024)
by: Formanek, Claude, et al.
Published: (2024)
From Parameters to Behaviors: Unsupervised Compression of the Policy Space
by: Tenedini, Davide, et al.
Published: (2025)
by: Tenedini, Davide, et al.
Published: (2025)
Similar Items
-
Selective Reincarnation: Offline-to-Online Multi-Agent Reinforcement Learning
by: Formanek, Claude, et al.
Published: (2023) -
Terra Nova: A Comprehensive Challenge Environment for Intelligent Agents
by: McInroe, Trevor
Published: (2025) -
Remembering the Markov Property in Cooperative MARL
by: Tessera, Kale-ab Abebe, et al.
Published: (2025) -
Efficient Offline Reinforcement Learning: First Imitate, then Improve
by: Jelley, Adam, et al.
Published: (2024) -
Probing Dec-POMDP Reasoning in Cooperative MARL
by: Tessera, Kale-ab, et al.
Published: (2026)