Dispelling the Mirage of Progress in Offline MARL through Standardised Baselines and Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Formanek, Claude, Tilbury, Callum Rhys, Beyers, Louise, Shock, Jonathan, Pretorius, Arnu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Coordination Failure in Cooperative Offline MARL
von: Tilbury, Callum Rhys, et al.
Veröffentlicht: (2024)
von: Tilbury, Callum Rhys, et al.
Veröffentlicht: (2024)
Putting Data at the Centre of Offline Multi-Agent Reinforcement Learning
von: Formanek, Claude, et al.
Veröffentlicht: (2024)
von: Formanek, Claude, et al.
Veröffentlicht: (2024)
Selective Reincarnation: Offline-to-Online Multi-Agent Reinforcement Learning
von: Formanek, Claude, et al.
Veröffentlicht: (2023)
von: Formanek, Claude, et al.
Veröffentlicht: (2023)
Opportunities of Reinforcement Learning in South Africa's Just Transition
von: Formanek, Claude, et al.
Veröffentlicht: (2024)
von: Formanek, Claude, et al.
Veröffentlicht: (2024)
Sable: a Performant, Efficient and Scalable Sequence Model for MARL
von: Mahjoub, Omayma, et al.
Veröffentlicht: (2024)
von: Mahjoub, Omayma, et al.
Veröffentlicht: (2024)
Generalisable Agents for Neural Network Optimisation
von: Tessera, Kale-ab, et al.
Veröffentlicht: (2023)
von: Tessera, Kale-ab, et al.
Veröffentlicht: (2023)
Multi-Agent Reinforcement Learning with Selective State-Space Models
von: Daniel, Jemma, et al.
Veröffentlicht: (2024)
von: Daniel, Jemma, et al.
Veröffentlicht: (2024)
Oryx: a Scalable Sequence Model for Many-Agent Coordination in Offline MARL
von: Formanek, Claude, et al.
Veröffentlicht: (2025)
von: Formanek, Claude, et al.
Veröffentlicht: (2025)
Learning Partial Action Replacement in Offline MARL
von: Jin, Yue, et al.
Veröffentlicht: (2026)
von: Jin, Yue, et al.
Veröffentlicht: (2026)
Memory-Enhanced Neural Solvers for Routing Problems
von: Chalumeau, Felix, et al.
Veröffentlicht: (2024)
von: Chalumeau, Felix, et al.
Veröffentlicht: (2024)
Dispelling the Curse of Singularities in Neural Network Optimizations
von: Cao, Hengjie, et al.
Veröffentlicht: (2026)
von: Cao, Hengjie, et al.
Veröffentlicht: (2026)
Combinatorial Optimization with Policy Adaptation using Latent Space Search
von: Chalumeau, Felix, et al.
Veröffentlicht: (2023)
von: Chalumeau, Felix, et al.
Veröffentlicht: (2023)
Partial Action Replacement: Tackling Distribution Shift in Offline MARL
von: Jin, Yue, et al.
Veröffentlicht: (2025)
von: Jin, Yue, et al.
Veröffentlicht: (2025)
Evaluating Evaluation Metrics -- The Mirage of Hallucination Detection
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
Characterizing MARL for Energy Control: A Multi-KPI Benchmark on the CityLearn Environment
von: Khouja, Aymen, et al.
Veröffentlicht: (2026)
von: Khouja, Aymen, et al.
Veröffentlicht: (2026)
Breaking the Performance Ceiling in Reinforcement Learning requires Inference Strategies
von: Chalumeau, Felix, et al.
Veröffentlicht: (2025)
von: Chalumeau, Felix, et al.
Veröffentlicht: (2025)
Beyond Shallow Behavior: Task-Efficient Value-Based Multi-Task Offline MARL via Skill Discovery
von: Wang, Xun, et al.
Veröffentlicht: (2025)
von: Wang, Xun, et al.
Veröffentlicht: (2025)
On the Mirage of Long-Range Dependency, with an Application to Integer Multiplication
von: Wei, Zichao
Veröffentlicht: (2026)
von: Wei, Zichao
Veröffentlicht: (2026)
Mirage: Model-Agnostic Graph Distillation for Graph Classification
von: Gupta, Mridul, et al.
Veröffentlicht: (2023)
von: Gupta, Mridul, et al.
Veröffentlicht: (2023)
A Unified Framework for Locality in Scalable MARL
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Modular Jets for Supervised Pipelines: Diagnosing Mirage vs Identifiability
von: Sanyal, Suman
Veröffentlicht: (2025)
von: Sanyal, Suman
Veröffentlicht: (2025)
CAMA: Exploring Collusive Adversarial Attacks in c-MARL
von: Niu, Men, et al.
Veröffentlicht: (2026)
von: Niu, Men, et al.
Veröffentlicht: (2026)
TIGER-MARL: Enhancing Multi-Agent Reinforcement Learning with Temporal Information through Graph-based Embeddings and Representations
von: Gupta, Nikunj, et al.
Veröffentlicht: (2025)
von: Gupta, Nikunj, et al.
Veröffentlicht: (2025)
Dataset Distillation for Offline Reinforcement Learning
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
Towards a Theory of AI Personhood
von: Ward, Francis Rhys
Veröffentlicht: (2025)
von: Ward, Francis Rhys
Veröffentlicht: (2025)
Efficiently Quantifying Individual Agent Importance in Cooperative MARL
von: Mahjoub, Omayma, et al.
Veröffentlicht: (2023)
von: Mahjoub, Omayma, et al.
Veröffentlicht: (2023)
The Mirage of Artificial Intelligence Terms of Use Restrictions
von: Henderson, Peter, et al.
Veröffentlicht: (2024)
von: Henderson, Peter, et al.
Veröffentlicht: (2024)
Mirage: A Multi-Level Superoptimizer for Tensor Programs
von: Wu, Mengdi, et al.
Veröffentlicht: (2024)
von: Wu, Mengdi, et al.
Veröffentlicht: (2024)
Mirage: An RNS-Based Photonic Accelerator for DNN Training
von: Demirkiran, Cansu, et al.
Veröffentlicht: (2023)
von: Demirkiran, Cansu, et al.
Veröffentlicht: (2023)
Continuous-Time Analysis of Adaptive Optimization and Normalization
von: Gould, Rhys, et al.
Veröffentlicht: (2024)
von: Gould, Rhys, et al.
Veröffentlicht: (2024)
The Mean is the Mirage: Entropy-Adaptive Model Merging under Heterogeneous Domain Shifts in Medical Imaging
von: Ambekar, Sameer, et al.
Veröffentlicht: (2026)
von: Ambekar, Sameer, et al.
Veröffentlicht: (2026)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
von: Rutherford, Alexander, et al.
Veröffentlicht: (2023)
von: Rutherford, Alexander, et al.
Veröffentlicht: (2023)
Investigating Relational State Abstraction in Collaborative MARL
von: Utke, Sharlin, et al.
Veröffentlicht: (2024)
von: Utke, Sharlin, et al.
Veröffentlicht: (2024)
Mildly Conservative Regularized Evaluation for Offline Reinforcement Learning
von: Chen, Haohui, et al.
Veröffentlicht: (2025)
von: Chen, Haohui, et al.
Veröffentlicht: (2025)
Evaluation-Time Policy Switching for Offline Reinforcement Learning
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
Unveiling Causal Reasoning in Large Language Models: Reality or Mirage?
von: Chi, Haoang, et al.
Veröffentlicht: (2025)
von: Chi, Haoang, et al.
Veröffentlicht: (2025)
Offline Policy Evaluation for Reinforcement Learning with Adaptively Collected Data
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
HyperMARL: Adaptive Hypernetworks for Multi-Agent RL
von: Tessera, Kale-ab Abebe, et al.
Veröffentlicht: (2024)
von: Tessera, Kale-ab Abebe, et al.
Veröffentlicht: (2024)
BenchMARL: Benchmarking Multi-Agent Reinforcement Learning
von: Bettini, Matteo, et al.
Veröffentlicht: (2023)
von: Bettini, Matteo, et al.
Veröffentlicht: (2023)
Offline Trajectory Optimization for Offline Reinforcement Learning
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Coordination Failure in Cooperative Offline MARL
von: Tilbury, Callum Rhys, et al.
Veröffentlicht: (2024) -
Putting Data at the Centre of Offline Multi-Agent Reinforcement Learning
von: Formanek, Claude, et al.
Veröffentlicht: (2024) -
Selective Reincarnation: Offline-to-Online Multi-Agent Reinforcement Learning
von: Formanek, Claude, et al.
Veröffentlicht: (2023) -
Opportunities of Reinforcement Learning in South Africa's Just Transition
von: Formanek, Claude, et al.
Veröffentlicht: (2024) -
Sable: a Performant, Efficient and Scalable Sequence Model for MARL
von: Mahjoub, Omayma, et al.
Veröffentlicht: (2024)