High entropy leads to symmetry equivariant policies in Dec-POMDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Forkel, Johannes, Ruhdorfer, Constantin, Beukman, Michael, Bulling, Andreas, Foerster, Jakob |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Yokai Learning Environment: Tracking Beliefs Over Space and Time
by: Ruhdorfer, Constantin, et al.
Published: (2025)
by: Ruhdorfer, Constantin, et al.
Published: (2025)
The Overcooked Generalisation Challenge: Evaluating Cooperation with Novel Partners in Unknown Environments Using Unsupervised Environment Design
by: Ruhdorfer, Constantin, et al.
Published: (2024)
by: Ruhdorfer, Constantin, et al.
Published: (2024)
Unsupervised Partner Design Enables Robust Ad-hoc Teamwork
by: Ruhdorfer, Constantin, et al.
Published: (2025)
by: Ruhdorfer, Constantin, et al.
Published: (2025)
Solving Collaborative Dec-POMDPs with Deep Reinforcement Learning Heuristics
by: Soffair, Nitsan
Published: (2022)
by: Soffair, Nitsan
Published: (2022)
Expected Return Symmetries
by: Muglich, Darius, et al.
Published: (2025)
by: Muglich, Darius, et al.
Published: (2025)
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
by: Omari, Bassel Al, et al.
Published: (2025)
by: Omari, Bassel Al, et al.
Published: (2025)
Risk-seeking conservative policy iteration with agent-state based policies for Dec-POMDPs with guaranteed convergence
by: Sinha, Amit, et al.
Published: (2026)
by: Sinha, Amit, et al.
Published: (2026)
Optimally Solving Simultaneous-Move Dec-POMDPs: The Sequential Central Planning Approach
by: Peralez, Johan, et al.
Published: (2024)
by: Peralez, Johan, et al.
Published: (2024)
Probing Dec-POMDP Reasoning in Cooperative MARL
by: Tessera, Kale-ab, et al.
Published: (2026)
by: Tessera, Kale-ab, et al.
Published: (2026)
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
by: Barde, Paul, et al.
Published: (2023)
by: Barde, Paul, et al.
Published: (2023)
Select to Perfect: Imitating desired behavior from large multi-agent data
by: Franzmeyer, Tim, et al.
Published: (2024)
by: Franzmeyer, Tim, et al.
Published: (2024)
Learning to Drive in New Cities Without Human Demonstrations
by: Wang, Zilin, et al.
Published: (2026)
by: Wang, Zilin, et al.
Published: (2026)
Federated Learning for Data-Driven Feedforward Control: A Case Study on Vehicle Lateral Dynamics
by: Weber, Jakob, et al.
Published: (2025)
by: Weber, Jakob, et al.
Published: (2025)
GRSN: Gated Recurrent Spiking Neurons for POMDPs and MARL
by: Qin, Lang, et al.
Published: (2024)
by: Qin, Lang, et al.
Published: (2024)
On-policy Actor-Critic Reinforcement Learning for Multi-UAV Exploration
by: Farid, Ali Moltajaei, et al.
Published: (2024)
by: Farid, Ali Moltajaei, et al.
Published: (2024)
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments
by: Zhaikhan, Ainur, et al.
Published: (2024)
by: Zhaikhan, Ainur, et al.
Published: (2024)
Learning safe, constrained policies via imitation learning: Connection to Probabilistic Inference and a Naive Algorithm
by: Papadopoulos, George, et al.
Published: (2025)
by: Papadopoulos, George, et al.
Published: (2025)
Towards Optimal Performance and Action Consistency Guarantees in Dec-POMDPs with Inconsistent Beliefs and Limited Communication
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
Generative adversarial imitation learning for robot swarms: Learning from human demonstrations and trained policies
by: Kraus, Mattes, et al.
Published: (2026)
by: Kraus, Mattes, et al.
Published: (2026)
Analysing the Sample Complexity of Opponent Shaping
by: Fung, Kitty, et al.
Published: (2024)
by: Fung, Kitty, et al.
Published: (2024)
LEED: A Highly Efficient and Scalable LLM-Empowered Expert Demonstrations Framework for Multi-Agent Reinforcement Learning
by: Duan, Tianyang, et al.
Published: (2025)
by: Duan, Tianyang, et al.
Published: (2025)
GRAND: Guidance, Rebalancing, and Assignment for Networked Dispatch in Multi-Agent Path Finding
by: Gaber, Johannes, et al.
Published: (2025)
by: Gaber, Johannes, et al.
Published: (2025)
Fairness Aware Reinforcement Learning via Proximal Policy Optimization
by: La Malfa, Gabriele, et al.
Published: (2025)
by: La Malfa, Gabriele, et al.
Published: (2025)
Using a single actor to output personalized policy for different intersections
by: Zhou, Kailing, et al.
Published: (2025)
by: Zhou, Kailing, et al.
Published: (2025)
Neural Network-Based Parameter Estimation of a Labour Market Agent-Based Model
by: Alves, M Lopes, et al.
Published: (2026)
by: Alves, M Lopes, et al.
Published: (2026)
Combat Urban Congestion via Collaboration: Heterogeneous GNN-based MARL for Coordinated Platooning and Traffic Signal Control
by: Peng, Xianyue, et al.
Published: (2023)
by: Peng, Xianyue, et al.
Published: (2023)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023)
by: Rutherford, Alexander, et al.
Published: (2023)
Soft Condorcet Optimization for Ranking of General Agents
by: Lanctot, Marc, et al.
Published: (2024)
by: Lanctot, Marc, et al.
Published: (2024)
Distributed Area Coverage with High Altitude Balloons Using Multi-Agent Reinforcement Learning
by: Haroon, Adam, et al.
Published: (2025)
by: Haroon, Adam, et al.
Published: (2025)
A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations
by: Flavin, Timothy, et al.
Published: (2026)
by: Flavin, Timothy, et al.
Published: (2026)
Enhancing Cooperative Multi-Agent Reinforcement Learning with State Modelling and Adversarial Exploration
by: Kontogiannis, Andreas, et al.
Published: (2025)
by: Kontogiannis, Andreas, et al.
Published: (2025)
Ad-Hoc Human-AI Coordination Challenge
by: Dizdarević, Tin, et al.
Published: (2025)
by: Dizdarević, Tin, et al.
Published: (2025)
Flexible Swarm Learning May Outpace Foundation Models in Essential Tasks
by: Samadi, Moein E., et al.
Published: (2025)
by: Samadi, Moein E., et al.
Published: (2025)
Mixed Traffic Control and Coordination from Pixels
by: Villarreal, Michael, et al.
Published: (2023)
by: Villarreal, Michael, et al.
Published: (2023)
Continuous-Time Value Iteration for Multi-Agent Reinforcement Learning
by: Wang, Xuefeng, et al.
Published: (2025)
by: Wang, Xuefeng, et al.
Published: (2025)
Sustainable Smart Farm Networks: Enhancing Resilience and Efficiency with Decision Theory-Guided Deep Reinforcement Learning
by: Chen, Dian, et al.
Published: (2025)
by: Chen, Dian, et al.
Published: (2025)
A Large Language Model for Feasible and Diverse Population Synthesis
by: Lim, Sung Yoo, et al.
Published: (2025)
by: Lim, Sung Yoo, et al.
Published: (2025)
Learning Graph Representation of Agent Diffusers
by: Djenouri, Youcef, et al.
Published: (2025)
by: Djenouri, Youcef, et al.
Published: (2025)
Large-Scale Mixed-Traffic and Intersection Control using Multi-agent Reinforcement Learning
by: Liu, Songyang, et al.
Published: (2025)
by: Liu, Songyang, et al.
Published: (2025)
Integrated Noise and Safety Management in UAM via A Unified Reinforcement Learning Framework
by: Murthy, Surya, et al.
Published: (2025)
by: Murthy, Surya, et al.
Published: (2025)
Similar Items
-
The Yokai Learning Environment: Tracking Beliefs Over Space and Time
by: Ruhdorfer, Constantin, et al.
Published: (2025) -
The Overcooked Generalisation Challenge: Evaluating Cooperation with Novel Partners in Unknown Environments Using Unsupervised Environment Design
by: Ruhdorfer, Constantin, et al.
Published: (2024) -
Unsupervised Partner Design Enables Robust Ad-hoc Teamwork
by: Ruhdorfer, Constantin, et al.
Published: (2025) -
Solving Collaborative Dec-POMDPs with Deep Reinforcement Learning Heuristics
by: Soffair, Nitsan
Published: (2022) -
Expected Return Symmetries
by: Muglich, Darius, et al.
Published: (2025)