The Benefits of Power Regularization in Cooperative Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Michelle, Dennis, Michael |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning Bilateral Team Formation in Cooperative Multi-Agent Reinforcement Learning
por: Moslemi, Koorosh, et al.
Publicado: (2025)
por: Moslemi, Koorosh, et al.
Publicado: (2025)
On the Complexity of Learning to Cooperate with Populations of Socially Rational Agents
por: Loftin, Robert, et al.
Publicado: (2024)
por: Loftin, Robert, et al.
Publicado: (2024)
Reinforcement Learning for Hanabi
por: Cohen, Nina, et al.
Publicado: (2025)
por: Cohen, Nina, et al.
Publicado: (2025)
Combining Tree-Search, Generative Models, and Nash Bargaining Concepts in Game-Theoretic Reinforcement Learning
por: Li, Zun, et al.
Publicado: (2023)
por: Li, Zun, et al.
Publicado: (2023)
Emergent Dominance Hierarchies in Reinforcement Learning Agents
por: Rachum, Ram, et al.
Publicado: (2024)
por: Rachum, Ram, et al.
Publicado: (2024)
Welfare and Fairness in Multi-objective Reinforcement Learning
por: Fan, Zimeng, et al.
Publicado: (2022)
por: Fan, Zimeng, et al.
Publicado: (2022)
Inverse Concave-Utility Reinforcement Learning is Inverse Game Theory
por: Çelikok, Mustafa Mert, et al.
Publicado: (2024)
por: Çelikok, Mustafa Mert, et al.
Publicado: (2024)
Cooperative Game-Theoretic Credit Assignment for Multi-Agent Policy Gradients via the Core
por: Ji, Mengda, et al.
Publicado: (2025)
por: Ji, Mengda, et al.
Publicado: (2025)
Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning
por: Kudo, Mikoto, et al.
Publicado: (2026)
por: Kudo, Mikoto, et al.
Publicado: (2026)
Preference-Based Multi-Agent Reinforcement Learning: Data Coverage and Algorithmic Techniques
por: Zhang, Natalia, et al.
Publicado: (2024)
por: Zhang, Natalia, et al.
Publicado: (2024)
A Black-box Approach for Non-stationary Multi-agent Reinforcement Learning
por: Jiang, Haozhe, et al.
Publicado: (2023)
por: Jiang, Haozhe, et al.
Publicado: (2023)
Independent RL for Cooperative-Competitive Agents: A Mean-Field Perspective
por: Zaman, Muhammad Aneeq uz, et al.
Publicado: (2024)
por: Zaman, Muhammad Aneeq uz, et al.
Publicado: (2024)
Adaptive Network Intervention for Complex Systems: A Hierarchical Graph Reinforcement Learning Approach
por: Chen, Qiliang, et al.
Publicado: (2024)
por: Chen, Qiliang, et al.
Publicado: (2024)
AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning
por: Ekpo, Promise, et al.
Publicado: (2025)
por: Ekpo, Promise, et al.
Publicado: (2025)
Policy Optimization finds Nash Equilibrium in Regularized General-Sum LQ Games
por: Zaman, Muhammad Aneeq uz, et al.
Publicado: (2024)
por: Zaman, Muhammad Aneeq uz, et al.
Publicado: (2024)
ELA: Exploited Level Augmentation for Offline Learning in Zero-Sum Games
por: Lei, Shiqi, et al.
Publicado: (2024)
por: Lei, Shiqi, et al.
Publicado: (2024)
Enhancing Cooperation through Selective Interaction and Long-term Experiences in Multi-Agent Reinforcement Learning
por: Ren, Tianyu, et al.
Publicado: (2024)
por: Ren, Tianyu, et al.
Publicado: (2024)
Agent-Temporal Credit Assignment for Optimal Policy Preservation in Sparse Multi-Agent Reinforcement Learning
por: Kapoor, Aditya, et al.
Publicado: (2024)
por: Kapoor, Aditya, et al.
Publicado: (2024)
Nash Learning from Human Feedback
por: Munos, Rémi, et al.
Publicado: (2023)
por: Munos, Rémi, et al.
Publicado: (2023)
Learning to Negotiate via Voluntary Commitment
por: Zhu, Shuhui, et al.
Publicado: (2025)
por: Zhu, Shuhui, et al.
Publicado: (2025)
Learning Mean Field Control on Sparse Graphs
por: Fabian, Christian, et al.
Publicado: (2025)
por: Fabian, Christian, et al.
Publicado: (2025)
Learning Unanimously Acceptable Lotteries via Queries
por: Choo, Davin, et al.
Publicado: (2026)
por: Choo, Davin, et al.
Publicado: (2026)
Separation Assurance between Heterogeneous Fleets of Small Unmanned Aerial Systems via Multi-Agent Reinforcement Learning
por: Sharifi, Iman, et al.
Publicado: (2026)
por: Sharifi, Iman, et al.
Publicado: (2026)
Bounded Rationality Equilibrium Learning in Mean Field Games
por: Eich, Yannick, et al.
Publicado: (2024)
por: Eich, Yannick, et al.
Publicado: (2024)
Learning Optimal Tax Design in Nonatomic Congestion Games
por: Cui, Qiwen, et al.
Publicado: (2024)
por: Cui, Qiwen, et al.
Publicado: (2024)
Online Learning of Counter Categories and Ratings in PvP Games
por: Lin, Chiu-Chou, et al.
Publicado: (2025)
por: Lin, Chiu-Chou, et al.
Publicado: (2025)
Offline Learning of Nash Stable Coalition Structures with Possibly Overlapping Coalitions
por: Cohen, Saar
Publicado: (2026)
por: Cohen, Saar
Publicado: (2026)
Learning Mean Field Games on Sparse Graphs: A Hybrid Graphex Approach
por: Fabian, Christian, et al.
Publicado: (2024)
por: Fabian, Christian, et al.
Publicado: (2024)
A Single Online Agent Can Efficiently Learn Mean Field Games
por: Zhang, Chenyu, et al.
Publicado: (2024)
por: Zhang, Chenyu, et al.
Publicado: (2024)
Stackelberg Learning from Human Feedback: Preference Optimization as a Sequential Game
por: Pásztor, Barna, et al.
Publicado: (2025)
por: Pásztor, Barna, et al.
Publicado: (2025)
Active Evaluation of General Agents: Problem Definition and Comparison of Baseline Algorithms
por: Lanctot, Marc, et al.
Publicado: (2026)
por: Lanctot, Marc, et al.
Publicado: (2026)
Distributed Stackelberg Strategies in State-based Potential Games for Autonomous Decentralized Learning Manufacturing Systems
por: Yuwono, Steve, et al.
Publicado: (2024)
por: Yuwono, Steve, et al.
Publicado: (2024)
Configurable Mirror Descent: Towards a Unification of Decision Making
por: Li, Pengdeng, et al.
Publicado: (2024)
por: Li, Pengdeng, et al.
Publicado: (2024)
Computational Performance of Deep Reinforcement Learning to find Nash Equilibria
por: Graf, Christoph, et al.
Publicado: (2021)
por: Graf, Christoph, et al.
Publicado: (2021)
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
por: Hu, Anran, et al.
Publicado: (2024)
por: Hu, Anran, et al.
Publicado: (2024)
Observation Interference in Partially Observable Assistance Games
por: Emmons, Scott, et al.
Publicado: (2024)
por: Emmons, Scott, et al.
Publicado: (2024)
Offline Fictitious Self-Play for Competitive Games
por: Chen, Jingxiao, et al.
Publicado: (2024)
por: Chen, Jingxiao, et al.
Publicado: (2024)
Strategic Classification With Externalities
por: Hossain, Safwan, et al.
Publicado: (2024)
por: Hossain, Safwan, et al.
Publicado: (2024)
Fusion-PSRO: Nash Policy Fusion for Policy Space Response Oracles
por: Lian, Jiesong, et al.
Publicado: (2024)
por: Lian, Jiesong, et al.
Publicado: (2024)
Analysing the Sample Complexity of Opponent Shaping
por: Fung, Kitty, et al.
Publicado: (2024)
por: Fung, Kitty, et al.
Publicado: (2024)
Ejemplares similares
-
Learning Bilateral Team Formation in Cooperative Multi-Agent Reinforcement Learning
por: Moslemi, Koorosh, et al.
Publicado: (2025) -
On the Complexity of Learning to Cooperate with Populations of Socially Rational Agents
por: Loftin, Robert, et al.
Publicado: (2024) -
Reinforcement Learning for Hanabi
por: Cohen, Nina, et al.
Publicado: (2025) -
Combining Tree-Search, Generative Models, and Nash Bargaining Concepts in Game-Theoretic Reinforcement Learning
por: Li, Zun, et al.
Publicado: (2023) -
Emergent Dominance Hierarchies in Reinforcement Learning Agents
por: Rachum, Ram, et al.
Publicado: (2024)