Efficient Reinforcement Learning for Global Decision Making in the Presence of Local Agents at Scale
Fuente:
arXiv
Saved in:
| Main Authors: | Anand, Emile, Qu, Guannan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling
by: Anand, Emile, et al.
Published: (2026)
by: Anand, Emile, et al.
Published: (2026)
Beyond Monotonicity: Revisiting Factorization Principles in Multi-Agent Q-Learning
by: Hu, Tianmeng, et al.
Published: (2025)
by: Hu, Tianmeng, et al.
Published: (2025)
ParamMem: Augmenting Language Agents with Parametric Reflective Memory
by: Yao, Tianjun, et al.
Published: (2026)
by: Yao, Tianjun, et al.
Published: (2026)
Towards General Negotiation Strategies with End-to-End Reinforcement Learning
by: Renting, Bram M., et al.
Published: (2024)
by: Renting, Bram M., et al.
Published: (2024)
A Principle of Targeted Intervention for Multi-Agent Reinforcement Learning
by: Liu, Anjie, et al.
Published: (2025)
by: Liu, Anjie, et al.
Published: (2025)
Risk-Sensitive Multi-Agent Reinforcement Learning in Network Aggregative Markov Games
by: Ghaemi, Hafez, et al.
Published: (2024)
by: Ghaemi, Hafez, et al.
Published: (2024)
Social Interpretable Reinforcement Learning
by: Custode, Leonardo Lucio, et al.
Published: (2024)
by: Custode, Leonardo Lucio, et al.
Published: (2024)
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
by: Castellini, Jacopo, et al.
Published: (2019)
by: Castellini, Jacopo, et al.
Published: (2019)
Individual Contributions as Intrinsic Exploration Scaffolds for Multi-agent Reinforcement Learning
by: Li, Xinran, et al.
Published: (2024)
by: Li, Xinran, et al.
Published: (2024)
Autonomous Decision Making for UAV Cooperative Pursuit-Evasion Game with Reinforcement Learning
by: Zhao, Yang, et al.
Published: (2024)
by: Zhao, Yang, et al.
Published: (2024)
Networked Agents in the Dark: Team Value Learning under Partial Observability
by: Varela, Guilherme S., et al.
Published: (2025)
by: Varela, Guilherme S., et al.
Published: (2025)
Optimally Solving Simultaneous-Move Dec-POMDPs: The Sequential Central Planning Approach
by: Peralez, Johan, et al.
Published: (2024)
by: Peralez, Johan, et al.
Published: (2024)
Difference Rewards Policy Gradients
by: Castellini, Jacopo, et al.
Published: (2020)
by: Castellini, Jacopo, et al.
Published: (2020)
Dynamic Attentional Context Scoping: Agent-Triggered Focus Sessions for Isolated Per-Agent Steering in Multi-Agent LLM Orchestration
by: Patel, Nickson
Published: (2026)
by: Patel, Nickson
Published: (2026)
Distributed Value Decomposition Networks with Networked Agents
by: Varela, Guilherme S., et al.
Published: (2025)
by: Varela, Guilherme S., et al.
Published: (2025)
Benefits and Limitations of Communication in Multi-Agent Reasoning
by: Rizvi-Martel, Michael, et al.
Published: (2025)
by: Rizvi-Martel, Michael, et al.
Published: (2025)
David vs. Goliath: Verifiable Agent-to-Agent Jailbreaking via Reinforcement Learning
by: Nellessen, Samuel, et al.
Published: (2026)
by: Nellessen, Samuel, et al.
Published: (2026)
Collaboration Promotes Group Resilience in Multi-Agent RL
by: Shraga, Ilai, et al.
Published: (2021)
by: Shraga, Ilai, et al.
Published: (2021)
Representational Collapse in Multi-Agent LLM Committees: Measurement and Diversity-Aware Consensus
by: Patel, Dipkumar
Published: (2026)
by: Patel, Dipkumar
Published: (2026)
Mean-Field Sampling for Cooperative Multi-Agent Reinforcement Learning
by: Anand, Emile, et al.
Published: (2024)
by: Anand, Emile, et al.
Published: (2024)
HieraMAS: Optimizing Intra-Node LLM Mixtures and Inter-Node Topology for Multi-Agent Systems
by: Yao, Tianjun, et al.
Published: (2026)
by: Yao, Tianjun, et al.
Published: (2026)
Centrally Coordinated Multi-Agent Reinforcement Learning for Power Grid Topology Control
by: de Mol, Barbera, et al.
Published: (2025)
by: de Mol, Barbera, et al.
Published: (2025)
On Convex Optimal Value Functions For POSGs
by: Cunha, Rafael F., et al.
Published: (2023)
by: Cunha, Rafael F., et al.
Published: (2023)
Policy Search, Retrieval, and Composition via Task Similarity in Collaborative Agentic Systems
by: Nath, Saptarshi, et al.
Published: (2025)
by: Nath, Saptarshi, et al.
Published: (2025)
A Systematic Study of Multi-Agent Deep Reinforcement Learning for Safe and Robust Autonomous Highway Ramp Entry
by: Schester, Larry, et al.
Published: (2024)
by: Schester, Larry, et al.
Published: (2024)
Analyzing Closed-loop Training Techniques for Realistic Traffic Agent Models in Autonomous Highway Driving Simulations
by: Bitzer, Matthias, et al.
Published: (2024)
by: Bitzer, Matthias, et al.
Published: (2024)
Characterizing MARL for Energy Control: A Multi-KPI Benchmark on the CityLearn Environment
by: Khouja, Aymen, et al.
Published: (2026)
by: Khouja, Aymen, et al.
Published: (2026)
Learning to Communicate Across Modalities: Perceptual Heterogeneity in Multi-Agent Systems
by: Pitzer, Naomi, et al.
Published: (2026)
by: Pitzer, Naomi, et al.
Published: (2026)
When Agents Disagree: The Selection Bottleneck in Multi-Agent LLM Pipelines
by: Maryanskyy, Artem
Published: (2026)
by: Maryanskyy, Artem
Published: (2026)
On the Fundamental Limitations of Decentralized Learnable Reward Shaping in Cooperative Multi-Agent Reinforcement Learning
by: Akella, Aditya
Published: (2025)
by: Akella, Aditya
Published: (2025)
AgentLTV: An Agent-Based Unified Search-and-Evolution Framework for Automated Lifetime Value Prediction
by: Wu, Chaowei, et al.
Published: (2026)
by: Wu, Chaowei, et al.
Published: (2026)
Mean-Field Reinforcement Learning without Synchrony
by: Yang, Shan
Published: (2026)
by: Yang, Shan
Published: (2026)
DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows
by: Gao, Yuxuan, et al.
Published: (2026)
by: Gao, Yuxuan, et al.
Published: (2026)
Feel-Good Thompson Sampling for Contextual Bandits: a Markov Chain Monte Carlo Showdown
by: Anand, Emile, et al.
Published: (2025)
by: Anand, Emile, et al.
Published: (2025)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
by: Zeng, Jack, et al.
Published: (2025)
by: Zeng, Jack, et al.
Published: (2025)
SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use
by: Patel, Hitesh Laxmichand, et al.
Published: (2025)
by: Patel, Hitesh Laxmichand, et al.
Published: (2025)
QTypeMix: Enhancing Multi-Agent Cooperative Strategies through Heterogeneous and Homogeneous Value Decomposition
by: Fu, Songchen, et al.
Published: (2024)
by: Fu, Songchen, et al.
Published: (2024)
Multi-Agent Decision-Focused Learning via Value-Aware Sequential Communication
by: Amoh, Benjamin, et al.
Published: (2026)
by: Amoh, Benjamin, et al.
Published: (2026)
Discovering Antagonists in Networks of Systems: Robot Deployment
by: Wenger, Ingeborg, et al.
Published: (2025)
by: Wenger, Ingeborg, et al.
Published: (2025)
A Knowledge-Based Language Model: Deducing Grammatical Knowledge in a Multi-Agent Language Acquisition Simulation
by: Shakouri, David Ph., et al.
Published: (2025)
by: Shakouri, David Ph., et al.
Published: (2025)
Similar Items
-
Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling
by: Anand, Emile, et al.
Published: (2026) -
Beyond Monotonicity: Revisiting Factorization Principles in Multi-Agent Q-Learning
by: Hu, Tianmeng, et al.
Published: (2025) -
ParamMem: Augmenting Language Agents with Parametric Reflective Memory
by: Yao, Tianjun, et al.
Published: (2026) -
Towards General Negotiation Strategies with End-to-End Reinforcement Learning
by: Renting, Bram M., et al.
Published: (2024) -
A Principle of Targeted Intervention for Multi-Agent Reinforcement Learning
by: Liu, Anjie, et al.
Published: (2025)