Counterfactual Multi-Agent Policy Gradients
Fuente:
arXiv
Saved in:
| Main Authors: | Foerster, Jakob, Farquhar, Gregory, Afouras, Triantafyllos, Nardelli, Nantas, Whiteson, Shimon |
|---|---|
| Format: | Preprint |
| Published: |
2017
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TAPE: Leveraging Agent Topology for Cooperative Multi-Agent Policy Gradient
by: Lou, Xingzhou, et al.
Published: (2023)
by: Lou, Xingzhou, et al.
Published: (2023)
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
by: Barde, Paul, et al.
Published: (2023)
by: Barde, Paul, et al.
Published: (2023)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023)
by: Rutherford, Alexander, et al.
Published: (2023)
The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind
by: Lupu, Andrei, et al.
Published: (2025)
by: Lupu, Andrei, et al.
Published: (2025)
Understanding Individual Agent Importance in Multi-Agent System via Counterfactual Reasoning
by: Chen, Jianming, et al.
Published: (2024)
by: Chen, Jianming, et al.
Published: (2024)
Counterfactual Effect Decomposition in Multi-Agent Sequential Decision Making
by: Triantafyllou, Stelios, et al.
Published: (2024)
by: Triantafyllou, Stelios, et al.
Published: (2024)
EnergAIze: Multi Agent Deep Deterministic Policy Gradient for Vehicle to Grid Energy Management
by: Fonseca, Tiago, et al.
Published: (2024)
by: Fonseca, Tiago, et al.
Published: (2024)
Counterfactual Reasoning for Causal Responsibility Attribution in Probabilistic Multi-Agent Systems
by: Mu, Chunyan, et al.
Published: (2026)
by: Mu, Chunyan, et al.
Published: (2026)
Adaptive Event-Triggered Policy Gradient for Multi-Agent Reinforcement Learning
by: Siddique, Umer, et al.
Published: (2025)
by: Siddique, Umer, et al.
Published: (2025)
AgentMixer: Multi-Agent Correlated Policy Factorization
by: Li, Zhiyuan, et al.
Published: (2024)
by: Li, Zhiyuan, et al.
Published: (2024)
Multi-Agent Guided Policy Optimization
by: Li, Yueheng, et al.
Published: (2025)
by: Li, Yueheng, et al.
Published: (2025)
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
by: Yang, Shan, et al.
Published: (2026)
by: Yang, Shan, et al.
Published: (2026)
Counterfactual-based Agent Influence Ranker for Agentic AI Workflows
by: Giloni, Amit, et al.
Published: (2025)
by: Giloni, Amit, et al.
Published: (2025)
Measuring Policy Distance for Multi-Agent Reinforcement Learning
by: Hu, Tianyi, et al.
Published: (2024)
by: Hu, Tianyi, et al.
Published: (2024)
Multi-Agent Reinforcement Learning Simulation for Environmental Policy Synthesis
by: Rudd-Jones, James, et al.
Published: (2025)
by: Rudd-Jones, James, et al.
Published: (2025)
SocialGFs: Learning Social Gradient Fields for Multi-Agent Reinforcement Learning
by: Long, Qian, et al.
Published: (2024)
by: Long, Qian, et al.
Published: (2024)
Agent-GSPO: Communication-Efficient Multi-Agent Systems via Group Sequence Policy Optimization
by: Fan, Yijia, et al.
Published: (2025)
by: Fan, Yijia, et al.
Published: (2025)
Influencing LLM Multi-Agent Dialogue via Policy-Parameterized Prompts
by: Bo, Hongbo, et al.
Published: (2026)
by: Bo, Hongbo, et al.
Published: (2026)
GRASP: Gradient Realignment via Active Shared Perception for Multi-Agent Collaborative Optimization
by: Zhou, Sihan, et al.
Published: (2026)
by: Zhou, Sihan, et al.
Published: (2026)
BMW Agents -- A Framework For Task Automation Through Multi-Agent Collaboration
by: Crawford, Noel, et al.
Published: (2024)
by: Crawford, Noel, et al.
Published: (2024)
Safe Equilibrium Policy Optimization for Strategic Agent Policies
by: Arumugam, Karthika, et al.
Published: (2026)
by: Arumugam, Karthika, et al.
Published: (2026)
Hierarchical Policy-Gradient Reinforcement Learning for Multi-Agent Shepherding Control of Non-Cohesive Targets
by: Covone, Stefano, et al.
Published: (2025)
by: Covone, Stefano, et al.
Published: (2025)
Maestro: Learning to Collaborate via Conditional Listwise Policy Optimization for Multi-Agent LLMs
by: Yang, Wei, et al.
Published: (2025)
by: Yang, Wei, et al.
Published: (2025)
The Yokai Learning Environment: Tracking Beliefs Over Space and Time
by: Ruhdorfer, Constantin, et al.
Published: (2025)
by: Ruhdorfer, Constantin, et al.
Published: (2025)
Select to Perfect: Imitating desired behavior from large multi-agent data
by: Franzmeyer, Tim, et al.
Published: (2024)
by: Franzmeyer, Tim, et al.
Published: (2024)
Evolving Many Worlds: Towards Open-Ended Discovery in Petri Dish NCA via Population-Based Training
by: Berdica, Uljad, et al.
Published: (2026)
by: Berdica, Uljad, et al.
Published: (2026)
Budget Allocation Policies for Real-Time Multi-Agent Path Finding
by: Beck, Raz, et al.
Published: (2025)
by: Beck, Raz, et al.
Published: (2025)
Agentic Discovery: Closing the Loop with Cooperative Agents
by: Pauloski, J. Gregory, et al.
Published: (2025)
by: Pauloski, J. Gregory, et al.
Published: (2025)
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System
by: Du, Haikuo, et al.
Published: (2025)
by: Du, Haikuo, et al.
Published: (2025)
Optimized Directed Roadmap Graph for Multi-Agent Path Finding Using Stochastic Gradient Descent
by: Henkel, Christian, et al.
Published: (2020)
by: Henkel, Christian, et al.
Published: (2020)
Inverse Attention Agents for Multi-Agent Systems
by: Long, Qian, et al.
Published: (2024)
by: Long, Qian, et al.
Published: (2024)
Simulation-Based Optimistic Policy Iteration For Multi-Agent MDPs with Kullback-Leibler Control Cost
by: Nakhleh, Khaled, et al.
Published: (2024)
by: Nakhleh, Khaled, et al.
Published: (2024)
Single-Agent Scaling Fails Multi-Agent Intelligence: Towards Foundation Models with Native Multi-Agent Intelligence
by: Hu, Shuyue, et al.
Published: (2025)
by: Hu, Shuyue, et al.
Published: (2025)
Learning to Drive in New Cities Without Human Demonstrations
by: Wang, Zilin, et al.
Published: (2026)
by: Wang, Zilin, et al.
Published: (2026)
MPAC: A Multi-Principal Agent Coordination Protocol for Interoperable Multi-Agent Collaboration
by: Qian, Kaiyang, et al.
Published: (2026)
by: Qian, Kaiyang, et al.
Published: (2026)
AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent
by: Luo, Yinyi, et al.
Published: (2026)
by: Luo, Yinyi, et al.
Published: (2026)
The Heterogeneous Multi-Agent Challenge
by: Dansereau, Charles, et al.
Published: (2025)
by: Dansereau, Charles, et al.
Published: (2025)
Multi-Agent Security Tax: Trading Off Security and Collaboration Capabilities in Multi-Agent Systems
by: Peigne-Lefebvre, Pierre, et al.
Published: (2025)
by: Peigne-Lefebvre, Pierre, et al.
Published: (2025)
BioAgents: Democratizing Bioinformatics Analysis with Multi-Agent Systems
by: Mehandru, Nikita, et al.
Published: (2025)
by: Mehandru, Nikita, et al.
Published: (2025)
Optimizing Agent Collaboration through Heuristic Multi-Agent Planning
by: Soffair, Nitsan
Published: (2023)
by: Soffair, Nitsan
Published: (2023)
Similar Items
-
TAPE: Leveraging Agent Topology for Cooperative Multi-Agent Policy Gradient
by: Lou, Xingzhou, et al.
Published: (2023) -
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
by: Barde, Paul, et al.
Published: (2023) -
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023) -
The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind
by: Lupu, Andrei, et al.
Published: (2025) -
Understanding Individual Agent Importance in Multi-Agent System via Counterfactual Reasoning
by: Chen, Jianming, et al.
Published: (2024)