Policy-Conditioned Policies for Multi-Agent Task Solving
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Yue, Zhu, Shuhui, Li, Wenhao, Li, Ang, Qiao, Dan, Poupart, Pascal, Zha, Hongyuan, Wang, Baoxiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Information Bargaining: Bilateral Commitment in Bayesian Persuasion
by: Lin, Yue, et al.
Published: (2025)
by: Lin, Yue, et al.
Published: (2025)
Talk, Judge, Cooperate: Gossip-Driven Indirect Reciprocity in Self-Interested LLM Agents
by: Zhu, Shuhui, et al.
Published: (2026)
by: Zhu, Shuhui, et al.
Published: (2026)
Learning to Negotiate via Voluntary Commitment
by: Zhu, Shuhui, et al.
Published: (2025)
by: Zhu, Shuhui, et al.
Published: (2025)
Verbalized Bayesian Persuasion
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
The Reciprocity Gradient
by: Lin, Yue, et al.
Published: (2026)
by: Lin, Yue, et al.
Published: (2026)
PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers
by: Li, Boning, et al.
Published: (2026)
by: Li, Boning, et al.
Published: (2026)
Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
by: Hennes, Daniel, et al.
Published: (2026)
by: Hennes, Daniel, et al.
Published: (2026)
Which Coauthor Should I Nominate in My 99 ICLR Submissions? A Mathematical Analysis of the ICLR 2026 Reciprocal Reviewer Nomination Policy
by: Song, Zhao, et al.
Published: (2025)
by: Song, Zhao, et al.
Published: (2025)
A Policy-Gradient Approach to Solving Imperfect-Information Games with Best-Iterate Convergence
by: Liu, Mingyang, et al.
Published: (2024)
by: Liu, Mingyang, et al.
Published: (2024)
A Task-Driven Multi-UAV Coalition Formation Mechanism
by: Lu, Xinpeng, et al.
Published: (2024)
by: Lu, Xinpeng, et al.
Published: (2024)
NePPO: Near-Potential Policy Optimization for General-Sum Multi-Agent Reinforcement Learning
by: Kalanther, Addison, et al.
Published: (2026)
by: Kalanther, Addison, et al.
Published: (2026)
Policy Abstraction and Nash Refinement in Tree-Exploiting PSRO
by: Konicki, Christine, et al.
Published: (2025)
by: Konicki, Christine, et al.
Published: (2025)
Strategic Resource Selection with Homophilic Agents
by: Harder, Jonathan Gadea, et al.
Published: (2023)
by: Harder, Jonathan Gadea, et al.
Published: (2023)
Cooperative Game-Theoretic Credit Assignment for Multi-Agent Policy Gradients via the Core
by: Ji, Mengda, et al.
Published: (2025)
by: Ji, Mengda, et al.
Published: (2025)
Bi-LSTM based Multi-Agent DRL with Computation-aware Pruning for Agent Twins Migration in Vehicular Embodied AI Networks
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
Tiny Multi-Agent DRL for Twins Migration in UAV Metaverses: A Multi-Leader Multi-Follower Stackelberg Game Approach
by: Kang, Jiawen, et al.
Published: (2024)
by: Kang, Jiawen, et al.
Published: (2024)
Solving Urban Network Security Games: Learning Platform, Benchmark, and Challenge for AI Research
by: Zhuang, Shuxin, et al.
Published: (2025)
by: Zhuang, Shuxin, et al.
Published: (2025)
A Computable Game-Theoretic Framework for Multi-Agent Theory of Mind
by: Zhu, Fengming, et al.
Published: (2025)
by: Zhu, Fengming, et al.
Published: (2025)
Policy Aggregation
by: Alamdari, Parand A., et al.
Published: (2024)
by: Alamdari, Parand A., et al.
Published: (2024)
On Corrigibility and Alignment in Multi Agent Games
by: Dable-Heath, Edmund, et al.
Published: (2025)
by: Dable-Heath, Edmund, et al.
Published: (2025)
Ev-Trust: An Evolutionarily Stable Trust Mechanism for Decentralized LLM-Based Multi-Agent Service Economies
by: Wang, Jiye, et al.
Published: (2025)
by: Wang, Jiye, et al.
Published: (2025)
Agent-Temporal Credit Assignment for Optimal Policy Preservation in Sparse Multi-Agent Reinforcement Learning
by: Kapoor, Aditya, et al.
Published: (2024)
by: Kapoor, Aditya, et al.
Published: (2024)
Automated Approach for Solving Infinite-state Polynomial Reachability Games
by: Chatterjee, Krishnendu, et al.
Published: (2026)
by: Chatterjee, Krishnendu, et al.
Published: (2026)
Bi-Level Policy Optimization with Nyström Hypergradients
by: Prakash, Arjun, et al.
Published: (2025)
by: Prakash, Arjun, et al.
Published: (2025)
Policy Space Response Oracles: A Survey
by: Bighashdel, Ariyan, et al.
Published: (2024)
by: Bighashdel, Ariyan, et al.
Published: (2024)
Competitive Multi-armed Bandit Games for Resource Sharing
by: Li, Hongbo, et al.
Published: (2025)
by: Li, Hongbo, et al.
Published: (2025)
Hypergame Rationalisability: Solving Agent Misalignment In Strategic Play
by: Trencsenyi, Vince
Published: (2025)
by: Trencsenyi, Vince
Published: (2025)
DIML: Differentiable Inverse Mechanism Learning from Behaviors of Multi-Agent Learning Trajectories
by: An, Zhiyu, et al.
Published: (2026)
by: An, Zhiyu, et al.
Published: (2026)
Puzzle it Out: Local-to-Global World Model for Offline Multi-Agent Reinforcement Learning
by: Li, Sijia, et al.
Published: (2026)
by: Li, Sijia, et al.
Published: (2026)
The Price of Paranoia: Robust Risk-Sensitive Cooperation in Non-Stationary Multi-Agent Reinforcement Learning
by: Ganguly, Deep Kumar, et al.
Published: (2026)
by: Ganguly, Deep Kumar, et al.
Published: (2026)
Cooperation Dynamics in Multi-Agent Systems: Exploring Game-Theoretic Scenarios with Mean-Field Equilibria
by: Sathi, Vaigarai, et al.
Published: (2023)
by: Sathi, Vaigarai, et al.
Published: (2023)
The Bakers and Millers Game with Restricted Locations
by: Krogmann, Simon, et al.
Published: (2025)
by: Krogmann, Simon, et al.
Published: (2025)
Strategic Facility Location with Clients that Minimize Total Waiting Time
by: Krogmann, Simon, et al.
Published: (2022)
by: Krogmann, Simon, et al.
Published: (2022)
Optimal Welfare in Noncooperative Network Formation under Attack
by: Doubez, Natan, et al.
Published: (2025)
by: Doubez, Natan, et al.
Published: (2025)
Institutional AI: Governing LLM Collusion in Multi-Agent Cournot Markets via Public Governance Graphs
by: Syrnikov, Marcantonio Bracale, et al.
Published: (2026)
by: Syrnikov, Marcantonio Bracale, et al.
Published: (2026)
Off-Policy Evaluation for Sequential Persuasion Process with Unobserved Confounding
by: S., Nishanth Venkatesh, et al.
Published: (2025)
by: S., Nishanth Venkatesh, et al.
Published: (2025)
Two-Stage Facility Location Games with Strategic Clients and Facilities
by: Krogmann, Simon, et al.
Published: (2021)
by: Krogmann, Simon, et al.
Published: (2021)
Scalable Mechanism Design for Multi-Agent Path Finding
by: Friedrich, Paul, et al.
Published: (2024)
by: Friedrich, Paul, et al.
Published: (2024)
Multi-Sender Persuasion: A Computational Perspective
by: Hossain, Safwan, et al.
Published: (2024)
by: Hossain, Safwan, et al.
Published: (2024)
Fusion-PSRO: Nash Policy Fusion for Policy Space Response Oracles
by: Lian, Jiesong, et al.
Published: (2024)
by: Lian, Jiesong, et al.
Published: (2024)
Similar Items
-
Information Bargaining: Bilateral Commitment in Bayesian Persuasion
by: Lin, Yue, et al.
Published: (2025) -
Talk, Judge, Cooperate: Gossip-Driven Indirect Reciprocity in Self-Interested LLM Agents
by: Zhu, Shuhui, et al.
Published: (2026) -
Learning to Negotiate via Voluntary Commitment
by: Zhu, Shuhui, et al.
Published: (2025) -
Verbalized Bayesian Persuasion
by: Li, Wenhao, et al.
Published: (2025) -
The Reciprocity Gradient
by: Lin, Yue, et al.
Published: (2026)