Maestro: Learning to Collaborate via Conditional Listwise Policy Optimization for Multi-Agent LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Wei, Pang, Jiacheng, Li, Shixuan, Bogdan, Paul, Tu, Stephen, Thomason, Jesse |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Deliberate: Meta-policy Collaboration for Agentic LLMs with Multi-agent Reinforcement Learning
by: Yang, Wei, et al.
Published: (2025)
by: Yang, Wei, et al.
Published: (2025)
Evaluating Creativity and Deception in Large Language Models: A Simulation Framework for Multi-Agent Balderdash
by: Hejabi, Parsa, et al.
Published: (2024)
by: Hejabi, Parsa, et al.
Published: (2024)
Collaborative Belief Reasoning with LLMs for Efficient Multi-Agent Collaboration
by: Wang, Zhimin, et al.
Published: (2025)
by: Wang, Zhimin, et al.
Published: (2025)
Agent-GSPO: Communication-Efficient Multi-Agent Systems via Group Sequence Policy Optimization
by: Fan, Yijia, et al.
Published: (2025)
by: Fan, Yijia, et al.
Published: (2025)
Multi-Agent Guided Policy Optimization
by: Li, Yueheng, et al.
Published: (2025)
by: Li, Yueheng, et al.
Published: (2025)
RobotFleet: An Open-Source Framework for Centralized Multi-Robot Task Planning
by: Gupta, Rohan, et al.
Published: (2025)
by: Gupta, Rohan, et al.
Published: (2025)
Optimizing Agent Collaboration through Heuristic Multi-Agent Planning
by: Soffair, Nitsan
Published: (2023)
by: Soffair, Nitsan
Published: (2023)
Weak-Link Optimization for Multi-Agent Reasoning and Collaboration
by: Bian, Haoyu, et al.
Published: (2026)
by: Bian, Haoyu, et al.
Published: (2026)
GRASP: Gradient Realignment via Active Shared Perception for Multi-Agent Collaborative Optimization
by: Zhou, Sihan, et al.
Published: (2026)
by: Zhou, Sihan, et al.
Published: (2026)
TwoStep: Multi-agent Task Planning using Classical Planners and Large Language Models
by: Bai, David, et al.
Published: (2024)
by: Bai, David, et al.
Published: (2024)
Measuring Policy Distance for Multi-Agent Reinforcement Learning
by: Hu, Tianyi, et al.
Published: (2024)
by: Hu, Tianyi, et al.
Published: (2024)
An Empirical Study of Multi-Agent Collaboration for Automated Research
by: Shen, Yang, et al.
Published: (2026)
by: Shen, Yang, et al.
Published: (2026)
Safe Equilibrium Policy Optimization for Strategic Agent Policies
by: Arumugam, Karthika, et al.
Published: (2026)
by: Arumugam, Karthika, et al.
Published: (2026)
CCL: Collaborative Curriculum Learning for Sparse-Reward Multi-Agent Reinforcement Learning via Co-evolutionary Task Evolution
by: Lin, Yufei, et al.
Published: (2025)
by: Lin, Yufei, et al.
Published: (2025)
Multi-Agent Reinforcement Learning Simulation for Environmental Policy Synthesis
by: Rudd-Jones, James, et al.
Published: (2025)
by: Rudd-Jones, James, et al.
Published: (2025)
Multi-Agent Collaboration via Evolving Orchestration
by: Dang, Yufan, et al.
Published: (2025)
by: Dang, Yufan, et al.
Published: (2025)
AgentCDM: Enhancing Multi-Agent Collaborative Decision-Making via ACH-Inspired Structured Reasoning
by: Zhao, Xuyang, et al.
Published: (2025)
by: Zhao, Xuyang, et al.
Published: (2025)
Influencing LLM Multi-Agent Dialogue via Policy-Parameterized Prompts
by: Bo, Hongbo, et al.
Published: (2026)
by: Bo, Hongbo, et al.
Published: (2026)
AgentMixer: Multi-Agent Correlated Policy Factorization
by: Li, Zhiyuan, et al.
Published: (2024)
by: Li, Zhiyuan, et al.
Published: (2024)
MPAC: A Multi-Principal Agent Coordination Protocol for Interoperable Multi-Agent Collaboration
by: Qian, Kaiyang, et al.
Published: (2026)
by: Qian, Kaiyang, et al.
Published: (2026)
Counterfactual Multi-Agent Policy Gradients
by: Foerster, Jakob, et al.
Published: (2017)
by: Foerster, Jakob, et al.
Published: (2017)
MALBO: Optimizing LLM-Based Multi-Agent Teams via Multi-Objective Bayesian Optimization
by: Sabbatella, Antonio
Published: (2025)
by: Sabbatella, Antonio
Published: (2025)
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
by: Yang, Shan, et al.
Published: (2026)
by: Yang, Shan, et al.
Published: (2026)
The Consensus Trap: Rescuing Multi-Agent LLMs from Adversarial Majorities via Token-Level Collaboration
by: Liu, Jiayuan, et al.
Published: (2026)
by: Liu, Jiayuan, et al.
Published: (2026)
Multi-Agent Security Tax: Trading Off Security and Collaboration Capabilities in Multi-Agent Systems
by: Peigne-Lefebvre, Pierre, et al.
Published: (2025)
by: Peigne-Lefebvre, Pierre, et al.
Published: (2025)
MACC: Multi-Agent Collaborative Competition for Scientific Exploration
by: Oyama, Satoshi, et al.
Published: (2026)
by: Oyama, Satoshi, et al.
Published: (2026)
BMW Agents -- A Framework For Task Automation Through Multi-Agent Collaboration
by: Crawford, Noel, et al.
Published: (2024)
by: Crawford, Noel, et al.
Published: (2024)
DDO: Dual-Decision Optimization for LLM-Based Medical Consultation via Multi-Agent Collaboration
by: Jia, Zhihao, et al.
Published: (2025)
by: Jia, Zhihao, et al.
Published: (2025)
TAPE: Leveraging Agent Topology for Cooperative Multi-Agent Policy Gradient
by: Lou, Xingzhou, et al.
Published: (2023)
by: Lou, Xingzhou, et al.
Published: (2023)
CREW-WILDFIRE: Benchmarking Agentic Multi-Agent Collaborations at Scale
by: Hyun, Jonathan, et al.
Published: (2025)
by: Hyun, Jonathan, et al.
Published: (2025)
HAWK: A Hierarchical Workflow Framework for Multi-Agent Collaboration
by: Cheng, Yuyang, et al.
Published: (2025)
by: Cheng, Yuyang, et al.
Published: (2025)
Beyond Frameworks: Unpacking Collaboration Strategies in Multi-Agent Systems
by: Wang, Haochun, et al.
Published: (2025)
by: Wang, Haochun, et al.
Published: (2025)
MultiAgent Collaboration Attack: Investigating Adversarial Attacks in Large Language Model Collaborations via Debate
by: Amayuelas, Alfonso, et al.
Published: (2024)
by: Amayuelas, Alfonso, et al.
Published: (2024)
Generalizable Agent Modeling for Agent Collaboration-Competition Adaptation with Multi-Retrieval and Dynamic Generation
by: Wang, Chenxu, et al.
Published: (2025)
by: Wang, Chenxu, et al.
Published: (2025)
OPTAGENT: Optimizing Multi-Agent LLM Interactions Through Verbal Reinforcement Learning for Enhanced Reasoning
by: Bi, Zhenyu, et al.
Published: (2025)
by: Bi, Zhenyu, et al.
Published: (2025)
Multi-source Plume Tracing via Multi-Agent Reinforcement Learning
by: Granadeno, Pedro Antonio Alarcon, et al.
Published: (2025)
by: Granadeno, Pedro Antonio Alarcon, et al.
Published: (2025)
Decentralized and Lifelong-Adaptive Multi-Agent Collaborative Learning
by: Tang, Shuo, et al.
Published: (2024)
by: Tang, Shuo, et al.
Published: (2024)
Data-Efficient Multi-Agent Spatial Planning with LLMs
by: Su, Huangyuan, et al.
Published: (2025)
by: Su, Huangyuan, et al.
Published: (2025)
Towards CausalGPT: A Multi-Agent Approach for Faithful Knowledge Reasoning via Promoting Causal Consistency in LLMs
by: Tang, Ziyi, et al.
Published: (2023)
by: Tang, Ziyi, et al.
Published: (2023)
Multi-Agent Reinforcement Learning with Communication-Constrained Priors
by: Yang, Guang, et al.
Published: (2025)
by: Yang, Guang, et al.
Published: (2025)
Similar Items
-
Learning to Deliberate: Meta-policy Collaboration for Agentic LLMs with Multi-agent Reinforcement Learning
by: Yang, Wei, et al.
Published: (2025) -
Evaluating Creativity and Deception in Large Language Models: A Simulation Framework for Multi-Agent Balderdash
by: Hejabi, Parsa, et al.
Published: (2024) -
Collaborative Belief Reasoning with LLMs for Efficient Multi-Agent Collaboration
by: Wang, Zhimin, et al.
Published: (2025) -
Agent-GSPO: Communication-Efficient Multi-Agent Systems via Group Sequence Policy Optimization
by: Fan, Yijia, et al.
Published: (2025) -
Multi-Agent Guided Policy Optimization
by: Li, Yueheng, et al.
Published: (2025)