UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Yiqun, Yang, Wei, Zhang, Erhan, Wang, Shijie, Liu, Qi, Niu, Zechun, Zhang, Bin, Li, Haitao, Li, Rui, Yan, Lingyong, Feng, Jinyuan, Qi, Biqing, Wei, Xiaochi, Gao, Yan, Wu, Yi, Hu, Yao, Mao, Jiaxin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Attack the Messages, Not the Agents: A Multi-round Adaptive Stealthy Tampering Framework for LLM-MAS
by: Yan, Bingyu, et al.
Published: (2025)
by: Yan, Bingyu, et al.
Published: (2025)
PTDE: Personalized Training with Distilled Execution for Multi-Agent Reinforcement Learning
by: Chen, Yiqun, et al.
Published: (2022)
by: Chen, Yiqun, et al.
Published: (2022)
OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search
by: Zhang, Erhan, et al.
Published: (2026)
by: Zhang, Erhan, et al.
Published: (2026)
SoK: Security of Autonomous LLM Agents in Agentic Commerce
by: Mao, Qian'ang, et al.
Published: (2026)
by: Mao, Qian'ang, et al.
Published: (2026)
MA2RL: Masked Autoencoders for Generalizable Multi-Agent Reinforcement Learning
by: Feng, Jinyuan, et al.
Published: (2025)
by: Feng, Jinyuan, et al.
Published: (2025)
Depending on yourself when you should: Mentoring LLM with RL agents to become the master in cybersecurity games
by: Yan, Yikuan, et al.
Published: (2024)
by: Yan, Yikuan, et al.
Published: (2024)
Enhancing LLM Problem Solving via Tutor-Student Multi-Agent Interaction
by: Özdemir, Nurullah Eymen, et al.
Published: (2026)
by: Özdemir, Nurullah Eymen, et al.
Published: (2026)
MAS-GPT: Training LLMs to Build LLM-based Multi-Agent Systems
by: Ye, Rui, et al.
Published: (2025)
by: Ye, Rui, et al.
Published: (2025)
GraphMASAL: A Graph-based Multi-Agent System for Adaptive Learning
by: Zeng, Biqing, et al.
Published: (2025)
by: Zeng, Biqing, et al.
Published: (2025)
Aegis:An Advanced LLM-Based Multi-Agent for Intelligent Functional Safety Engineering
by: Shi, Lu, et al.
Published: (2024)
by: Shi, Lu, et al.
Published: (2024)
PeroMAS: A Multi-agent System of Perovskite Material Discovery
by: Wang, Yishu, et al.
Published: (2026)
by: Wang, Yishu, et al.
Published: (2026)
IBGP: Imperfect Byzantine Generals Problem for Zero-Shot Robustness in Communicative Multi-Agent Systems
by: Mao, Yihuan, et al.
Published: (2024)
by: Mao, Yihuan, et al.
Published: (2024)
Attrition-Aware Adaptation for Multi-Agent Patrolling
by: Goeckner, Anthony, et al.
Published: (2023)
by: Goeckner, Anthony, et al.
Published: (2023)
City-LEO: Toward Transparent City Management Using LLM with End-to-End Optimization
by: Jiao, Zihao, et al.
Published: (2024)
by: Jiao, Zihao, et al.
Published: (2024)
Towards Transparent and Incentive-Compatible Collaboration in Decentralized LLM Multi-Agent Systems: A Blockchain-Driven Approach
by: Qi, Minfeng, et al.
Published: (2025)
by: Qi, Minfeng, et al.
Published: (2025)
Single-Agent Scaling Fails Multi-Agent Intelligence: Towards Foundation Models with Native Multi-Agent Intelligence
by: Hu, Shuyue, et al.
Published: (2025)
by: Hu, Shuyue, et al.
Published: (2025)
AgentNet: Decentralized Evolutionary Coordination for LLM-based Multi-Agent Systems
by: Yang, Yingxuan, et al.
Published: (2025)
by: Yang, Yingxuan, et al.
Published: (2025)
LLM-Foraging: Large Language Models for Decentralized Swarm Robot Foraging
by: Li, Peihan, et al.
Published: (2026)
by: Li, Peihan, et al.
Published: (2026)
Coupling Agent-Based Simulations and VR universes: the case of GAMA and Unity
by: Drogoul, Alexis, et al.
Published: (2025)
by: Drogoul, Alexis, et al.
Published: (2025)
PartnerMAS: An LLM Hierarchical Multi-Agent Framework for Business Partner Selection on High-Dimensional Features
by: Li, Lingyao, et al.
Published: (2025)
by: Li, Lingyao, et al.
Published: (2025)
SkillMAS: Skill Co-Evolution with LLM-based Multi-Agent System
by: Pan, Shuai, et al.
Published: (2026)
by: Pan, Shuai, et al.
Published: (2026)
Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent Systems
by: Yan, Bingyu, et al.
Published: (2025)
by: Yan, Bingyu, et al.
Published: (2025)
MAS-Shield: A Defense Framework for Secure and Efficient LLM MAS
by: Wang, Kaixiang, et al.
Published: (2025)
by: Wang, Kaixiang, et al.
Published: (2025)
Bridging Training and Execution via Dynamic Directed Graph-Based Communication in Cooperative Multi-Agent Systems
by: Zhang, Zhuohui, et al.
Published: (2024)
by: Zhang, Zhuohui, et al.
Published: (2024)
Adversarial Yet Cooperative: Multi-Perspective Reasoning in Retrieved-Augmented Language Models
by: Xu, Can, et al.
Published: (2026)
by: Xu, Can, et al.
Published: (2026)
Graphon Mean-Field Control for Cooperative Multi-Agent Reinforcement Learning
by: Hu, Yuanquan, et al.
Published: (2022)
by: Hu, Yuanquan, et al.
Published: (2022)
HieraMAS: Optimizing Intra-Node LLM Mixtures and Inter-Node Topology for Multi-Agent Systems
by: Yao, Tianjun, et al.
Published: (2026)
by: Yao, Tianjun, et al.
Published: (2026)
OpenMines: A Light and Comprehensive Mining Simulation Environment for Truck Dispatching
by: Meng, Shi, et al.
Published: (2024)
by: Meng, Shi, et al.
Published: (2024)
ToMacVF : Temporal Macro-action Value Factorization for Asynchronous Multi-Agent Reinforcement Learning
by: Zhang, Wenjing, et al.
Published: (2025)
by: Zhang, Wenjing, et al.
Published: (2025)
An Improved Multi-Agent Algorithm for Cooperative and Competitive Environments by Identifying and Encouraging Cooperation among Agents
by: Qi, Junjie, et al.
Published: (2025)
by: Qi, Junjie, et al.
Published: (2025)
PIVOT: Bridging Planning and Execution in LLM Agents via Trajectory Refinement
by: Zhang, Tuo, et al.
Published: (2026)
by: Zhang, Tuo, et al.
Published: (2026)
EvoRoute: Experience-Driven Self-Routing LLM Agent Systems
by: Zhang, Guibin, et al.
Published: (2026)
by: Zhang, Guibin, et al.
Published: (2026)
Causal-Inspired Multi-Agent Decision-Making via Graph Reinforcement Learning
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
DTPPO: Dual-Transformer Encoder-based Proximal Policy Optimization for Multi-UAV Navigation in Unseen Complex Environments
by: Wei, Anning, et al.
Published: (2024)
by: Wei, Anning, et al.
Published: (2024)
Guiding LLM-Based Human Mobility Simulation with Mobility Measures from Shared Data
by: Yan, Hua, et al.
Published: (2026)
by: Yan, Hua, et al.
Published: (2026)
Multi-Agent Path Finding via Offline RL and LLM Collaboration
by: Atasever, Merve, et al.
Published: (2025)
by: Atasever, Merve, et al.
Published: (2025)
LLM-Enhanced Multi-Agent Reinforcement Learning with Expert Workflow for Real-Time P2P Energy Trading
by: Lou, Chengwei, et al.
Published: (2025)
by: Lou, Chengwei, et al.
Published: (2025)
DarwinTOD: LLM-driven Lifelong Self-evolution for Task-oriented Dialog Systems
by: Zhang, Shuyu, et al.
Published: (2026)
by: Zhang, Shuyu, et al.
Published: (2026)
On the external concurrency of current BDI frameworks for MAS
by: Baiardi, Martina, et al.
Published: (2024)
by: Baiardi, Martina, et al.
Published: (2024)
Multi-Agent Motion Planning with Bézier Curve Optimization under Kinodynamic Constraints
by: Yan, Jingtian, et al.
Published: (2023)
by: Yan, Jingtian, et al.
Published: (2023)
Similar Items
-
Attack the Messages, Not the Agents: A Multi-round Adaptive Stealthy Tampering Framework for LLM-MAS
by: Yan, Bingyu, et al.
Published: (2025) -
PTDE: Personalized Training with Distilled Execution for Multi-Agent Reinforcement Learning
by: Chen, Yiqun, et al.
Published: (2022) -
OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search
by: Zhang, Erhan, et al.
Published: (2026) -
SoK: Security of Autonomous LLM Agents in Agentic Commerce
by: Mao, Qian'ang, et al.
Published: (2026) -
MA2RL: Masked Autoencoders for Generalizable Multi-Agent Reinforcement Learning
by: Feng, Jinyuan, et al.
Published: (2025)