LLM-Mediated Guidance of MARL Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Siedler, Philipp D., Gemp, Ian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation
by: Siedler, Philipp D.
Published: (2026)
by: Siedler, Philipp D.
Published: (2026)
HIVEX: A High-Impact Environment Suite for Multi-Agent Research (extended version)
by: Siedler, Philipp Dominic
Published: (2025)
by: Siedler, Philipp Dominic
Published: (2025)
Learning to Communicate and Collaborate in a Competitive Multi-Agent Setup to Clean the Ocean from Macroplastics
by: Siedler, Philipp Dominic
Published: (2023)
by: Siedler, Philipp Dominic
Published: (2023)
Self-Resource Allocation in Multi-Agent LLM Systems
by: Amayuelas, Alfonso, et al.
Published: (2025)
by: Amayuelas, Alfonso, et al.
Published: (2025)
MASLab: A Unified and Comprehensive Codebase for LLM-based Multi-Agent Systems
by: Ye, Rui, et al.
Published: (2025)
by: Ye, Rui, et al.
Published: (2025)
Epistemic Context Learning: Building Trust the Right Way in LLM-Based Multi-Agent Systems
by: Zhou, Ruiwen, et al.
Published: (2026)
by: Zhou, Ruiwen, et al.
Published: (2026)
RCR-Router: Efficient Role-Aware Context Routing for Multi-Agent LLM Systems with Structured Memory
by: Liu, Jun, et al.
Published: (2025)
by: Liu, Jun, et al.
Published: (2025)
UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems
by: Chen, Yiqun, et al.
Published: (2026)
by: Chen, Yiqun, et al.
Published: (2026)
Scheming Ability in LLM-to-LLM Strategic Interactions
by: Pham, Thao
Published: (2025)
by: Pham, Thao
Published: (2025)
From Competition to Coordination: Market Making as a Scalable Framework for Safe and Aligned Multi-Agent LLM Systems
by: Gho, Brendan, et al.
Published: (2025)
by: Gho, Brendan, et al.
Published: (2025)
LLMs Working in Harmony: A Survey on the Technological Aspects of Building Effective LLM-Based Multi Agent Systems
by: Aratchige, R. M., et al.
Published: (2025)
by: Aratchige, R. M., et al.
Published: (2025)
Diversity Collapse in Multi-Agent LLM Systems: Structural Coupling and Collective Failure in Open-Ended Idea Generation
by: Chen, Nuo, et al.
Published: (2026)
by: Chen, Nuo, et al.
Published: (2026)
Layered Chain-of-Thought Prompting for Multi-Agent LLM Systems: A Comprehensive Approach to Explainable Large Language Models
by: Sanwal, Manish
Published: (2025)
by: Sanwal, Manish
Published: (2025)
LLM-as-RNN: A Recurrent Language Model for Memory Updates and Sequence Prediction
by: Lu, Yuxing, et al.
Published: (2026)
by: Lu, Yuxing, et al.
Published: (2026)
Towards a Science of Collective AI: LLM-based Multi-Agent Systems Need a Transition from Blind Trial-and-Error to Rigorous Science
by: Fan, Jingru, et al.
Published: (2026)
by: Fan, Jingru, et al.
Published: (2026)
Adaptive Memory Admission Control for LLM Agents
by: Zhang, Guilin, et al.
Published: (2026)
by: Zhang, Guilin, et al.
Published: (2026)
Solving a Million-Step LLM Task with Zero Errors
by: Meyerson, Elliot, et al.
Published: (2025)
by: Meyerson, Elliot, et al.
Published: (2025)
Cross-Task Experiential Learning on LLM-based Multi-Agent Collaboration
by: Li, Yilong, et al.
Published: (2025)
by: Li, Yilong, et al.
Published: (2025)
Time to Talk: LLM Agents for Asynchronous Group Communication in Mafia Games
by: Eckhaus, Niv, et al.
Published: (2025)
by: Eckhaus, Niv, et al.
Published: (2025)
GUARDIAN: Safeguarding LLM Multi-Agent Collaborations with Temporal Graph Modeling
by: Zhou, Jialong, et al.
Published: (2025)
by: Zhou, Jialong, et al.
Published: (2025)
Experience Compression Spectrum: Unifying Memory, Skills, and Rules in LLM Agents
by: Zhang, Xing, et al.
Published: (2026)
by: Zhang, Xing, et al.
Published: (2026)
iMAD: Intelligent Multi-Agent Debate for Efficient and Accurate LLM Inference
by: Fan, Wei, et al.
Published: (2025)
by: Fan, Wei, et al.
Published: (2025)
Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM Collectives
by: Ko, Changgeon, et al.
Published: (2026)
by: Ko, Changgeon, et al.
Published: (2026)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
by: Liu, Zijun, et al.
Published: (2023)
by: Liu, Zijun, et al.
Published: (2023)
LLM Review: Enhancing Creative Writing via Blind Peer Review Feedback
by: Li, Weiyue, et al.
Published: (2026)
by: Li, Weiyue, et al.
Published: (2026)
From Assumptions to Actions: Turning LLM Reasoning into Uncertainty-Aware Planning for Embodied Agents
by: Seo, SeungWon, et al.
Published: (2026)
by: Seo, SeungWon, et al.
Published: (2026)
Exploring Plan Space through Conversation: An Agentic Framework for LLM-Mediated Explanations in Planning
by: Fouilhé, Guilhem, et al.
Published: (2026)
by: Fouilhé, Guilhem, et al.
Published: (2026)
Stated Preference for Interaction and Continued Engagement (SPICE): Evaluating an LLM's Willingness to Re-engage in Conversation
by: Rost, Thomas Manuel, et al.
Published: (2025)
by: Rost, Thomas Manuel, et al.
Published: (2025)
MARBLE: A Multi-Agent Rule-Based LLM Reasoning Engine for Accident Severity Prediction
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
DDO: Dual-Decision Optimization for LLM-Based Medical Consultation via Multi-Agent Collaboration
by: Jia, Zhihao, et al.
Published: (2025)
by: Jia, Zhihao, et al.
Published: (2025)
Convex Markov Games: A New Frontier for Multi-Agent Reinforcement Learning
by: Gemp, Ian, et al.
Published: (2024)
by: Gemp, Ian, et al.
Published: (2024)
Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents
by: Sethi, Khushal
Published: (2026)
by: Sethi, Khushal
Published: (2026)
Trust, Lies, and Long Memories: Emergent Social Dynamics and Reputation in Multi-Round Avalon with LLM Agents
by: Ellawela, Suveen
Published: (2026)
by: Ellawela, Suveen
Published: (2026)
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
by: Sinha, Aarush, et al.
Published: (2026)
by: Sinha, Aarush, et al.
Published: (2026)
RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents
by: Rosati, Riccardo, et al.
Published: (2026)
by: Rosati, Riccardo, et al.
Published: (2026)
Grammar Search for Multi-Agent Systems
by: Singh, Mayank, et al.
Published: (2025)
by: Singh, Mayank, et al.
Published: (2025)
Process-Supervised Reinforcement Learning for Interactive Multimodal Tool-Use Agents
by: Tan, Weiting, et al.
Published: (2025)
by: Tan, Weiting, et al.
Published: (2025)
Enhancing Multi-Agent Consensus through Third-Party LLM Integration: Analyzing Uncertainty and Mitigating Hallucinations in Large Language Models
by: Duan, Zhihua, et al.
Published: (2024)
by: Duan, Zhihua, et al.
Published: (2024)
Unmasking Conversational Bias in AI Multiagent Systems
by: Coppolillo, Erica, et al.
Published: (2025)
by: Coppolillo, Erica, et al.
Published: (2025)
Adaptive Multi-Agent Response Refinement in Conversational Systems
by: Jeong, Soyeong, et al.
Published: (2025)
by: Jeong, Soyeong, et al.
Published: (2025)
Similar Items
-
Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation
by: Siedler, Philipp D.
Published: (2026) -
HIVEX: A High-Impact Environment Suite for Multi-Agent Research (extended version)
by: Siedler, Philipp Dominic
Published: (2025) -
Learning to Communicate and Collaborate in a Competitive Multi-Agent Setup to Clean the Ocean from Macroplastics
by: Siedler, Philipp Dominic
Published: (2023) -
Self-Resource Allocation in Multi-Agent LLM Systems
by: Amayuelas, Alfonso, et al.
Published: (2025) -
MASLab: A Unified and Comprehensive Codebase for LLM-based Multi-Agent Systems
by: Ye, Rui, et al.
Published: (2025)