Learning Multi-Robot Coordination through Locality-Based Factorized Multi-Agent Actor-Critic Algorithm
Fuente:
arXiv
Saved in:
| Main Authors: | Shek, Chak Lam, Bedi, Amrit Singh, Basak, Anjon, Novoseller, Ellen, Waytowich, Nick, Narayanan, Priya, Manocha, Dinesh, Tokekar, Pratap |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decision-Oriented Learning Using Differentiable Submodular Maximization for Multi-Robot Coordination
by: Shi, Guangyao, et al.
Published: (2023)
by: Shi, Guangyao, et al.
Published: (2023)
LANCAR: Leveraging Language for Context-Aware Robot Locomotion in Unstructured Environments
by: Shek, Chak Lam, et al.
Published: (2023)
by: Shek, Chak Lam, et al.
Published: (2023)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
by: Shek, Chak Lam, et al.
Published: (2025)
by: Shek, Chak Lam, et al.
Published: (2025)
Multi-Agent Trust Region Policy Optimisation: A Joint Constraint Approach
by: Shek, Chak Lam, et al.
Published: (2025)
by: Shek, Chak Lam, et al.
Published: (2025)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)
by: Chakraborty, Souradip, et al.
Published: (2023)
When to Localize? A Risk-Constrained Reinforcement Learning Approach
by: Shek, Chak Lam, et al.
Published: (2024)
by: Shek, Chak Lam, et al.
Published: (2024)
Beyond Joint Demonstrations: Personalized Expert Guidance for Efficient Multi-Agent Reinforcement Learning
by: Yu, Peihong, et al.
Published: (2024)
by: Yu, Peihong, et al.
Published: (2024)
DMCA: Dense Multi-agent Navigation using Attention and Communication
by: Arul, Senthil Hariharan, et al.
Published: (2022)
by: Arul, Senthil Hariharan, et al.
Published: (2022)
Multi-LLM QA with Embodied Exploration
by: Patel, Bhrij, et al.
Published: (2024)
by: Patel, Bhrij, et al.
Published: (2024)
Contextual Neural Moving Horizon Estimation for Robust Quadrotor Control in Varying Conditions
by: Torshizi, Kasra, et al.
Published: (2025)
by: Torshizi, Kasra, et al.
Published: (2025)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
by: Barakat, Anas, et al.
Published: (2024)
by: Barakat, Anas, et al.
Published: (2024)
Personalized Embodied Navigation for Portable Object Finding
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
Code Comprehension then Auditing for Unsupervised LLM Evaluation
by: Patel, Bhrij, et al.
Published: (2024)
by: Patel, Bhrij, et al.
Published: (2024)
Improving Zero-Shot ObjectNav with Generative Communication
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
Closing the Gap: Achieving Global Convergence (Last Iterate) of Actor-Critic under Markovian Sampling with Neural Network Parametrization
by: Gaur, Mudit, et al.
Published: (2024)
by: Gaur, Mudit, et al.
Published: (2024)
R2BC: Multi-Agent Imitation Learning from Single-Agent Demonstrations
by: Mattson, Connor, et al.
Published: (2025)
by: Mattson, Connor, et al.
Published: (2025)
Safety Recovery in Reasoning Models Is Only a Few Early Steering Steps Away
by: Ghosal, Soumya Suvra, et al.
Published: (2026)
by: Ghosal, Soumya Suvra, et al.
Published: (2026)
SABER: A Stealthy Agentic Black-Box Attack Framework for Vision-Language-Action Models
by: Wu, Xiyang, et al.
Published: (2026)
by: Wu, Xiyang, et al.
Published: (2026)
Fast Biconnectivity Restoration in Multi-Robot Systems for Robust Communication Maintenance
by: Ishat-E-Rabban, Md, et al.
Published: (2020)
by: Ishat-E-Rabban, Md, et al.
Published: (2020)
Pre-Trained Masked Image Model for Mobile Robot Navigation
by: Sharma, Vishnu Dutt, et al.
Published: (2023)
by: Sharma, Vishnu Dutt, et al.
Published: (2023)
MO-Playground: Massively Parallelized Multi-Objective Reinforcement Learning for Robotics
by: Janwani, Neil, et al.
Published: (2026)
by: Janwani, Neil, et al.
Published: (2026)
On the Vulnerability of LLM/VLM-Controlled Robotics
by: Wu, Xiyang, et al.
Published: (2024)
by: Wu, Xiyang, et al.
Published: (2024)
AG-CVG: Coverage Planning with a Mobile Recharging UGV and an Energy-Constrained UAV
by: Karapetyan, Nare, et al.
Published: (2023)
by: Karapetyan, Nare, et al.
Published: (2023)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
by: Patel, Bhrij, et al.
Published: (2023)
by: Patel, Bhrij, et al.
Published: (2023)
PARL: A Unified Framework for Policy Alignment in Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)
by: Chakraborty, Souradip, et al.
Published: (2023)
Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment
by: Ghosal, Soumya Suvra, et al.
Published: (2024)
by: Ghosal, Soumya Suvra, et al.
Published: (2024)
Active Asymmetric Multi-Agent Multimodal Learning under Uncertainty
by: Liu, Rui, et al.
Published: (2026)
by: Liu, Rui, et al.
Published: (2026)
Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings
by: Powell, Keenan, et al.
Published: (2026)
by: Powell, Keenan, et al.
Published: (2026)
Human-in-the-Loop Multi-Robot Information Gathering with Inverse Submodular Maximization
by: Shi, Guangyao, et al.
Published: (2024)
by: Shi, Guangyao, et al.
Published: (2024)
Achieving Zero Constraint Violation for Constrained Reinforcement Learning via Conservative Natural Policy Gradient Primal-Dual Algorithm
by: Bai, Qinbo, et al.
Published: (2022)
by: Bai, Qinbo, et al.
Published: (2022)
KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning
by: Shek, Chak Lam, et al.
Published: (2026)
by: Shek, Chak Lam, et al.
Published: (2026)
AI Cap-and-Trade: Efficiency Incentives for Accessibility and Sustainability
by: Bornstein, Marco, et al.
Published: (2026)
by: Bornstein, Marco, et al.
Published: (2026)
Fast k-connectivity Restoration in Multi-Robot Systems for Robust Communication Maintenance
by: Ishat-E-Rabban, Md, et al.
Published: (2024)
by: Ishat-E-Rabban, Md, et al.
Published: (2024)
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time
by: Chehade, Mohamad, et al.
Published: (2025)
by: Chehade, Mohamad, et al.
Published: (2025)
Transfer Q Star: Principled Decoding for LLM Alignment
by: Chakraborty, Souradip, et al.
Published: (2024)
by: Chakraborty, Souradip, et al.
Published: (2024)
Rating-based Reinforcement Learning
by: White, Devin, et al.
Published: (2023)
by: White, Devin, et al.
Published: (2023)
When to Localize? A POMDP Approach
by: Williams, Troi, et al.
Published: (2024)
by: Williams, Troi, et al.
Published: (2024)
MaxMin-RLHF: Alignment with Diverse Human Preferences
by: Chakraborty, Souradip, et al.
Published: (2024)
by: Chakraborty, Souradip, et al.
Published: (2024)
Towards Global Optimality for Practical Average Reward Reinforcement Learning without Mixing Time Oracles
by: Patel, Bhrij, et al.
Published: (2024)
by: Patel, Bhrij, et al.
Published: (2024)
Similar Items
-
Decision-Oriented Learning Using Differentiable Submodular Maximization for Multi-Robot Coordination
by: Shi, Guangyao, et al.
Published: (2023) -
LANCAR: Leveraging Language for Context-Aware Robot Locomotion in Unstructured Environments
by: Shek, Chak Lam, et al.
Published: (2023) -
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
by: Shek, Chak Lam, et al.
Published: (2025) -
Multi-Agent Trust Region Policy Optimisation: A Joint Constraint Approach
by: Shek, Chak Lam, et al.
Published: (2025) -
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)