Toward Scalable Multirobot Control: Fast Policy Learning in Distributed MPC
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Xinglong, Pan, Wei, Li, Cong, Xu, Xin, Wang, Xiangke, Zhang, Ronghua, Hu, Dewen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
von: Ma, Yunchang, et al.
Veröffentlicht: (2025)
von: Ma, Yunchang, et al.
Veröffentlicht: (2025)
SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-based Humanoid Control
von: Zhang, Jingyan, et al.
Veröffentlicht: (2026)
von: Zhang, Jingyan, et al.
Veröffentlicht: (2026)
SpikingSoft: A Spiking Neuron Controller for Bio-inspired Locomotion with Soft Snake Robots
von: Zhang, Chuhan, et al.
Veröffentlicht: (2025)
von: Zhang, Chuhan, et al.
Veröffentlicht: (2025)
SeFA-Policy: Fast and Accurate Visuomotor Policy Learning with Selective Flow Alignment
von: Xue, Rong, et al.
Veröffentlicht: (2025)
von: Xue, Rong, et al.
Veröffentlicht: (2025)
Fast Policy Learning for 6-DOF Position Control of Underwater Vehicles
von: Tunçay, Sümer, et al.
Veröffentlicht: (2025)
von: Tunçay, Sümer, et al.
Veröffentlicht: (2025)
CoVO-MPC: Theoretical Analysis of Sampling-based MPC and Optimal Covariance Design
von: Yi, Zeji, et al.
Veröffentlicht: (2024)
von: Yi, Zeji, et al.
Veröffentlicht: (2024)
DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion
von: Fan, Yahao, et al.
Veröffentlicht: (2025)
von: Fan, Yahao, et al.
Veröffentlicht: (2025)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint
von: Lin, Haotian, et al.
Veröffentlicht: (2025)
von: Lin, Haotian, et al.
Veröffentlicht: (2025)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
TD-MPC2: Scalable, Robust World Models for Continuous Control
von: Hansen, Nicklas, et al.
Veröffentlicht: (2023)
von: Hansen, Nicklas, et al.
Veröffentlicht: (2023)
One-Step Diffusion Policy: Fast Visuomotor Policies via Diffusion Distillation
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
MPC-based Deep Reinforcement Learning Method for Space Robotic Control with Fuel Sloshing Mitigation
von: Ramezani, Mahya, et al.
Veröffentlicht: (2025)
von: Ramezani, Mahya, et al.
Veröffentlicht: (2025)
Reactive Diffusion Policy: Slow-Fast Visual-Tactile Policy Learning for Contact-Rich Manipulation
von: Xue, Han, et al.
Veröffentlicht: (2025)
von: Xue, Han, et al.
Veröffentlicht: (2025)
Lyapunov Stability Learning with Nonlinear Control via Inductive Biases
von: Lu, Yupu, et al.
Veröffentlicht: (2025)
von: Lu, Yupu, et al.
Veröffentlicht: (2025)
PolaRiS: Scalable Real-to-Sim Evaluations for Generalist Robot Policies
von: Jain, Arhan, et al.
Veröffentlicht: (2025)
von: Jain, Arhan, et al.
Veröffentlicht: (2025)
SOMTP: Self-Supervised Learning-Based Optimizer for MPC-Based Safe Trajectory Planning Problems in Robotics
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
MPC-Inspired Reinforcement Learning for Verifiable Model-Free Control
von: Lu, Yiwen, et al.
Veröffentlicht: (2023)
von: Lu, Yiwen, et al.
Veröffentlicht: (2023)
NaviMaster: Learning a Unified Policy for GUI and Embodied Navigation Tasks
von: Luo, Zhihao, et al.
Veröffentlicht: (2025)
von: Luo, Zhihao, et al.
Veröffentlicht: (2025)
DR-MPC: Deep Residual Model Predictive Control for Real-world Social Navigation
von: Han, James R., et al.
Veröffentlicht: (2024)
von: Han, James R., et al.
Veröffentlicht: (2024)
SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
Autonomous Vehicle Lateral Control Using Deep Reinforcement Learning with MPC-PID Demonstration
von: Wu, Chengdong, et al.
Veröffentlicht: (2025)
von: Wu, Chengdong, et al.
Veröffentlicht: (2025)
Robot Policy Learning with Temporal Optimal Transport Reward
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
Robot Fleet Learning via Policy Merging
von: Wang, Lirui, et al.
Veröffentlicht: (2023)
von: Wang, Lirui, et al.
Veröffentlicht: (2023)
3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
von: Ze, Yanjie, et al.
Veröffentlicht: (2024)
von: Ze, Yanjie, et al.
Veröffentlicht: (2024)
Compose Your Policies! Improving Diffusion-based or Flow-based Robot Policies via Test-time Distribution-level Composition
von: Cao, Jiahang, et al.
Veröffentlicht: (2025)
von: Cao, Jiahang, et al.
Veröffentlicht: (2025)
Mini Diffuser: Fast Multi-task Diffusion Policy Training Using Two-level Mini-batches
von: Hu, Yutong, et al.
Veröffentlicht: (2025)
von: Hu, Yutong, et al.
Veröffentlicht: (2025)
Generalized Advantage Estimation for Distributional Policy Gradients
von: Shaik, Shahil, et al.
Veröffentlicht: (2025)
von: Shaik, Shahil, et al.
Veröffentlicht: (2025)
Scalable Exploration for High-Dimensional Continuous Control via Value-Guided Flow
von: Wei, Yunyue, et al.
Veröffentlicht: (2026)
von: Wei, Yunyue, et al.
Veröffentlicht: (2026)
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents
von: Kuzmenko, Dmytro, et al.
Veröffentlicht: (2025)
von: Kuzmenko, Dmytro, et al.
Veröffentlicht: (2025)
Scalable Dexterous Robot Learning with AR-based Remote Human-Robot Interactions
von: Yang, Yicheng, et al.
Veröffentlicht: (2026)
von: Yang, Yicheng, et al.
Veröffentlicht: (2026)
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies
von: Patil, Sarvesh, et al.
Veröffentlicht: (2026)
von: Patil, Sarvesh, et al.
Veröffentlicht: (2026)
From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
Fast and Robust Visuomotor Riemannian Flow Matching Policy
von: Ding, Haoran, et al.
Veröffentlicht: (2024)
von: Ding, Haoran, et al.
Veröffentlicht: (2024)
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
von: Zhang, Tonghe, et al.
Veröffentlicht: (2025)
von: Zhang, Tonghe, et al.
Veröffentlicht: (2025)
CaRL: Learning Scalable Planning Policies with Simple Rewards
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
Towards Learning Scalable Agile Dynamic Motion Planning for Robosoccer Teams with Policy Optimization
von: Ho, Brandon, et al.
Veröffentlicht: (2025)
von: Ho, Brandon, et al.
Veröffentlicht: (2025)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2026)
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2026)
ArrayBot: Reinforcement Learning for Generalizable Distributed Manipulation through Touch
von: Xue, Zhengrong, et al.
Veröffentlicht: (2023)
von: Xue, Zhengrong, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
von: Ma, Yunchang, et al.
Veröffentlicht: (2025) -
SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-based Humanoid Control
von: Zhang, Jingyan, et al.
Veröffentlicht: (2026) -
SpikingSoft: A Spiking Neuron Controller for Bio-inspired Locomotion with Soft Snake Robots
von: Zhang, Chuhan, et al.
Veröffentlicht: (2025) -
SeFA-Policy: Fast and Accurate Visuomotor Policy Learning with Selective Flow Alignment
von: Xue, Rong, et al.
Veröffentlicht: (2025) -
Fast Policy Learning for 6-DOF Position Control of Underwater Vehicles
von: Tunçay, Sümer, et al.
Veröffentlicht: (2025)