Direct Post-Training Preference Alignment for Multi-Agent Motion Generation Models Using Implicit Feedback from Pre-training Demonstrations
Fuente:
arXiv
Saved in:
| Main Authors: | Tian, Ran, Goel, Kratarth |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Integrating Controllable Motion Skills from Demonstrations
by: Liao, Honghao, et al.
Published: (2024)
by: Liao, Honghao, et al.
Published: (2024)
Representation Alignment from Human Feedback for Cross-Embodiment Reward Learning from Mixed-Quality Demonstrations
by: Mattson, Connor, et al.
Published: (2024)
by: Mattson, Connor, et al.
Published: (2024)
Leveraging Pre-trained Large Language Models with Refined Prompting for Online Task and Motion Planning
by: Guo, Huihui, et al.
Published: (2025)
by: Guo, Huihui, et al.
Published: (2025)
Learning Reward and Policy Jointly from Demonstration and Preference Improves Alignment
by: Li, Chenliang, et al.
Published: (2024)
by: Li, Chenliang, et al.
Published: (2024)
Generating Physically Realistic and Directable Human Motions from Multi-Modal Inputs
by: Shrestha, Aayam, et al.
Published: (2025)
by: Shrestha, Aayam, et al.
Published: (2025)
Scaling Motion Forecasting Models with Ensemble Distillation
by: Ettinger, Scott, et al.
Published: (2024)
by: Ettinger, Scott, et al.
Published: (2024)
Scaling Laws of Motion Forecasting and Planning -- Technical Report
by: Baniodeh, Mustafa, et al.
Published: (2025)
by: Baniodeh, Mustafa, et al.
Published: (2025)
Steerable Adversarial Scenario Generation through Test-Time Preference Alignment
by: Nie, Tong, et al.
Published: (2025)
by: Nie, Tong, et al.
Published: (2025)
Lifelike Agility and Play in Quadrupedal Robots using Reinforcement Learning and Generative Pre-trained Models
by: Han, Lei, et al.
Published: (2023)
by: Han, Lei, et al.
Published: (2023)
Learning to Discern: Imitating Heterogeneous Human Demonstrations with Preference and Representation Learning
by: Kuhar, Sachit, et al.
Published: (2023)
by: Kuhar, Sachit, et al.
Published: (2023)
AeroVerse: UAV-Agent Benchmark Suite for Simulating, Pre-training, Finetuning, and Evaluating Aerospace Embodied World Models
by: Yao, Fanglong, et al.
Published: (2024)
by: Yao, Fanglong, et al.
Published: (2024)
Aligning Humans and Robots via Reinforcement Learning from Implicit Human Feedback
by: Kim, Suzie, et al.
Published: (2025)
by: Kim, Suzie, et al.
Published: (2025)
Reinforcement Learning from Implicit Neural Feedback for Human-Aligned Robot Control
by: Kim, Suzie
Published: (2025)
by: Kim, Suzie
Published: (2025)
FoldPath: End-to-End Object-Centric Motion Generation via Modulated Implicit Paths
by: Rabino, Paolo, et al.
Published: (2025)
by: Rabino, Paolo, et al.
Published: (2025)
R2BC: Multi-Agent Imitation Learning from Single-Agent Demonstrations
by: Mattson, Connor, et al.
Published: (2025)
by: Mattson, Connor, et al.
Published: (2025)
Implicit Kinodynamic Motion Retargeting for Human-to-humanoid Imitation Learning
by: Chen, Xingyu, et al.
Published: (2025)
by: Chen, Xingyu, et al.
Published: (2025)
Pre-training Auto-regressive Robotic Models with 4D Representations
by: Niu, Dantong, et al.
Published: (2025)
by: Niu, Dantong, et al.
Published: (2025)
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
by: Tian, Ran, et al.
Published: (2024)
by: Tian, Ran, et al.
Published: (2024)
Learning Sequential Kinematic Models from Demonstrations for Multi-Jointed Articulated Objects
by: Gupta, Anmol, et al.
Published: (2025)
by: Gupta, Anmol, et al.
Published: (2025)
Multi-Agent Planning Using Visual Language Models
by: Brienza, Michele, et al.
Published: (2024)
by: Brienza, Michele, et al.
Published: (2024)
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm
by: Kim, Hyeonjun, et al.
Published: (2025)
by: Kim, Hyeonjun, et al.
Published: (2025)
ARCap: Collecting High-quality Human Demonstrations for Robot Learning with Augmented Reality Feedback
by: Chen, Sirui, et al.
Published: (2024)
by: Chen, Sirui, et al.
Published: (2024)
Forecast-PEFT: Parameter-Efficient Fine-Tuning for Pre-trained Motion Forecasting Models
by: Wang, Jifeng, et al.
Published: (2024)
by: Wang, Jifeng, et al.
Published: (2024)
Hierarchical Vision Language Action Model Using Success and Failure Demonstrations
by: Park, Jeongeun, et al.
Published: (2025)
by: Park, Jeongeun, et al.
Published: (2025)
Parallels Between VLA Model Post-Training and Human Motor Learning: Progress, Challenges, and Trends
by: Xiang, Tian-Yu, et al.
Published: (2025)
by: Xiang, Tian-Yu, et al.
Published: (2025)
Learning from Demonstration with Implicit Nonlinear Dynamics Models
by: Fagan, Peter David, et al.
Published: (2024)
by: Fagan, Peter David, et al.
Published: (2024)
Neural-Assisted in-Motion Self-Heading Alignment
by: Yampolsky, Zeev, et al.
Published: (2026)
by: Yampolsky, Zeev, et al.
Published: (2026)
Multi-Agent Inverse Q-Learning from Demonstrations
by: Haynam, Nathaniel, et al.
Published: (2025)
by: Haynam, Nathaniel, et al.
Published: (2025)
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
by: Li, Haoyun, et al.
Published: (2025)
by: Li, Haoyun, et al.
Published: (2025)
Autonomous Implicit Indoor Scene Reconstruction with Frontier Exploration
by: Zeng, Jing, et al.
Published: (2024)
by: Zeng, Jing, et al.
Published: (2024)
Evaluating VLMs' Spatial Reasoning Over Robot Motion: A Step Towards Robot Planning with Motion Preferences
by: Wu, Wenxi, et al.
Published: (2026)
by: Wu, Wenxi, et al.
Published: (2026)
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning
by: Zhou, Gaoyue, et al.
Published: (2024)
by: Zhou, Gaoyue, et al.
Published: (2024)
Learning Pivoting Manipulation with Force and Vision Feedback Using Optimization-based Demonstrations
by: Shirai, Yuki, et al.
Published: (2025)
by: Shirai, Yuki, et al.
Published: (2025)
Zero-Shot Metric Depth Estimation via Monocular Visual-Inertial Rescaling for Autonomous Aerial Navigation
by: Yang, Steven, et al.
Published: (2025)
by: Yang, Steven, et al.
Published: (2025)
Improving Pre-Trained Vision-Language-Action Policies with Model-Based Search
by: Neary, Cyrus, et al.
Published: (2025)
by: Neary, Cyrus, et al.
Published: (2025)
RICL: Adding In-Context Adaptability to Pre-Trained Vision-Language-Action Models
by: Sridhar, Kaustubh, et al.
Published: (2025)
by: Sridhar, Kaustubh, et al.
Published: (2025)
Learning Novel Skills from Language-Generated Demonstrations
by: Jin, Ao-Qun, et al.
Published: (2024)
by: Jin, Ao-Qun, et al.
Published: (2024)
Efficient Construction of Implicit Surface Models From a Single Image for Motion Generation
by: Chu, Wei-Teng, et al.
Published: (2025)
by: Chu, Wei-Teng, et al.
Published: (2025)
Motion Control in Multi-Rotor Aerial Robots Using Deep Reinforcement Learning
by: Shetty, Gaurav, et al.
Published: (2025)
by: Shetty, Gaurav, et al.
Published: (2025)
Similar Items
-
Integrating Controllable Motion Skills from Demonstrations
by: Liao, Honghao, et al.
Published: (2024) -
Representation Alignment from Human Feedback for Cross-Embodiment Reward Learning from Mixed-Quality Demonstrations
by: Mattson, Connor, et al.
Published: (2024) -
Leveraging Pre-trained Large Language Models with Refined Prompting for Online Task and Motion Planning
by: Guo, Huihui, et al.
Published: (2025) -
Learning Reward and Policy Jointly from Demonstration and Preference Improves Alignment
by: Li, Chenliang, et al.
Published: (2024) -
Generating Physically Realistic and Directable Human Motions from Multi-Modal Inputs
by: Shrestha, Aayam, et al.
Published: (2025)