Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Tian, Ran, Wu, Yilin, Xu, Chenfeng, Tomizuka, Masayoshi, Malik, Jitendra, Bajcsy, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What Matters to You? Towards Visual Representation Alignment for Robot Learning
by: Tian, Ran, et al.
Published: (2023)
by: Tian, Ran, et al.
Published: (2023)
Reimagination with Test-time Observation Interventions: Distractor-Robust World Model Predictions for Visual Model Predictive Control
by: Chen, Yuxin, et al.
Published: (2025)
by: Chen, Yuxin, et al.
Published: (2025)
Adapting by Analogy: OOD Generalization of Visuomotor Policies via Functional Correspondence
by: Gupta, Pranay, et al.
Published: (2025)
by: Gupta, Pranay, et al.
Published: (2025)
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
by: Wu, Yilin, et al.
Published: (2025)
by: Wu, Yilin, et al.
Published: (2025)
Position: Good Embodied Reward Models Need Bad Behavior Data
by: Tian, Ran, et al.
Published: (2026)
by: Tian, Ran, et al.
Published: (2026)
Fast Visuomotor Policy for Robotic Manipulation
by: Jia, Jingkai, et al.
Published: (2025)
by: Jia, Jingkai, et al.
Published: (2025)
NeRFs in Robotics: A Survey
by: Wang, Guangming, et al.
Published: (2024)
by: Wang, Guangming, et al.
Published: (2024)
DSLO: Deep Sequence LiDAR Odometry Based on Inconsistent Spatio-temporal Propagation
by: Zhang, Huixin, et al.
Published: (2024)
by: Zhang, Huixin, et al.
Published: (2024)
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024)
by: Zhao, Yinuo, et al.
Published: (2024)
Multi-Camera View Scaling for Data-Efficient Robot Imitation Learning
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
ROCKET-2: Steering Visuomotor Policy via Cross-View Goal Alignment
by: Cai, Shaofei, et al.
Published: (2025)
by: Cai, Shaofei, et al.
Published: (2025)
FLASH: Efficient Visuomotor Policy via Sparse Sampling
by: Bai, Jiaqi, et al.
Published: (2026)
by: Bai, Jiaqi, et al.
Published: (2026)
H$^3$DP: Triply-Hierarchical Diffusion Policy for Visuomotor Learning
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
Learning Predictive Visuomotor Coordination
by: Jia, Wenqi, et al.
Published: (2025)
by: Jia, Wenqi, et al.
Published: (2025)
Referring-Aware Visuomotor Policy Learning for Closed-Loop Manipulation
by: Ma, Jiahua, et al.
Published: (2026)
by: Ma, Jiahua, et al.
Published: (2026)
The Sound of Simulation: Learning Multimodal Sim-to-Real Robot Policies with Generative Audio
by: Wang, Renhao, et al.
Published: (2025)
by: Wang, Renhao, et al.
Published: (2025)
Rodrigues Network for Learning Robot Actions
by: Zhang, Jialiang, et al.
Published: (2025)
by: Zhang, Jialiang, et al.
Published: (2025)
Dreamitate: Real-World Visuomotor Policy Learning via Video Generation
by: Liang, Junbang, et al.
Published: (2024)
by: Liang, Junbang, et al.
Published: (2024)
CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction
by: Gong, Zhefei, et al.
Published: (2024)
by: Gong, Zhefei, et al.
Published: (2024)
GRAPE: Generalizing Robot Policy via Preference Alignment
by: Zhang, Zijian, et al.
Published: (2024)
by: Zhang, Zijian, et al.
Published: (2024)
CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion
by: Ma, Jiahua, et al.
Published: (2025)
by: Ma, Jiahua, et al.
Published: (2025)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
by: Liu, Yijun, et al.
Published: (2025)
by: Liu, Yijun, et al.
Published: (2025)
DINOv3-Diffusion Policy: Self-Supervised Large Visual Model for Visuomotor Diffusion Policy Learning
by: Egbe, ThankGod, et al.
Published: (2025)
by: Egbe, ThankGod, et al.
Published: (2025)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
by: Wang, Yixiao, et al.
Published: (2024)
by: Wang, Yixiao, et al.
Published: (2024)
Seeing to Act, Prompting to Specify: A Bayesian Factorization of Vision Language Action Policy
by: Xu, Kechun, et al.
Published: (2025)
by: Xu, Kechun, et al.
Published: (2025)
3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
by: Ze, Yanjie, et al.
Published: (2024)
by: Ze, Yanjie, et al.
Published: (2024)
When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering
by: Yuan, Jessie, et al.
Published: (2026)
by: Yuan, Jessie, et al.
Published: (2026)
Agent-to-Sim: Learning Interactive Behavior Models from Casual Longitudinal Videos
by: Yang, Gengshan, et al.
Published: (2024)
by: Yang, Gengshan, et al.
Published: (2024)
FreqPolicy: Frequency Autoregressive Visuomotor Policy with Continuous Tokens
by: Zhong, Yiming, et al.
Published: (2025)
by: Zhong, Yiming, et al.
Published: (2025)
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
by: Xu, Yifu, et al.
Published: (2026)
by: Xu, Yifu, et al.
Published: (2026)
Adaptive Human Trajectory Prediction via Latent Corridors
by: Thakkar, Neerja, et al.
Published: (2023)
by: Thakkar, Neerja, et al.
Published: (2023)
Testing Human-Hand Segmentation on In-Distribution and Out-of-Distribution Data in Human-Robot Interactions Using a Deep Ensemble Model
by: Jalayer, Reza, et al.
Published: (2025)
by: Jalayer, Reza, et al.
Published: (2025)
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
by: Yuan, Puzhen, et al.
Published: (2025)
by: Yuan, Puzhen, et al.
Published: (2025)
ViTaS: Visual Tactile Soft Fusion Contrastive Learning for Visuomotor Learning
by: Tian, Yufeng, et al.
Published: (2026)
by: Tian, Yufeng, et al.
Published: (2026)
StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement
by: Seo, Junwon, et al.
Published: (2026)
by: Seo, Junwon, et al.
Published: (2026)
Crossway Diffusion: Improving Diffusion-based Visuomotor Policy via Self-supervised Learning
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Video Diffusion Alignment via Reward Gradients
by: Prabhudesai, Mihir, et al.
Published: (2024)
by: Prabhudesai, Mihir, et al.
Published: (2024)
Collaborative Representation Learning for Alignment of Tactile, Language, and Vision Modalities
by: Zhou, Yiyun, et al.
Published: (2025)
by: Zhou, Yiyun, et al.
Published: (2025)
Bridging the Sim2Real Gap: Vision Encoder Pre-Training for Visuomotor Policy Transfer
by: Yardi, Yash, et al.
Published: (2025)
by: Yardi, Yash, et al.
Published: (2025)
PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models
by: Guo, Dingkun, et al.
Published: (2024)
by: Guo, Dingkun, et al.
Published: (2024)
Similar Items
-
What Matters to You? Towards Visual Representation Alignment for Robot Learning
by: Tian, Ran, et al.
Published: (2023) -
Reimagination with Test-time Observation Interventions: Distractor-Robust World Model Predictions for Visual Model Predictive Control
by: Chen, Yuxin, et al.
Published: (2025) -
Adapting by Analogy: OOD Generalization of Visuomotor Policies via Functional Correspondence
by: Gupta, Pranay, et al.
Published: (2025) -
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
by: Wu, Yilin, et al.
Published: (2025) -
Position: Good Embodied Reward Models Need Bad Behavior Data
by: Tian, Ran, et al.
Published: (2026)