FDPP: Fine-tune Diffusion Policy with Human Preference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yuxin, Jha, Devesh K., Tomizuka, Masayoshi, Romeres, Diego |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Pivoting Manipulation with Force and Vision Feedback Using Optimization-based Demonstrations
von: Shirai, Yuki, et al.
Veröffentlicht: (2025)
von: Shirai, Yuki, et al.
Veröffentlicht: (2025)
PPGuide: Steering Diffusion Policies with Performance Predictive Guidance
von: Wang, Zixing, et al.
Veröffentlicht: (2026)
von: Wang, Zixing, et al.
Veröffentlicht: (2026)
DADP: Domain Adaptive Diffusion Policy
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
Testing Human-Hand Segmentation on In-Distribution and Out-of-Distribution Data in Human-Robot Interactions Using a Deep Ensemble Model
von: Jalayer, Reza, et al.
Veröffentlicht: (2025)
von: Jalayer, Reza, et al.
Veröffentlicht: (2025)
Robust In-Hand Manipulation with Extrinsic Contacts
von: Liang, Boyuan, et al.
Veröffentlicht: (2024)
von: Liang, Boyuan, et al.
Veröffentlicht: (2024)
BeTAIL: Behavior Transformer Adversarial Imitation Learning from Human Racing Gameplay
von: Weaver, Catherine, et al.
Veröffentlicht: (2024)
von: Weaver, Catherine, et al.
Veröffentlicht: (2024)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
RecoveryChaining: Learning Local Recovery Policies for Robust Manipulation
von: Vats, Shivam, et al.
Veröffentlicht: (2024)
von: Vats, Shivam, et al.
Veröffentlicht: (2024)
Grounded Relational Inference: Domain Knowledge Driven Explainable Autonomous Driving
von: Tang, Chen, et al.
Veröffentlicht: (2021)
von: Tang, Chen, et al.
Veröffentlicht: (2021)
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm
von: Kim, Hyeonjun, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonjun, et al.
Veröffentlicht: (2025)
Behavioral Mode Discovery for Fine-tuning Multimodal Generative Policies
von: Longhini, Alberta, et al.
Veröffentlicht: (2026)
von: Longhini, Alberta, et al.
Veröffentlicht: (2026)
REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning
von: Gu, Zhaoyuan, et al.
Veröffentlicht: (2026)
von: Gu, Zhaoyuan, et al.
Veröffentlicht: (2026)
SkillDiffuser: Interpretable Hierarchical Planning via Skill Abstractions in Diffusion-Based Task Execution
von: Liang, Zhixuan, et al.
Veröffentlicht: (2023)
von: Liang, Zhixuan, et al.
Veröffentlicht: (2023)
FOSP: Fine-tuning Offline Safe Policy through World Models
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
von: Zhang, Tonghe, et al.
Veröffentlicht: (2025)
von: Zhang, Tonghe, et al.
Veröffentlicht: (2025)
RoDiF: Robust Direct Fine-Tuning of Diffusion Policies with Corrupted Human Feedback
von: Vatsa, Amitesh, et al.
Veröffentlicht: (2026)
von: Vatsa, Amitesh, et al.
Veröffentlicht: (2026)
LANGTRAJ: Diffusion Model and Dataset for Language-Conditioned Trajectory Simulation
von: Chang, Wei-Jer, et al.
Veröffentlicht: (2025)
von: Chang, Wei-Jer, et al.
Veröffentlicht: (2025)
CLAW: Composable Language-Annotated Whole-body Motion Generation
von: Cao, Jianuo, et al.
Veröffentlicht: (2026)
von: Cao, Jianuo, et al.
Veröffentlicht: (2026)
Embedding Morphology into Transformers for Cross-Robot Policy Learning
von: Suzuki, Kei, et al.
Veröffentlicht: (2026)
von: Suzuki, Kei, et al.
Veröffentlicht: (2026)
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
von: Tian, Ran, et al.
Veröffentlicht: (2024)
von: Tian, Ran, et al.
Veröffentlicht: (2024)
Adaptive Linear Path Model-Based Diffusion
von: Shimizu, Yutaka, et al.
Veröffentlicht: (2026)
von: Shimizu, Yutaka, et al.
Veröffentlicht: (2026)
A Black-Box Physics-Informed Estimator based on Gaussian Process Regression for Robot Inverse Dynamics Identification
von: Giacomuzzos, Giulio, et al.
Veröffentlicht: (2023)
von: Giacomuzzos, Giulio, et al.
Veröffentlicht: (2023)
Bootstrap Off-policy with World Model
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
Latent Embedding Adaptation for Human Preference Alignment in Diffusion Planners
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2025)
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2025)
DexHandDiff: Interaction-aware Diffusion Planning for Adaptive Dexterous Manipulation
von: Liang, Zhixuan, et al.
Veröffentlicht: (2024)
von: Liang, Zhixuan, et al.
Veröffentlicht: (2024)
LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines
von: Cherian, Anoop, et al.
Veröffentlicht: (2024)
von: Cherian, Anoop, et al.
Veröffentlicht: (2024)
Learning global control of underactuated systems with Model-Based Reinforcement Learning
von: Turcato, Niccolò, et al.
Veröffentlicht: (2025)
von: Turcato, Niccolò, et al.
Veröffentlicht: (2025)
Uncertainty Comes for Free: Human-in-the-Loop Policies with Diffusion Models
von: He, Zhanpeng, et al.
Veröffentlicht: (2025)
von: He, Zhanpeng, et al.
Veröffentlicht: (2025)
DemoDiffusion: One-Shot Human Imitation using pre-trained Diffusion Policy
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
Tactile Estimation of Extrinsic Contact Patch for Stable Placement
von: Ota, Kei, et al.
Veröffentlicht: (2023)
von: Ota, Kei, et al.
Veröffentlicht: (2023)
Diffusion Policy through Conditional Proximal Policy Optimization
von: Liu, Ben, et al.
Veröffentlicht: (2026)
von: Liu, Ben, et al.
Veröffentlicht: (2026)
Diffusion Policy Policy Optimization
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
Dichotomous Diffusion Policy Optimization
von: Liang, Ruiming, et al.
Veröffentlicht: (2025)
von: Liang, Ruiming, et al.
Veröffentlicht: (2025)
CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies
von: Chen, Keyu, et al.
Veröffentlicht: (2026)
von: Chen, Keyu, et al.
Veröffentlicht: (2026)
Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning
von: Park, Byoungwoo, et al.
Veröffentlicht: (2026)
von: Park, Byoungwoo, et al.
Veröffentlicht: (2026)
Equivariant Diffusion Policy
von: Wang, Dian, et al.
Veröffentlicht: (2024)
von: Wang, Dian, et al.
Veröffentlicht: (2024)
One-Step Diffusion Policy: Fast Visuomotor Policies via Diffusion Distillation
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
Residual Policy Gradient: A Reward View of KL-regularized Objective
von: Wang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2025)
Active Fine-Tuning of Multi-Task Policies
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
MEReQ: Max-Ent Residual-Q Inverse RL for Sample-Efficient Alignment from Intervention
von: Chen, Yuxin, et al.
Veröffentlicht: (2024)
von: Chen, Yuxin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning Pivoting Manipulation with Force and Vision Feedback Using Optimization-based Demonstrations
von: Shirai, Yuki, et al.
Veröffentlicht: (2025) -
PPGuide: Steering Diffusion Policies with Performance Predictive Guidance
von: Wang, Zixing, et al.
Veröffentlicht: (2026) -
DADP: Domain Adaptive Diffusion Policy
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026) -
Testing Human-Hand Segmentation on In-Distribution and Out-of-Distribution Data in Human-Robot Interactions Using a Deep Ensemble Model
von: Jalayer, Reza, et al.
Veröffentlicht: (2025) -
Robust In-Hand Manipulation with Extrinsic Contacts
von: Liang, Boyuan, et al.
Veröffentlicht: (2024)