Good in Bad (GiB): Sifting Through End-user Demonstrations for Learning a Better Policy
Fuente:
arXiv
Saved in:
| Main Authors: | Sojib, Noushad, Ghattas, Ola, Begum, Momotaz |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Efficient Metric for Data Quality Measurement in Imitation Learning
by: Sojib, Noushad, et al.
Published: (2026)
by: Sojib, Noushad, et al.
Published: (2026)
To Do or Not to Do: Ensuring the Safety of Visuomotor Policies Learned from Demonstrations
by: Ahmed, Riad, et al.
Published: (2026)
by: Ahmed, Riad, et al.
Published: (2026)
A Principled Approach for Creating High-fidelity Synthetic Demonstrations for Imitation Learning
by: Akash, Moniruzzaman, et al.
Published: (2026)
by: Akash, Moniruzzaman, et al.
Published: (2026)
TAIL-Safe: Task-Agnostic Safety Monitoring for Imitation Learning Policies
by: Ahmed, Riad, et al.
Published: (2026)
by: Ahmed, Riad, et al.
Published: (2026)
Trajectory-Consistent Flow Matching for Robust Visuomotor Policy Learning
by: Ahmed, Riad, et al.
Published: (2026)
by: Ahmed, Riad, et al.
Published: (2026)
Probabilistic Safety Guarantee for Stochastic Control Systems Using Average Reward MDPs
by: Omidi, Saber, et al.
Published: (2025)
by: Omidi, Saber, et al.
Published: (2025)
Position: Good Embodied Reward Models Need Bad Behavior Data
by: Tian, Ran, et al.
Published: (2026)
by: Tian, Ran, et al.
Published: (2026)
Instrumentation for Better Demonstrations: A Case Study
by: Proesmans, Remko, et al.
Published: (2025)
by: Proesmans, Remko, et al.
Published: (2025)
Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt
by: Zhu, Xiang, et al.
Published: (2025)
by: Zhu, Xiang, et al.
Published: (2025)
End-to-End Multi-Task Policy Learning from NMPC for Quadruped Locomotion
by: Sajja, Anudeep, et al.
Published: (2025)
by: Sajja, Anudeep, et al.
Published: (2025)
GiAnt: A Bio-Inspired Hexapod for Adaptive Terrain Navigation and Object Detection
by: Bhuiyan, Aasfee Mosharraf, et al.
Published: (2025)
by: Bhuiyan, Aasfee Mosharraf, et al.
Published: (2025)
DIRIGENt: End-To-End Robotic Imitation of Human Demonstrations Based on a Diffusion Model
by: Spisak, Josua, et al.
Published: (2025)
by: Spisak, Josua, et al.
Published: (2025)
Domain Adaptation of Visual Policies with a Single Demonstration
by: Wang, Weiyao, et al.
Published: (2024)
by: Wang, Weiyao, et al.
Published: (2024)
Talk Through It: End User Directed Manipulation Learning
by: Winge, Carl, et al.
Published: (2024)
by: Winge, Carl, et al.
Published: (2024)
DemoGen: Synthetic Demonstration Generation for Data-Efficient Visuomotor Policy Learning
by: Xue, Zhengrong, et al.
Published: (2025)
by: Xue, Zhengrong, et al.
Published: (2025)
State-Conditional Adversarial Learning: An Off-Policy Visual Domain Transfer Method for End-to-End Imitation Learning
by: Liu, Yuxiang, et al.
Published: (2025)
by: Liu, Yuxiang, et al.
Published: (2025)
Raising Body Ownership in End-to-End Visuomotor Policy Learning via Robot-Centric Pooling
by: Zhuang, Zheyu, et al.
Published: (2024)
by: Zhuang, Zheyu, et al.
Published: (2024)
SAIL: Faster-than-Demonstration Execution of Imitation Learning Policies
by: Arachchige, Nadun Ranawaka, et al.
Published: (2025)
by: Arachchige, Nadun Ranawaka, et al.
Published: (2025)
WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations
by: Freeman, Harry, et al.
Published: (2026)
by: Freeman, Harry, et al.
Published: (2026)
End-to-End Humanoid Robot Safe and Comfortable Locomotion Policy
by: Wang, Zifan, et al.
Published: (2025)
by: Wang, Zifan, et al.
Published: (2025)
Gentle Manipulation Policy Learning via Demonstrations from VLM Planned Atomic Skills
by: Zhou, Jiayu, et al.
Published: (2025)
by: Zhou, Jiayu, et al.
Published: (2025)
Embodiment-Agnostic Navigation Policy Trained with Visual Demonstrations
by: Curtis, Nimrod, et al.
Published: (2024)
by: Curtis, Nimrod, et al.
Published: (2024)
BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies
by: Singh, Anya, et al.
Published: (2026)
by: Singh, Anya, et al.
Published: (2026)
Unified Humanoid Fall-Safety Policy from a Few Demonstrations
by: Xu, Zhengjie, et al.
Published: (2025)
by: Xu, Zhengjie, et al.
Published: (2025)
Learning Diffusion Policies from Demonstrations For Compliant Contact-rich Manipulation
by: Aburub, Malek, et al.
Published: (2024)
by: Aburub, Malek, et al.
Published: (2024)
SINGER: An Onboard Generalist Vision-Language Navigation Policy for Drones
by: Adang, Maximilian, et al.
Published: (2025)
by: Adang, Maximilian, et al.
Published: (2025)
Learning from Demonstration with Hierarchical Policy Abstractions Toward High-Performance and Courteous Autonomous Racing
by: Chung, Chanyoung, et al.
Published: (2024)
by: Chung, Chanyoung, et al.
Published: (2024)
From a Single Demonstration to a General Policy for Contact-Rich Manipulation
by: Li, Xing, et al.
Published: (2026)
by: Li, Xing, et al.
Published: (2026)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
by: Zhu, Minjie, et al.
Published: (2025)
by: Zhu, Minjie, et al.
Published: (2025)
Robotic System for Chemical Experiment Automation with Dual Demonstration of End-effector and Jig Operations
by: Sasaki, Hikaru, et al.
Published: (2025)
by: Sasaki, Hikaru, et al.
Published: (2025)
Demonstration Based Explainable AI for Learning from Demonstration Methods
by: Gu, Morris, et al.
Published: (2024)
by: Gu, Morris, et al.
Published: (2024)
Is Your Imitation Learning Policy Better than Mine? Policy Comparison with Near-Optimal Stopping
by: Snyder, David, et al.
Published: (2025)
by: Snyder, David, et al.
Published: (2025)
DockAnywhere: Data-Efficient Visuomotor Policy Learning for Mobile Manipulation via Novel Demonstration Generation
by: Shan, Ziyu, et al.
Published: (2026)
by: Shan, Ziyu, et al.
Published: (2026)
SADP: Subgoal-Aware Diffusion Policy for Explainable Robots Learned from Foundation Model Generated Demonstrations
by: Hu, Site, et al.
Published: (2026)
by: Hu, Site, et al.
Published: (2026)
RwoR: Generating Robot Demonstrations from Human Hand Collection for Policy Learning without Robot
by: Heng, Liang, et al.
Published: (2025)
by: Heng, Liang, et al.
Published: (2025)
SuFIA-BC: Generating High Quality Demonstration Data for Visuomotor Policy Learning in Surgical Subtasks
by: Moghani, Masoud, et al.
Published: (2025)
by: Moghani, Masoud, et al.
Published: (2025)
DVDP: An End-to-End Policy for Mobile Robot Visual Docking with RGB-D Perception
by: Min, Haohan, et al.
Published: (2025)
by: Min, Haohan, et al.
Published: (2025)
Action Images: End-to-End Policy Learning via Multiview Video Generation
by: Zhen, Haoyu, et al.
Published: (2026)
by: Zhen, Haoyu, et al.
Published: (2026)
DiffE2E: Rethinking End-to-End Driving with a Hybrid Action Diffusion and Supervised Policy
by: Zhao, Rui, et al.
Published: (2025)
by: Zhao, Rui, et al.
Published: (2025)
Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?
by: Onoda, Ku, et al.
Published: (2026)
by: Onoda, Ku, et al.
Published: (2026)
Similar Items
-
An Efficient Metric for Data Quality Measurement in Imitation Learning
by: Sojib, Noushad, et al.
Published: (2026) -
To Do or Not to Do: Ensuring the Safety of Visuomotor Policies Learned from Demonstrations
by: Ahmed, Riad, et al.
Published: (2026) -
A Principled Approach for Creating High-fidelity Synthetic Demonstrations for Imitation Learning
by: Akash, Moniruzzaman, et al.
Published: (2026) -
TAIL-Safe: Task-Agnostic Safety Monitoring for Imitation Learning Policies
by: Ahmed, Riad, et al.
Published: (2026) -
Trajectory-Consistent Flow Matching for Robust Visuomotor Policy Learning
by: Ahmed, Riad, et al.
Published: (2026)