Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Kevin, Scalise, Rosario, Winston, Cleah, Agrawal, Ayush, Zhang, Yunchu, Baijal, Rohan, Grotz, Markus, Boots, Byron, Burchfiel, Benjamin, Itkina, Masha, Shah, Paarth, Gupta, Abhishek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Sample Long Range Path Planning under Sensing Uncertainty for Off-Road Autonomous Driving
by: Schmittle, Matt, et al.
Published: (2024)
by: Schmittle, Matt, et al.
Published: (2024)
Long Range Navigator (LRN): Extending robot planning horizons beyond metric maps
by: Schmittle, Matt, et al.
Published: (2025)
by: Schmittle, Matt, et al.
Published: (2025)
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
by: Xu, Chen, et al.
Published: (2025)
by: Xu, Chen, et al.
Published: (2025)
Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets
by: Zhu, Chuning, et al.
Published: (2025)
by: Zhu, Chuning, et al.
Published: (2025)
CCIL: Continuity-based Data Augmentation for Corrective Imitation Learning
by: Ke, Liyiming, et al.
Published: (2023)
by: Ke, Liyiming, et al.
Published: (2023)
VAMOS: A Hierarchical Vision-Language-Action Model for Capability-Modulated and Steerable Navigation
by: Castro, Mateo Guaman, et al.
Published: (2025)
by: Castro, Mateo Guaman, et al.
Published: (2025)
How Generalizable Is My Behavior Cloning Policy? A Statistical Approach to Trustworthy Performance Evaluation
by: Vincent, Joseph A., et al.
Published: (2024)
by: Vincent, Joseph A., et al.
Published: (2024)
Model Predictive Adversarial Imitation Learning for Planning from Observation
by: Han, Tyler, et al.
Published: (2025)
by: Han, Tyler, et al.
Published: (2025)
Parental Guidance: Efficient Lifelong Learning through Evolutionary Distillation
by: Zhang, Octi, et al.
Published: (2025)
by: Zhang, Octi, et al.
Published: (2025)
Overcoming the Sim-to-Real Gap: Leveraging Simulation to Learn to Explore for Real-World RL
by: Wagenmaker, Andrew, et al.
Published: (2024)
by: Wagenmaker, Andrew, et al.
Published: (2024)
Planning from Observation and Interaction
by: Han, Tyler, et al.
Published: (2026)
by: Han, Tyler, et al.
Published: (2026)
Active Vision for Scene Understanding
by: Grotz, Markus
Published: (2022)
by: Grotz, Markus
Published: (2022)
LOPR: Latent Occupancy PRediction using Generative Models
by: Lange, Bernard, et al.
Published: (2022)
by: Lange, Bernard, et al.
Published: (2022)
Balance Equation-based Distributionally Robust Offline Imitation Learning
by: Agrawal, Rishabh, et al.
Published: (2025)
by: Agrawal, Rishabh, et al.
Published: (2025)
Dynamics Models in the Aggressive Off-Road Driving Regime
by: Han, Tyler, et al.
Published: (2024)
by: Han, Tyler, et al.
Published: (2024)
Robot Learning as an Empirical Science: Best Practices for Policy Evaluation
by: Kress-Gazit, Hadas, et al.
Published: (2024)
by: Kress-Gazit, Hadas, et al.
Published: (2024)
Is Your Imitation Learning Policy Better than Mine? Policy Comparison with Near-Optimal Stopping
by: Snyder, David, et al.
Published: (2025)
by: Snyder, David, et al.
Published: (2025)
Markov Balance Satisfaction Improves Performance in Strictly Batch Offline Imitation Learning
by: Agrawal, Rishabh, et al.
Published: (2024)
by: Agrawal, Rishabh, et al.
Published: (2024)
Dynamic Non-Prehensile Object Transport via Model-Predictive Reinforcement Learning
by: Jawale, Neel, et al.
Published: (2024)
by: Jawale, Neel, et al.
Published: (2024)
ATK: Automatic Task-driven Keypoint Selection for Robust Policy Learning
by: Zhang, Yunchu, et al.
Published: (2025)
by: Zhang, Yunchu, et al.
Published: (2025)
Self-supervised Multi-future Occupancy Forecasting for Autonomous Driving
by: Lange, Bernard, et al.
Published: (2024)
by: Lange, Bernard, et al.
Published: (2024)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
by: Hoang, Huy, et al.
Published: (2025)
by: Hoang, Huy, et al.
Published: (2025)
When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited
by: Agrawal, Rishabh, et al.
Published: (2026)
by: Agrawal, Rishabh, et al.
Published: (2026)
PerAct2: Benchmarking and Learning for Robotic Bimanual Manipulation Tasks
by: Grotz, Markus, et al.
Published: (2024)
by: Grotz, Markus, et al.
Published: (2024)
Wheeled Lab: Modern Sim2Real for Low-cost, Open-source Wheeled Robotics
by: Han, Tyler, et al.
Published: (2025)
by: Han, Tyler, et al.
Published: (2025)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
by: Hatch, Kyle B., et al.
Published: (2024)
by: Hatch, Kyle B., et al.
Published: (2024)
V-STRONG: Visual Self-Supervised Traversability Learning for Off-road Navigation
by: Jung, Sanghun, et al.
Published: (2023)
by: Jung, Sanghun, et al.
Published: (2023)
Offline Imitation Learning with Variational Counterfactual Reasoning
by: He, Bowei, et al.
Published: (2023)
by: He, Bowei, et al.
Published: (2023)
DITTO: Offline Imitation Learning with World Models
by: DeMoss, Branton, et al.
Published: (2023)
by: DeMoss, Branton, et al.
Published: (2023)
CUPID: Curating Data your Robot Loves with Influence Functions
by: Agia, Christopher, et al.
Published: (2025)
by: Agia, Christopher, et al.
Published: (2025)
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies
by: Patil, Sarvesh, et al.
Published: (2026)
by: Patil, Sarvesh, et al.
Published: (2026)
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
by: Corrado, Nicholas E., et al.
Published: (2023)
by: Corrado, Nicholas E., et al.
Published: (2023)
Steering Your Diffusion Policy with Latent Space Reinforcement Learning
by: Wagenmaker, Andrew, et al.
Published: (2025)
by: Wagenmaker, Andrew, et al.
Published: (2025)
A Study on Issues and Challenges of Women Empowerment in India
by: Sanjay Baijal
Published: (2017)
by: Sanjay Baijal
Published: (2017)
Consumer Behaviour towards Mobile Tele Services: A Case Study
by: Sanjay Baijal
Published: (2017)
by: Sanjay Baijal
Published: (2017)
A Study on the Efficacy of the Public Distribution System in India
by: Sanjay Baijal
Published: (2017)
by: Sanjay Baijal
Published: (2017)
Offline Imitation Learning with Model-based Reverse Augmentation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
Efficient Offline Reinforcement Learning: First Imitate, then Improve
by: Jelley, Adam, et al.
Published: (2024)
by: Jelley, Adam, et al.
Published: (2024)
How to Leverage Diverse Demonstrations in Offline Imitation Learning
by: Yue, Sheng, et al.
Published: (2024)
by: Yue, Sheng, et al.
Published: (2024)
Offline Imitation Learning Through Graph Search and Retrieval
by: Yin, Zhao-Heng, et al.
Published: (2024)
by: Yin, Zhao-Heng, et al.
Published: (2024)
Similar Items
-
Multi-Sample Long Range Path Planning under Sensing Uncertainty for Off-Road Autonomous Driving
by: Schmittle, Matt, et al.
Published: (2024) -
Long Range Navigator (LRN): Extending robot planning horizons beyond metric maps
by: Schmittle, Matt, et al.
Published: (2025) -
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
by: Xu, Chen, et al.
Published: (2025) -
Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets
by: Zhu, Chuning, et al.
Published: (2025) -
CCIL: Continuity-based Data Augmentation for Corrective Imitation Learning
by: Ke, Liyiming, et al.
Published: (2023)