Is Your Imitation Learning Policy Better than Mine? Policy Comparison with Near-Optimal Stopping
Fuente:
arXiv
Saved in:
| Main Authors: | Snyder, David, Hancock, Asher James, Badithela, Apurva, Dixon, Emma, Miller, Patrick, Ambrus, Rares Andrei, Majumdar, Anirudha, Itkina, Masha, Nishimura, Haruki |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Binary Success: Sample-Efficient and Statistically Rigorous Robot Policy Comparison
by: Snyder, David, et al.
Published: (2026)
by: Snyder, David, et al.
Published: (2026)
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
by: Xu, Chen, et al.
Published: (2025)
by: Xu, Chen, et al.
Published: (2025)
Reliable and Scalable Robot Policy Evaluation with Imperfect Simulators
by: Badithela, Apurva, et al.
Published: (2025)
by: Badithela, Apurva, et al.
Published: (2025)
How Generalizable Is My Behavior Cloning Policy? A Statistical Approach to Trustworthy Performance Evaluation
by: Vincent, Joseph A., et al.
Published: (2024)
by: Vincent, Joseph A., et al.
Published: (2024)
STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation
by: Goli, Hossein, et al.
Published: (2025)
by: Goli, Hossein, et al.
Published: (2025)
Run-time Observation Interventions Make Vision-Language-Action Models More Visually Robust
by: Hancock, Asher J., et al.
Published: (2024)
by: Hancock, Asher J., et al.
Published: (2024)
Explore until Confident: Efficient Exploration for Embodied Question Answering
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
Video Generation Models in Robotics -- Applications, Research Challenges, Future Directions
by: Mei, Zhiting, et al.
Published: (2026)
by: Mei, Zhiting, et al.
Published: (2026)
Actions as Language: Fine-Tuning VLMs into VLAs Without Catastrophic Forgetting
by: Hancock, Asher J., et al.
Published: (2025)
by: Hancock, Asher J., et al.
Published: (2025)
SAFE: Multitask Failure Detection for Vision-Language-Action Models
by: Gu, Qiao, et al.
Published: (2025)
by: Gu, Qiao, et al.
Published: (2025)
Guiding Data Collection via Factored Scaling Curves
by: Zha, Lihan, et al.
Published: (2025)
by: Zha, Lihan, et al.
Published: (2025)
CUPID: Curating Data your Robot Loves with Influence Functions
by: Agia, Christopher, et al.
Published: (2025)
by: Agia, Christopher, et al.
Published: (2025)
Video Generators are Robot Policies
by: Liang, Junbang, et al.
Published: (2025)
by: Liang, Junbang, et al.
Published: (2025)
Espresso: High Compression For Rich Extraction From Videos for Your Vision-Language Model
by: Yu, Keunwoo Peter, et al.
Published: (2024)
by: Yu, Keunwoo Peter, et al.
Published: (2024)
Deceptive Risk Minimization: Out-of-Distribution Generalization by Deceiving Distribution Shift Detectors
by: Majumdar, Anirudha
Published: (2025)
by: Majumdar, Anirudha
Published: (2025)
LOPR: Latent Occupancy PRediction using Generative Models
by: Lange, Bernard, et al.
Published: (2022)
by: Lange, Bernard, et al.
Published: (2022)
Impact of Different Failures on a Robot's Perceived Reliability
by: Violette, Andrew, et al.
Published: (2026)
by: Violette, Andrew, et al.
Published: (2026)
Privacy-Preserving Map-Free Exploration for Confirming the Absence of a Radioactive Source
by: Lepowsky, Eric, et al.
Published: (2024)
by: Lepowsky, Eric, et al.
Published: (2024)
PlayWorld: Learning Robot World Models from Autonomous Play
by: Yin, Tenny, et al.
Published: (2026)
by: Yin, Tenny, et al.
Published: (2026)
Diffusion Policy Policy Optimization
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
Predictive Red Teaming: Breaking Policies Without Breaking Robots
by: Majumdar, Anirudha, et al.
Published: (2025)
by: Majumdar, Anirudha, et al.
Published: (2025)
Self-supervised Multi-future Occupancy Forecasting for Autonomous Driving
by: Lange, Bernard, et al.
Published: (2024)
by: Lange, Bernard, et al.
Published: (2024)
SAIL: Faster-than-Demonstration Execution of Imitation Learning Policies
by: Arachchige, Nadun Ranawaka, et al.
Published: (2025)
by: Arachchige, Nadun Ranawaka, et al.
Published: (2025)
LAP: Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transfer
by: Zha, Lihan, et al.
Published: (2026)
by: Zha, Lihan, et al.
Published: (2026)
Australian Wage Policy: Infancy and Adolescence
by: Hancock, Keith
Published: (2015)
by: Hancock, Keith
Published: (2015)
Generalized Early Stopping in Evolutionary Direct Policy Search
by: Arza, Etor, et al.
Published: (2023)
by: Arza, Etor, et al.
Published: (2023)
My Statistics is Better than Yours
by: Benhaïem, Simon
Published: (2024)
by: Benhaïem, Simon
Published: (2024)
ReFiNe: Recursive Field Networks for Cross-modal Multi-scene Representation
by: Zakharov, Sergey, et al.
Published: (2024)
by: Zakharov, Sergey, et al.
Published: (2024)
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
by: Guizilini, Vitor, et al.
Published: (2024)
by: Guizilini, Vitor, et al.
Published: (2024)
Residual Q-Learning: Offline and Online Policy Customization without Value
by: Li, Chenran, et al.
Published: (2023)
by: Li, Chenran, et al.
Published: (2023)
The Impact of Class Uncertainty Propagation in Perception-Based Motion Planning
by: Shah, Jibran Iqbal, et al.
Published: (2026)
by: Shah, Jibran Iqbal, et al.
Published: (2026)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
by: Huang, Kevin, et al.
Published: (2025)
by: Huang, Kevin, et al.
Published: (2025)
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
SciFi-Benchmark: Leveraging Science Fiction To Improve Robot Behavior
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
Geometry-aware Policy Imitation
by: Li, Yiming, et al.
Published: (2025)
by: Li, Yiming, et al.
Published: (2025)
"Better Ask for Forgiveness than Permission": Practices and Policies of AI Disclosure in Freelance Work
by: Hwang, Angel Hsing-Chi, et al.
Published: (2026)
by: Hwang, Angel Hsing-Chi, et al.
Published: (2026)
wd1: Weighted Policy Optimization for Reasoning in Diffusion Language Models
by: Tang, Xiaohang, et al.
Published: (2025)
by: Tang, Xiaohang, et al.
Published: (2025)
Career Don’t Stop Believing: Career Empowerment as a Mediator between Hope and Organizational Outcomes
by: Daphna Shwartz-Asher
Published: (2023)
by: Daphna Shwartz-Asher
Published: (2023)
Globally Stable Neural Imitation Policies
by: Abyaneh, Amin, et al.
Published: (2024)
by: Abyaneh, Amin, et al.
Published: (2024)
Similar Items
-
Beyond Binary Success: Sample-Efficient and Statistically Rigorous Robot Policy Comparison
by: Snyder, David, et al.
Published: (2026) -
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
by: Xu, Chen, et al.
Published: (2025) -
Reliable and Scalable Robot Policy Evaluation with Imperfect Simulators
by: Badithela, Apurva, et al.
Published: (2025) -
How Generalizable Is My Behavior Cloning Policy? A Statistical Approach to Trustworthy Performance Evaluation
by: Vincent, Joseph A., et al.
Published: (2024) -
STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation
by: Goli, Hossein, et al.
Published: (2025)