NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Albaba, Mert, Li, Chenhao, Diomataris, Markos, Taheri, Omid, Krause, Andreas, Black, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Moving by Looking: Towards Vision-Driven Avatar Motion Generation
by: Diomataris, Markos, et al.
Published: (2025)
by: Diomataris, Markos, et al.
Published: (2025)
WANDR: Intention-guided Human Motion Generation
by: Diomataris, Markos, et al.
Published: (2024)
by: Diomataris, Markos, et al.
Published: (2024)
MotionFix: Text-Driven 3D Human Motion Editing
by: Athanasiou, Nikos, et al.
Published: (2024)
by: Athanasiou, Nikos, et al.
Published: (2024)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
by: Yin, Zhenhan, et al.
Published: (2025)
by: Yin, Zhenhan, et al.
Published: (2025)
VILP: Imitation Learning with Latent Video Planning
by: Xu, Zhengtong, et al.
Published: (2025)
by: Xu, Zhengtong, et al.
Published: (2025)
Robust Instant Policy: Leveraging Student's t-Regression Model for Robust In-context Imitation Learning of Robot Manipulation
by: Oh, Hanbit, et al.
Published: (2025)
by: Oh, Hanbit, et al.
Published: (2025)
3D Whole-body Grasp Synthesis with Directional Controllability
by: Paschalidis, Georgios, et al.
Published: (2024)
by: Paschalidis, Georgios, et al.
Published: (2024)
EgoMimic: Scaling Imitation Learning via Egocentric Video
by: Kareer, Simar, et al.
Published: (2024)
by: Kareer, Simar, et al.
Published: (2024)
Multi-Transmotion: Pre-trained Model for Human Motion Prediction
by: Gao, Yang, et al.
Published: (2024)
by: Gao, Yang, et al.
Published: (2024)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
by: Zhang, Junjie, et al.
Published: (2024)
by: Zhang, Junjie, et al.
Published: (2024)
Stem-OB: Generalizable Visual Imitation Learning with Stem-Like Convergent Observation through Diffusion Inversion
by: Hu, Kaizhe, et al.
Published: (2024)
by: Hu, Kaizhe, et al.
Published: (2024)
RoboMirror: Understand Before You Imitate for Video to Humanoid Locomotion
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
One-shot Video Imitation via Parameterized Symbolic Abstraction Graphs
by: Wang, Jianren, et al.
Published: (2024)
by: Wang, Jianren, et al.
Published: (2024)
FUSION: Full-Body Unified Motion Prior for Body and Hands via Diffusion
by: Duran, Enes, et al.
Published: (2026)
by: Duran, Enes, et al.
Published: (2026)
UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations
by: Kim, Hanjung, et al.
Published: (2025)
by: Kim, Hanjung, et al.
Published: (2025)
One-Shot Dual-Arm Imitation Learning
by: Wang, Yilong, et al.
Published: (2025)
by: Wang, Yilong, et al.
Published: (2025)
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
by: Barcellona, Leonardo, et al.
Published: (2024)
by: Barcellona, Leonardo, et al.
Published: (2024)
DIVER: Reinforced Diffusion Breaks Imitation Bottlenecks in End-to-End Autonomous Driving
by: Song, Ziying, et al.
Published: (2025)
by: Song, Ziying, et al.
Published: (2025)
GASP: Unifying Geometric and Semantic Self-Supervised Pre-training for Autonomous Driving
by: Ljungbergh, William, et al.
Published: (2025)
by: Ljungbergh, William, et al.
Published: (2025)
RILe: Reinforced Imitation Learning
by: Albaba, Mert, et al.
Published: (2024)
by: Albaba, Mert, et al.
Published: (2024)
Slot-Level Robotic Placement via Visual Imitation from Single Human Video
by: Shan, Dandan, et al.
Published: (2025)
by: Shan, Dandan, et al.
Published: (2025)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Failure Identification in Imitation Learning Via Statistical and Semantic Filtering
by: Rolland, Quentin, et al.
Published: (2026)
by: Rolland, Quentin, et al.
Published: (2026)
Towards Fusing Point Cloud and Visual Representations for Imitation Learning
by: Donat, Atalay, et al.
Published: (2025)
by: Donat, Atalay, et al.
Published: (2025)
Lifelong Imitation Learning with Multimodal Latent Replay and Incremental Adjustment
by: Yu, Fanqi, et al.
Published: (2026)
by: Yu, Fanqi, et al.
Published: (2026)
Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained Transformers
by: Wang, Lirui, et al.
Published: (2024)
by: Wang, Lirui, et al.
Published: (2024)
Imitating What Works: Simulation-Filtered Modular Policy Learning from Human Videos
by: Zhai, Albert J., et al.
Published: (2026)
by: Zhai, Albert J., et al.
Published: (2026)
Predicting 4D Hand Trajectory from Monocular Videos
by: Ye, Yufei, et al.
Published: (2025)
by: Ye, Yufei, et al.
Published: (2025)
Mitigating the Human-Robot Domain Discrepancy in Visual Pre-training for Robotic Manipulation
by: Zhou, Jiaming, et al.
Published: (2024)
by: Zhou, Jiaming, et al.
Published: (2024)
LUMOS: Language-Conditioned Imitation Learning with World Models
by: Nematollahi, Iman, et al.
Published: (2025)
by: Nematollahi, Iman, et al.
Published: (2025)
FUNCTO: Function-Centric One-Shot Imitation Learning for Tool Manipulation
by: Tang, Chao, et al.
Published: (2025)
by: Tang, Chao, et al.
Published: (2025)
Contrastive Imitation Learning for Language-guided Multi-Task Robotic Manipulation
by: Ma, Teli, et al.
Published: (2024)
by: Ma, Teli, et al.
Published: (2024)
Learning Sidewalk Autopilot from Multi-Scale Imitation with Corrective Behavior Expansion
by: He, Honglin, et al.
Published: (2026)
by: He, Honglin, et al.
Published: (2026)
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
by: Patel, Shivansh, et al.
Published: (2025)
by: Patel, Shivansh, et al.
Published: (2025)
DITTO: Demonstration Imitation by Trajectory Transformation
by: Heppert, Nick, et al.
Published: (2024)
by: Heppert, Nick, et al.
Published: (2024)
Beyond Imitation: Learning Safe End-to-End Autonomous Driving from Hard Negatives
by: Wang, Junli, et al.
Published: (2026)
by: Wang, Junli, et al.
Published: (2026)
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model
by: Li, Peiyan, et al.
Published: (2026)
by: Li, Peiyan, et al.
Published: (2026)
Instant Policy: In-Context Imitation Learning via Graph Diffusion
by: Vosylius, Vitalis, et al.
Published: (2024)
by: Vosylius, Vitalis, et al.
Published: (2024)
Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning
by: Zhang, Xiaoyu, et al.
Published: (2024)
by: Zhang, Xiaoyu, et al.
Published: (2024)
Keypoint Abstraction using Large Models for Object-Relative Imitation Learning
by: Fang, Xiaolin, et al.
Published: (2024)
by: Fang, Xiaolin, et al.
Published: (2024)
Similar Items
-
Moving by Looking: Towards Vision-Driven Avatar Motion Generation
by: Diomataris, Markos, et al.
Published: (2025) -
WANDR: Intention-guided Human Motion Generation
by: Diomataris, Markos, et al.
Published: (2024) -
MotionFix: Text-Driven 3D Human Motion Editing
by: Athanasiou, Nikos, et al.
Published: (2024) -
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
by: Yin, Zhenhan, et al.
Published: (2025) -
VILP: Imitation Learning with Latent Video Planning
by: Xu, Zhengtong, et al.
Published: (2025)