EgoMimic: Scaling Imitation Learning via Egocentric Video
Fuente:
arXiv
Saved in:
| Main Authors: | Kareer, Simar, Patel, Dhruv, Punamiya, Ryan, Mathur, Pranay, Cheng, Shuo, Wang, Chen, Hoffman, Judy, Xu, Danfei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EgoBridge: Domain Adaptation for Generalizable Imitation from Egocentric Human Data
by: Punamiya, Ryan, et al.
Published: (2025)
by: Punamiya, Ryan, et al.
Published: (2025)
EMMA: Scaling Mobile Manipulation via Egocentric Human Data
by: Zhu, Lawrence Y., et al.
Published: (2025)
by: Zhu, Lawrence Y., et al.
Published: (2025)
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
by: Punamiya, Ryan, et al.
Published: (2026)
by: Punamiya, Ryan, et al.
Published: (2026)
Emergence of Human to Robot Transfer in Vision-Language-Action Models
by: Kareer, Simar, et al.
Published: (2025)
by: Kareer, Simar, et al.
Published: (2025)
EgoAVFlow: Robot Policy Learning with Active Vision from Human Egocentric Videos via 3D Flow
by: Cho, Daesol, et al.
Published: (2026)
by: Cho, Daesol, et al.
Published: (2026)
EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
by: Hoque, Ryan, et al.
Published: (2025)
by: Hoque, Ryan, et al.
Published: (2025)
ImMimic: Cross-Domain Imitation from Human Videos via Mapping and Interpolation
by: Liu, Yangcen, et al.
Published: (2025)
by: Liu, Yangcen, et al.
Published: (2025)
We're Not Using Videos Effectively: An Updated Domain Adaptive Video Segmentation Baseline
by: Kareer, Simar, et al.
Published: (2024)
by: Kareer, Simar, et al.
Published: (2024)
Neural Visibility Field for Uncertainty-Driven Active Mapping
by: Xue, Shangjie, et al.
Published: (2024)
by: Xue, Shangjie, et al.
Published: (2024)
EgoScale: Scaling Dexterous Manipulation with Diverse Egocentric Human Data
by: Zheng, Ruijie, et al.
Published: (2026)
by: Zheng, Ruijie, et al.
Published: (2026)
Ego-Grounding for Personalized Question-Answering in Egocentric Videos
by: Xiao, Junbin, et al.
Published: (2026)
by: Xiao, Junbin, et al.
Published: (2026)
Learning to Discern: Imitating Heterogeneous Human Demonstrations with Preference and Representation Learning
by: Kuhar, Sachit, et al.
Published: (2023)
by: Kuhar, Sachit, et al.
Published: (2023)
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
by: Yang, Ruihan, et al.
Published: (2025)
by: Yang, Ruihan, et al.
Published: (2025)
OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation
by: Jawaid, Ahad, et al.
Published: (2025)
by: Jawaid, Ahad, et al.
Published: (2025)
MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence
by: Tang, Chao, et al.
Published: (2025)
by: Tang, Chao, et al.
Published: (2025)
Uncertainty-driven 3D Gaussian Splatting Active Mapping via Anisotropic Visibility Field
by: Xue, Shangjie, et al.
Published: (2026)
by: Xue, Shangjie, et al.
Published: (2026)
C3DM: Constrained-Context Conditional Diffusion Models for Imitation Learning
by: Saxena, Vaibhav, et al.
Published: (2023)
by: Saxena, Vaibhav, et al.
Published: (2023)
DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning
by: Jiang, Zhenyu, et al.
Published: (2024)
by: Jiang, Zhenyu, et al.
Published: (2024)
EgoSpot:Egocentric Multimodal Control for Hands-Free Mobile Manipulation
by: Zhang, Ganlin, et al.
Published: (2023)
by: Zhang, Ganlin, et al.
Published: (2023)
VILP: Imitation Learning with Latent Video Planning
by: Xu, Zhengtong, et al.
Published: (2025)
by: Xu, Zhengtong, et al.
Published: (2025)
EgoKit: Towards Unified Low-Cost Egocentric Data Collection with Heterogeneous Devices
by: Yu, Liuchuan, et al.
Published: (2026)
by: Yu, Liuchuan, et al.
Published: (2026)
EgoActor: Grounding Task Planning into Spatial-aware Egocentric Actions for Humanoid Robots via Visual-Language Models
by: Bai, Yu, et al.
Published: (2026)
by: Bai, Yu, et al.
Published: (2026)
LatentMimic: Terrain-Adaptive Locomotion via Latent Space Imitation
by: Wang, Zhiquan, et al.
Published: (2026)
by: Wang, Zhiquan, et al.
Published: (2026)
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
EgoEvGesture: Gesture Recognition Based on Egocentric Event Camera
by: Wang, Luming, et al.
Published: (2025)
by: Wang, Luming, et al.
Published: (2025)
TAVIS: A Benchmark for Egocentric Active Vision and Anticipatory Gaze in Imitation Learning
by: Spigler, Giacomo
Published: (2026)
by: Spigler, Giacomo
Published: (2026)
Advancing Egocentric Video Question Answering with Multimodal Large Language Models
by: Patel, Alkesh, et al.
Published: (2025)
by: Patel, Alkesh, et al.
Published: (2025)
EasyMimic: A Low-Cost Framework for Robot Imitation Learning from Human Videos
by: Zhang, Tao, et al.
Published: (2026)
by: Zhang, Tao, et al.
Published: (2026)
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
by: Wang, Zhi, et al.
Published: (2026)
by: Wang, Zhi, et al.
Published: (2026)
Learning Predictive Visuomotor Coordination
by: Jia, Wenqi, et al.
Published: (2025)
by: Jia, Wenqi, et al.
Published: (2025)
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
by: Patel, Shivansh, et al.
Published: (2025)
by: Patel, Shivansh, et al.
Published: (2025)
MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos
by: Gavryushin, Alexey, et al.
Published: (2025)
by: Gavryushin, Alexey, et al.
Published: (2025)
IndEgo: A Dataset of Industrial Scenarios and Collaborative Work for Egocentric Assistants
by: Chavan, Vivek, et al.
Published: (2025)
by: Chavan, Vivek, et al.
Published: (2025)
Ego to World: Collaborative Spatial Reasoning in Embodied Systems via Reinforcement Learning
by: Zhou, Heng, et al.
Published: (2026)
by: Zhou, Heng, et al.
Published: (2026)
One-shot Video Imitation via Parameterized Symbolic Abstraction Graphs
by: Wang, Jianren, et al.
Published: (2024)
by: Wang, Jianren, et al.
Published: (2024)
RoboMirror: Understand Before You Imitate for Video to Humanoid Locomotion
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
In-N-On: Scaling Egocentric Manipulation with in-the-wild and on-task Data
by: Cai, Xiongyi, et al.
Published: (2025)
by: Cai, Xiongyi, et al.
Published: (2025)
Multi-Camera View Scaling for Data-Efficient Robot Imitation Learning
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
Scalable Benchmarking and Robust Learning for Noise-Free Ego-Motion and 3D Reconstruction from Noisy Video
by: Xu, Xiaohao, et al.
Published: (2025)
by: Xu, Xiaohao, et al.
Published: (2025)
NOD-TAMP: Generalizable Long-Horizon Planning with Neural Object Descriptors
by: Cheng, Shuo, et al.
Published: (2023)
by: Cheng, Shuo, et al.
Published: (2023)
Similar Items
-
EgoBridge: Domain Adaptation for Generalizable Imitation from Egocentric Human Data
by: Punamiya, Ryan, et al.
Published: (2025) -
EMMA: Scaling Mobile Manipulation via Egocentric Human Data
by: Zhu, Lawrence Y., et al.
Published: (2025) -
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
by: Punamiya, Ryan, et al.
Published: (2026) -
Emergence of Human to Robot Transfer in Vision-Language-Action Models
by: Kareer, Simar, et al.
Published: (2025) -
EgoAVFlow: Robot Policy Learning with Active Vision from Human Egocentric Videos via 3D Flow
by: Cho, Daesol, et al.
Published: (2026)