In-N-On: Scaling Egocentric Manipulation with in-the-wild and on-task Data
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Xiongyi, Qiu, Ri-Zhao, Chen, Geng, Wei, Lai, Liu, Isabella, Huang, Tianshu, Cheng, Xuxin, Wang, Xiaolong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual Whole-Body Control for Legged Loco-Manipulation
by: Liu, Minghuan, et al.
Published: (2024)
by: Liu, Minghuan, et al.
Published: (2024)
GraspSplats: Efficient Manipulation with 3D Feature Splatting
by: Ji, Mazeyu, et al.
Published: (2024)
by: Ji, Mazeyu, et al.
Published: (2024)
HMC: Learning Heterogeneous Meta-Control for Contact-Rich Loco-Manipulation
by: Wei, Lai, et al.
Published: (2025)
by: Wei, Lai, et al.
Published: (2025)
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
by: Yang, Ruihan, et al.
Published: (2025)
by: Yang, Ruihan, et al.
Published: (2025)
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
by: Punamiya, Ryan, et al.
Published: (2026)
by: Punamiya, Ryan, et al.
Published: (2026)
GSWorld: Closed-Loop Photo-Realistic Simulation Suite for Robotic Manipulation
by: Jiang, Guangqi, et al.
Published: (2025)
by: Jiang, Guangqi, et al.
Published: (2025)
EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
by: Hoque, Ryan, et al.
Published: (2025)
by: Hoque, Ryan, et al.
Published: (2025)
Exploring the Limits of Vision-Language-Action Manipulations in Cross-task Generalization
by: Zhou, Jiaming, et al.
Published: (2025)
by: Zhou, Jiaming, et al.
Published: (2025)
M3: 3D-Spatial MultiModal Memory
by: Zou, Xueyan, et al.
Published: (2025)
by: Zou, Xueyan, et al.
Published: (2025)
EgoMimic: Scaling Imitation Learning via Egocentric Video
by: Kareer, Simar, et al.
Published: (2024)
by: Kareer, Simar, et al.
Published: (2024)
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Learning Generalizable Feature Fields for Mobile Manipulation
by: Qiu, Ri-Zhao, et al.
Published: (2024)
by: Qiu, Ri-Zhao, et al.
Published: (2024)
OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation
by: Jawaid, Ahad, et al.
Published: (2025)
by: Jawaid, Ahad, et al.
Published: (2025)
EgoSpot:Egocentric Multimodal Control for Hands-Free Mobile Manipulation
by: Zhang, Ganlin, et al.
Published: (2023)
by: Zhang, Ganlin, et al.
Published: (2023)
WildLMa: Long Horizon Loco-Manipulation in the Wild
by: Qiu, Ri-Zhao, et al.
Published: (2024)
by: Qiu, Ri-Zhao, et al.
Published: (2024)
Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos
by: Yuan, Chengbo, et al.
Published: (2024)
by: Yuan, Chengbo, et al.
Published: (2024)
MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos
by: Gavryushin, Alexey, et al.
Published: (2025)
by: Gavryushin, Alexey, et al.
Published: (2025)
Benchmarking Egocentric Visual-Inertial SLAM at City Scale
by: Krishnan, Anusha, et al.
Published: (2025)
by: Krishnan, Anusha, et al.
Published: (2025)
SEER-VAR: Semantic Egocentric Environment Reasoner for Vehicle Augmented Reality
by: Lai, Yuzhi, et al.
Published: (2025)
by: Lai, Yuzhi, et al.
Published: (2025)
Helpful DoggyBot: Open-World Object Fetching using Legged Robots and Vision-Language Models
by: Wu, Qi, et al.
Published: (2024)
by: Wu, Qi, et al.
Published: (2024)
CheckManual: A New Challenge and Benchmark for Manual-based Appliance Manipulation
by: Long, Yuxing, et al.
Published: (2025)
by: Long, Yuxing, et al.
Published: (2025)
ManiGaussian: Dynamic Gaussian Splatting for Multi-task Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2024)
by: Lu, Guanxing, et al.
Published: (2024)
Humanoid Policy ~ Human Policy
by: Qiu, Ri-Zhao, et al.
Published: (2025)
by: Qiu, Ri-Zhao, et al.
Published: (2025)
Lucid-XR: An Extended-Reality Data Engine for Robotic Manipulation
by: Ravan, Yajvan, et al.
Published: (2026)
by: Ravan, Yajvan, et al.
Published: (2026)
EgoKit: Towards Unified Low-Cost Egocentric Data Collection with Heterogeneous Devices
by: Yu, Liuchuan, et al.
Published: (2026)
by: Yu, Liuchuan, et al.
Published: (2026)
Ensuring Force Safety in Vision-Guided Robotic Manipulation via Implicit Tactile Calibration
by: Wei, Lai, et al.
Published: (2024)
by: Wei, Lai, et al.
Published: (2024)
Scaling Cross-Environment Failure Reasoning Data for Vision-Language Robotic Manipulation
by: Pacaud, Paul, et al.
Published: (2025)
by: Pacaud, Paul, et al.
Published: (2025)
SAGE: Bridging Semantic and Actionable Parts for GEneralizable Manipulation of Articulated Objects
by: Geng, Haoran, et al.
Published: (2023)
by: Geng, Haoran, et al.
Published: (2023)
ACE: A Cross-Platform Visual-Exoskeletons System for Low-Cost Dexterous Teleoperation
by: Yang, Shiqi, et al.
Published: (2024)
by: Yang, Shiqi, et al.
Published: (2024)
CyberDemo: Augmenting Simulated Human Demonstration for Real-World Dexterous Manipulation
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
AoE: Always-on Egocentric Human Video Collection for Embodied AI
by: Yang, Bowen, et al.
Published: (2026)
by: Yang, Bowen, et al.
Published: (2026)
METIS: Multi-Source Egocentric Training for Integrated Dexterous Vision-Language-Action Model
by: Fu, Yankai, et al.
Published: (2025)
by: Fu, Yankai, et al.
Published: (2025)
The Monado SLAM Dataset for Egocentric Visual-Inertial Tracking
by: de Mayo, Mateo, et al.
Published: (2025)
by: de Mayo, Mateo, et al.
Published: (2025)
Robot Synesthesia: In-Hand Manipulation with Visuotactile Sensing
by: Yuan, Ying, et al.
Published: (2023)
by: Yuan, Ying, et al.
Published: (2023)
RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation
by: Han, Songhao, et al.
Published: (2025)
by: Han, Songhao, et al.
Published: (2025)
Dex1B: Learning with 1B Demonstrations for Dexterous Manipulation
by: Ye, Jianglong, et al.
Published: (2025)
by: Ye, Jianglong, et al.
Published: (2025)
Simultaneous Localization and Affordance Prediction of Tasks from Egocentric Video
by: Chavis, Zachary, et al.
Published: (2024)
by: Chavis, Zachary, et al.
Published: (2024)
Egocentric World Model for Photorealistic Hand-Object Interaction Synthesis
by: Li, Dayou, et al.
Published: (2026)
by: Li, Dayou, et al.
Published: (2026)
Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views
by: Ma, Junyi, et al.
Published: (2025)
by: Ma, Junyi, et al.
Published: (2025)
SM$^3$: Self-Supervised Multi-task Modeling with Multi-view 2D Images for Articulated Objects
by: Wang, Haowen, et al.
Published: (2024)
by: Wang, Haowen, et al.
Published: (2024)
Similar Items
-
Visual Whole-Body Control for Legged Loco-Manipulation
by: Liu, Minghuan, et al.
Published: (2024) -
GraspSplats: Efficient Manipulation with 3D Feature Splatting
by: Ji, Mazeyu, et al.
Published: (2024) -
HMC: Learning Heterogeneous Meta-Control for Contact-Rich Loco-Manipulation
by: Wei, Lai, et al.
Published: (2025) -
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
by: Yang, Ruihan, et al.
Published: (2025) -
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
by: Punamiya, Ryan, et al.
Published: (2026)