Saved in:
| Main Authors: | Zhu, Yichen, Feng, Feifei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.05199 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Visual Robotic Manipulation with Depth-Aware Pretraining
by: Wang, Wanying, et al.
Published: (2024)
by: Wang, Wanying, et al.
Published: (2024)
Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation
by: Zhu, Minjie, et al.
Published: (2024)
by: Zhu, Minjie, et al.
Published: (2024)
Object-Centric Instruction Augmentation for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos
by: Gavryushin, Alexey, et al.
Published: (2025)
by: Gavryushin, Alexey, et al.
Published: (2025)
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
by: Zhu, Minjie, et al.
Published: (2024)
by: Zhu, Minjie, et al.
Published: (2024)
Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt
by: Zhu, Xiang, et al.
Published: (2025)
by: Zhu, Xiang, et al.
Published: (2025)
EgoDemoGen: Egocentric Demonstration Generation for Viewpoint Generalization in Robotic Manipulation
by: Xu, Yuan, et al.
Published: (2025)
by: Xu, Yuan, et al.
Published: (2025)
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
by: Wen, Junjie, et al.
Published: (2025)
by: Wen, Junjie, et al.
Published: (2025)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration
by: Shi, Modi, et al.
Published: (2026)
by: Shi, Modi, et al.
Published: (2026)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
by: Zhu, Minjie, et al.
Published: (2025)
by: Zhu, Minjie, et al.
Published: (2025)
EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
by: Hoque, Ryan, et al.
Published: (2025)
by: Hoque, Ryan, et al.
Published: (2025)
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
by: Zeng, Qiyuan, et al.
Published: (2025)
by: Zeng, Qiyuan, et al.
Published: (2025)
You Only Teach Once: Learn One-Shot Bimanual Robotic Manipulation from Video Demonstrations
by: Zhou, Huayi, et al.
Published: (2025)
by: Zhou, Huayi, et al.
Published: (2025)
EMMA: Scaling Mobile Manipulation via Egocentric Human Data
by: Zhu, Lawrence Y., et al.
Published: (2025)
by: Zhu, Lawrence Y., et al.
Published: (2025)
Don't Let Your Robot be Harmful: Responsible Robotic Manipulation via Safety-as-Policy
by: Ni, Minheng, et al.
Published: (2024)
by: Ni, Minheng, et al.
Published: (2024)
MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos
by: Shah, Rutav, et al.
Published: (2025)
by: Shah, Rutav, et al.
Published: (2025)
EgoPush: Learning End-to-End Egocentric Multi-Object Rearrangement for Mobile Robots
by: An, Boyuan, et al.
Published: (2026)
by: An, Boyuan, et al.
Published: (2026)
MMRo: Are Multimodal LLMs Eligible as the Brain for In-Home Robotics?
by: Li, Jinming, et al.
Published: (2024)
by: Li, Jinming, et al.
Published: (2024)
Learning Generalizable Language-Conditioned Cloth Manipulation from Long Demonstrations
by: Zhao, Hanyi, et al.
Published: (2025)
by: Zhao, Hanyi, et al.
Published: (2025)
EgoAVFlow: Robot Policy Learning with Active Vision from Human Egocentric Videos via 3D Flow
by: Cho, Daesol, et al.
Published: (2026)
by: Cho, Daesol, et al.
Published: (2026)
Human Preference Modeling Using Visual Motion Prediction Improves Robot Skill Learning from Egocentric Human Video
by: Verghese, Mrinal, et al.
Published: (2026)
by: Verghese, Mrinal, et al.
Published: (2026)
Learning Granular Media Avalanche Behavior for Indirectly Manipulating Obstacles on a Granular Slope
by: Hu, Haodi, et al.
Published: (2024)
by: Hu, Haodi, et al.
Published: (2024)
Goal State Generation for Robotic Manipulation Based on Linguistically Guided Hybrid Gaussian Diffusion
by: Xu, Yichen, et al.
Published: (2024)
by: Xu, Yichen, et al.
Published: (2024)
Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning
by: Tirumala, Dhruva, et al.
Published: (2024)
by: Tirumala, Dhruva, et al.
Published: (2024)
EgoScale: Scaling Dexterous Manipulation with Diverse Egocentric Human Data
by: Zheng, Ruijie, et al.
Published: (2026)
by: Zheng, Ruijie, et al.
Published: (2026)
EgoMI: Learning Active Vision and Whole-Body Manipulation from Egocentric Human Demonstrations
by: Yu, Justin, et al.
Published: (2025)
by: Yu, Justin, et al.
Published: (2025)
HAND Me the Data: Fast Robot Adaptation via Hand Path Retrieval
by: Hong, Matthew, et al.
Published: (2025)
by: Hong, Matthew, et al.
Published: (2025)
Retrieval-Augmented Embodied Agents
by: Zhu, Yichen, et al.
Published: (2024)
by: Zhu, Yichen, et al.
Published: (2024)
Learning Efficient Robotic Garment Manipulation with Standardization
by: Zhou, Changshi, et al.
Published: (2025)
by: Zhou, Changshi, et al.
Published: (2025)
HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model
by: Zhu, Xiang, et al.
Published: (2026)
by: Zhu, Xiang, et al.
Published: (2026)
Information Filtering via Variational Regularization for Robot Manipulation
by: Zhang, Jinhao, et al.
Published: (2026)
by: Zhang, Jinhao, et al.
Published: (2026)
Efficient Sensorimotor Learning for Open-world Robot Manipulation
by: Zhu, Yifeng
Published: (2025)
by: Zhu, Yifeng
Published: (2025)
Is Diversity All You Need for Scalable Robotic Manipulation?
by: Shi, Modi, et al.
Published: (2025)
by: Shi, Modi, et al.
Published: (2025)
ConLA: Contrastive Latent Action Learning from Human Videos for Robotic Manipulation
by: Dai, Weisheng, et al.
Published: (2026)
by: Dai, Weisheng, et al.
Published: (2026)
WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations
by: Freeman, Harry, et al.
Published: (2026)
by: Freeman, Harry, et al.
Published: (2026)
UniDex: A Robot Foundation Suite for Universal Dexterous Hand Control from Egocentric Human Videos
by: Zhang, Gu, et al.
Published: (2026)
by: Zhang, Gu, et al.
Published: (2026)
Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
Similar Items
-
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
by: Li, Gen, et al.
Published: (2024) -
Visual Robotic Manipulation with Depth-Aware Pretraining
by: Wang, Wanying, et al.
Published: (2024) -
Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation
by: Zhu, Minjie, et al.
Published: (2024) -
Object-Centric Instruction Augmentation for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024) -
MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos
by: Gavryushin, Alexey, et al.
Published: (2025)