H2R: A Human-to-Robot Data Augmentation for Robot Pre-training from Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Guangrun, Lyu, Yaoxu, Liu, Zhuoyang, Hou, Chengkai, Zhang, Jieyu, Zhang, Shanghang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
URDF-Anything+: End-to-End Generation for Simulation-Ready Articulated Assets
by: Wu, Zhuangzhe, et al.
Published: (2026)
by: Wu, Zhuangzhe, et al.
Published: (2026)
Demo-JEPA: Joint-Embedding Predictive Architecture for One-shot Cross-Embodiment Imitation
by: He, Jingyang, et al.
Published: (2026)
by: He, Jingyang, et al.
Published: (2026)
Mask World Model: Predicting What Matters for Robust Robot Policy Learning
by: Lou, Yunfan, et al.
Published: (2026)
by: Lou, Yunfan, et al.
Published: (2026)
LaST$_{0}$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model
by: Liu, Zhuoyang, et al.
Published: (2026)
by: Liu, Zhuoyang, et al.
Published: (2026)
RwoR: Generating Robot Demonstrations from Human Hand Collection for Policy Learning without Robot
by: Heng, Liang, et al.
Published: (2025)
by: Heng, Liang, et al.
Published: (2025)
URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language Model
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
SpikePingpong: Spike Vision-based Fast-Slow Pingpong Robot System
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
ManualVLA: A Unified VLA Model for Chain-of-Thought Manual Generation and Robotic Manipulation
by: Gu, Chenyang, et al.
Published: (2025)
by: Gu, Chenyang, et al.
Published: (2025)
DexH2R: Task-oriented Dexterous Manipulation from Human to Robots
by: Zhao, Shuqi, et al.
Published: (2024)
by: Zhao, Shuqi, et al.
Published: (2024)
MotionTrans: Human VR Data Enable Motion-Level Learning for Robotic Manipulation Policies
by: Yuan, Chengbo, et al.
Published: (2025)
by: Yuan, Chengbo, et al.
Published: (2025)
Human2Robot: Learning Robot Actions from Paired Human-Robot Videos
by: Xie, Sicheng, et al.
Published: (2025)
by: Xie, Sicheng, et al.
Published: (2025)
Robo-Dopamine: General Process Reward Modeling for High-Precision Robotic Manipulation
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
Mitigating the Human-Robot Domain Discrepancy in Visual Pre-training for Robotic Manipulation
by: Zhou, Jiaming, et al.
Published: (2024)
by: Zhou, Jiaming, et al.
Published: (2024)
RoboAct-CLIP: Video-Driven Pre-training of Atomic Action Understanding for Robotics
by: Zhang, Zhiyuan, et al.
Published: (2025)
by: Zhang, Zhiyuan, et al.
Published: (2025)
Learning Physics from Pretrained Video Models: A Multimodal Continuous and Sequential World Interaction Models for Robotic Manipulation
by: Song, Zijian, et al.
Published: (2026)
by: Song, Zijian, et al.
Published: (2026)
MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic Manipulation
by: Liu, Zhuoyang, et al.
Published: (2025)
by: Liu, Zhuoyang, et al.
Published: (2025)
Spatiotemporal Predictive Pre-training for Robotic Motor Control
by: Yang, Jiange, et al.
Published: (2024)
by: Yang, Jiange, et al.
Published: (2024)
MobileH2R: Learning Generalizable Human to Mobile Robot Handover Exclusively from Scalable and Diverse Synthetic Data
by: Wang, Zifan, et al.
Published: (2025)
by: Wang, Zifan, et al.
Published: (2025)
RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Human-to-Robot Interaction: Learning from Video Demonstration for Robot Imitation
by: Canh, Thanh Nguyen, et al.
Published: (2026)
by: Canh, Thanh Nguyen, et al.
Published: (2026)
VIP: Vision Instructed Pre-training for Robotic Manipulation
by: Li, Zhuoling, et al.
Published: (2024)
by: Li, Zhuoling, et al.
Published: (2024)
Sim-and-Human Co-training for Data-Efficient and Generalizable Robotic Manipulation
by: Fang, Kaipeng, et al.
Published: (2026)
by: Fang, Kaipeng, et al.
Published: (2026)
4D Visual Pre-training for Robot Learning
by: Hou, Chengkai, et al.
Published: (2025)
by: Hou, Chengkai, et al.
Published: (2025)
AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation
by: Chen, Sixiang, et al.
Published: (2025)
by: Chen, Sixiang, et al.
Published: (2025)
ARMADA: Augmented Reality for Robot Manipulation and Robot-Free Data Acquisition
by: Nechyporenko, Nataliya, et al.
Published: (2024)
by: Nechyporenko, Nataliya, et al.
Published: (2024)
Video2Act: A Dual-System Video Diffusion Policy with Robotic Spatio-Motional Modeling
by: Jia, Yueru, et al.
Published: (2025)
by: Jia, Yueru, et al.
Published: (2025)
H2R-Grounder: A Paired-Data-Free Paradigm for Translating Human Interaction Videos into Physically Grounded Robot Videos
by: Ci, Hai, et al.
Published: (2025)
by: Ci, Hai, et al.
Published: (2025)
PoseDiff: A Unified Diffusion Model Bridging Robot Pose Estimation and Video-to-Action Control
by: Zhang, Haozhuo, et al.
Published: (2025)
by: Zhang, Haozhuo, et al.
Published: (2025)
GUIDES: Guidance Using Instructor-Distilled Embeddings for Pre-trained Robot Policy Enhancement
by: Gao, Minquan, et al.
Published: (2025)
by: Gao, Minquan, et al.
Published: (2025)
RoboMIND 2.0: A Multimodal, Bimanual Mobile Manipulation Dataset for Generalizable Embodied Intelligence
by: Hou, Chengkai, et al.
Published: (2025)
by: Hou, Chengkai, et al.
Published: (2025)
DexH2R: A Benchmark for Dynamic Dexterous Grasping in Human-to-Robot Handover
by: Wang, Youzhuo, et al.
Published: (2025)
by: Wang, Youzhuo, et al.
Published: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
by: Jiang, Guangqi, et al.
Published: (2024)
by: Jiang, Guangqi, et al.
Published: (2024)
Robotic Manipulation is Vision-to-Geometry Mapping ($f(v) \rightarrow G$): Vision-Geometry Backbones over Language and Video Models
by: Song, Zijian, et al.
Published: (2026)
by: Song, Zijian, et al.
Published: (2026)
Lifelike Agility and Play in Quadrupedal Robots using Reinforcement Learning and Generative Pre-trained Models
by: Han, Lei, et al.
Published: (2023)
by: Han, Lei, et al.
Published: (2023)
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation
by: Xu, Qinwen, et al.
Published: (2026)
by: Xu, Qinwen, et al.
Published: (2026)
Physical Human-Robot Interaction for Grasping in Augmented Reality via Rigid-Soft Robot Synergy
by: Huang, Huishi, et al.
Published: (2026)
by: Huang, Huishi, et al.
Published: (2026)
A Central Motor System Inspired Pre-training Reinforcement Learning for Robotic Control
by: Zhang, Pei, et al.
Published: (2023)
by: Zhang, Pei, et al.
Published: (2023)
Robot Learning from Human Videos: A Survey
by: Ma, Junyi, et al.
Published: (2026)
by: Ma, Junyi, et al.
Published: (2026)
RoboOS-NeXT: A Unified Memory-based Framework for Lifelong, Scalable, and Robust Multi-Robot Collaboration
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
Similar Items
-
URDF-Anything+: End-to-End Generation for Simulation-Ready Articulated Assets
by: Wu, Zhuangzhe, et al.
Published: (2026) -
Demo-JEPA: Joint-Embedding Predictive Architecture for One-shot Cross-Embodiment Imitation
by: He, Jingyang, et al.
Published: (2026) -
Mask World Model: Predicting What Matters for Robust Robot Policy Learning
by: Lou, Yunfan, et al.
Published: (2026) -
LaST$_{0}$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model
by: Liu, Zhuoyang, et al.
Published: (2026) -
RwoR: Generating Robot Demonstrations from Human Hand Collection for Policy Learning without Robot
by: Heng, Liang, et al.
Published: (2025)