Demo-JEPA: Joint-Embedding Predictive Architecture for One-shot Cross-Embodiment Imitation
Fuente:
arXiv
Saved in:
| Main Authors: | He, Jingyang, Li, Guangrun, Zhang, Jieyu, Hou, Chengkai, Che, Zhengping, Zhang, Shanghang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
H2R: A Human-to-Robot Data Augmentation for Robot Pre-training from Videos
by: Li, Guangrun, et al.
Published: (2025)
by: Li, Guangrun, et al.
Published: (2025)
HEX: Humanoid-Aligned Experts for Cross-Embodiment Whole-Body Manipulation
by: Bai, Shuanghao, et al.
Published: (2026)
by: Bai, Shuanghao, et al.
Published: (2026)
URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language Model
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
URDF-Anything+: End-to-End Generation for Simulation-Ready Articulated Assets
by: Wu, Zhuangzhe, et al.
Published: (2026)
by: Wu, Zhuangzhe, et al.
Published: (2026)
LEGATO: Cross-Embodiment Imitation Using a Grasping Tool
by: Seo, Mingyo, et al.
Published: (2024)
by: Seo, Mingyo, et al.
Published: (2024)
JEPA-VLA: Video Predictive Embedding is Needed for VLA Models
by: Miao, Shangchen, et al.
Published: (2026)
by: Miao, Shangchen, et al.
Published: (2026)
MOTIF: Learning Action Motifs for Few-shot Cross-Embodiment Transfer
by: Zhi, Heng, et al.
Published: (2026)
by: Zhi, Heng, et al.
Published: (2026)
DemoDiffusion: One-Shot Human Imitation using pre-trained Diffusion Policy
by: Park, Sungjae, et al.
Published: (2025)
by: Park, Sungjae, et al.
Published: (2025)
UMIGen: A Unified Framework for Egocentric Point Cloud Generation and Cross-Embodiment Robotic Imitation Learning
by: Huang, Yan, et al.
Published: (2025)
by: Huang, Yan, et al.
Published: (2025)
H-Zero: Cross-Humanoid Locomotion Pretraining Enables Few-shot Novel Embodiment Transfer
by: Lin, Yunfeng, et al.
Published: (2025)
by: Lin, Yunfeng, et al.
Published: (2025)
LAP: Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transfer
by: Zha, Lihan, et al.
Published: (2026)
by: Zha, Lihan, et al.
Published: (2026)
RoboOS: A Hierarchical Embodied Framework for Cross-Embodiment and Multi-Agent Collaboration
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
Label-Efficient Grasp Joint Prediction with Point-JEPA
by: Guzelkabaagac, Jed, et al.
Published: (2025)
by: Guzelkabaagac, Jed, et al.
Published: (2025)
LaST$_{0}$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model
by: Liu, Zhuoyang, et al.
Published: (2026)
by: Liu, Zhuoyang, et al.
Published: (2026)
Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing
by: Li, Zhiyuan, et al.
Published: (2026)
by: Li, Zhiyuan, et al.
Published: (2026)
UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations
by: Kim, Hanjung, et al.
Published: (2025)
by: Kim, Hanjung, et al.
Published: (2025)
GET-Zero: Graph Embodiment Transformer for Zero-shot Embodiment Generalization
by: Patel, Austin, et al.
Published: (2024)
by: Patel, Austin, et al.
Published: (2024)
One-Policy-Fits-All: Geometry-Aware Action Latents for Cross-Embodiment Manipulation
by: Mu, Juncheng, et al.
Published: (2026)
by: Mu, Juncheng, et al.
Published: (2026)
XMoP: Whole-Body Control Policy for Zero-shot Cross-Embodiment Neural Motion Planning
by: Rath, Prabin Kumar, et al.
Published: (2024)
by: Rath, Prabin Kumar, et al.
Published: (2024)
LACE: Latent Visual Representation for Cross-Embodiment Learning
by: Jang, Yoo Sung, et al.
Published: (2026)
by: Jang, Yoo Sung, et al.
Published: (2026)
One Demo Is All It Takes: Planning Domain Derivation with LLMs from A Single Demonstration
by: Huang, Jinbang, et al.
Published: (2025)
by: Huang, Jinbang, et al.
Published: (2025)
RoboMIND 2.0: A Multimodal, Bimanual Mobile Manipulation Dataset for Generalizable Embodied Intelligence
by: Hou, Chengkai, et al.
Published: (2025)
by: Hou, Chengkai, et al.
Published: (2025)
Scaling Cross-Embodiment World Models for Dexterous Manipulation
by: He, Zihao, et al.
Published: (2025)
by: He, Zihao, et al.
Published: (2025)
Learning Adaptive Cross-Embodiment Visuomotor Policy with Contrastive Prompt Orchestration
by: Zhang, Yuhang, et al.
Published: (2026)
by: Zhang, Yuhang, et al.
Published: (2026)
One-shot Video Imitation via Parameterized Symbolic Abstraction Graphs
by: Wang, Jianren, et al.
Published: (2024)
by: Wang, Jianren, et al.
Published: (2024)
Latent Action Diffusion for Cross-Embodiment Manipulation
by: Bauer, Erik, et al.
Published: (2025)
by: Bauer, Erik, et al.
Published: (2025)
DSeq-JEPA: Discriminative Sequential Joint-Embedding Predictive Architecture
by: He, Xiangteng, et al.
Published: (2025)
by: He, Xiangteng, et al.
Published: (2025)
UniBYD: A Unified Framework for Learning Robotic Manipulation Across Embodiments Beyond Imitation of Human Demonstrations
by: Yuan, Tingyu, et al.
Published: (2025)
by: Yuan, Tingyu, et al.
Published: (2025)
Crossing the Human-Robot Embodiment Gap with Sim-to-Real RL using One Human Demonstration
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
CE-Nav: Flow-Guided Reinforcement Refinement for Cross-Embodiment Local Navigation
by: Yang, Kai, et al.
Published: (2025)
by: Yang, Kai, et al.
Published: (2025)
Mirage: Cross-Embodiment Zero-Shot Policy Transfer with Cross-Painting
by: Chen, Lawrence Yunliang, et al.
Published: (2024)
by: Chen, Lawrence Yunliang, et al.
Published: (2024)
How to Utilize Failure Demo Data?: Effective Data Selection for Imitation Learning Using Distribution Differences in Attention Mechanism
by: Miyamoto, Kana, et al.
Published: (2026)
by: Miyamoto, Kana, et al.
Published: (2026)
VoroNav: Voronoi-based Zero-shot Object Navigation with Large Language Model
by: Wu, Pengying, et al.
Published: (2024)
by: Wu, Pengying, et al.
Published: (2024)
AdaMorph: Unified Motion Retargeting via Embodiment-Aware Adaptive Transformers
by: Zhang, Haoyu, et al.
Published: (2026)
by: Zhang, Haoyu, et al.
Published: (2026)
Meta-Controller: Few-Shot Imitation of Unseen Embodiments and Tasks in Continuous Control
by: Cho, Seongwoong, et al.
Published: (2024)
by: Cho, Seongwoong, et al.
Published: (2024)
RoboWheel: A Data Engine from Real-World Human Demonstrations for Cross-Embodiment Robotic Learning
by: Zhang, Yuhong, et al.
Published: (2025)
by: Zhang, Yuhong, et al.
Published: (2025)
Cross-Embodiment Dexterous Grasping with Reinforcement Learning
by: Yuan, Haoqi, et al.
Published: (2024)
by: Yuan, Haoqi, et al.
Published: (2024)
Being-H0.5: Scaling Human-Centric Robot Learning for Cross-Embodiment Generalization
by: Luo, Hao, et al.
Published: (2026)
by: Luo, Hao, et al.
Published: (2026)
Shadow: Leveraging Segmentation Masks for Cross-Embodiment Policy Transfer
by: Lepert, Marion, et al.
Published: (2025)
by: Lepert, Marion, et al.
Published: (2025)
Predictive Reachability for Embodiment Selection in Mobile Manipulation Behaviors
by: Feng, Xiaoxu, et al.
Published: (2024)
by: Feng, Xiaoxu, et al.
Published: (2024)
Similar Items
-
H2R: A Human-to-Robot Data Augmentation for Robot Pre-training from Videos
by: Li, Guangrun, et al.
Published: (2025) -
HEX: Humanoid-Aligned Experts for Cross-Embodiment Whole-Body Manipulation
by: Bai, Shuanghao, et al.
Published: (2026) -
URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language Model
by: Li, Zhe, et al.
Published: (2025) -
URDF-Anything+: End-to-End Generation for Simulation-Ready Articulated Assets
by: Wu, Zhuangzhe, et al.
Published: (2026) -
LEGATO: Cross-Embodiment Imitation Using a Grasping Tool
by: Seo, Mingyo, et al.
Published: (2024)