UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Hanjung, Kang, Jaehyun, Kang, Hyolim, Cho, Meedeum, Kim, Seon Joo, Lee, Youngwoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VISAGE: Video Instance Segmentation with Appearance-Guided Enhancement
von: Kim, Hanjung, et al.
Veröffentlicht: (2023)
von: Kim, Hanjung, et al.
Veröffentlicht: (2023)
ActionSwitch: Class-agnostic Detection of Simultaneous Actions in Streaming Videos
von: Kang, Hyolim, et al.
Veröffentlicht: (2024)
von: Kang, Hyolim, et al.
Veröffentlicht: (2024)
Open-ended Hierarchical Streaming Video Understanding with Vision Language Models
von: Kang, Hyolim, et al.
Veröffentlicht: (2025)
von: Kang, Hyolim, et al.
Veröffentlicht: (2025)
Object Aware Egocentric Online Action Detection
von: An, Joungbin, et al.
Veröffentlicht: (2024)
von: An, Joungbin, et al.
Veröffentlicht: (2024)
Query-Calibrated Segmental Admission for Descriptor-Agnostic LiDAR Loop Closure in Repetitive Environments
von: Kim, Jaehyun, et al.
Veröffentlicht: (2025)
von: Kim, Jaehyun, et al.
Veröffentlicht: (2025)
UniPrototype: Humn-Robot Skill Learning with Uniform Prototypes
von: Hu, Xiao, et al.
Veröffentlicht: (2025)
von: Hu, Xiao, et al.
Veröffentlicht: (2025)
CEDex: Cross-Embodiment Dexterous Grasp Generation at Scale from Human-like Contact Representations
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2026)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2026)
Exploring Scalability of Self-Training for Open-Vocabulary Temporal Action Localization
von: Hyun, Jeongseok, et al.
Veröffentlicht: (2024)
von: Hyun, Jeongseok, et al.
Veröffentlicht: (2024)
OKAMI: Teaching Humanoid Robots Manipulation Skills through Single Video Imitation
von: Li, Jinhan, et al.
Veröffentlicht: (2024)
von: Li, Jinhan, et al.
Veröffentlicht: (2024)
FreeAction: Training-Free Techniques for Enhanced Fidelity of Trajectory-to-Video Generation
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning
von: Kim, Jinwoo, et al.
Veröffentlicht: (2023)
von: Kim, Jinwoo, et al.
Veröffentlicht: (2023)
Slot-Level Robotic Placement via Visual Imitation from Single Human Video
von: Shan, Dandan, et al.
Veröffentlicht: (2025)
von: Shan, Dandan, et al.
Veröffentlicht: (2025)
LOTUS: Continual Imitation Learning for Robot Manipulation Through Unsupervised Skill Discovery
von: Wan, Weikang, et al.
Veröffentlicht: (2023)
von: Wan, Weikang, et al.
Veröffentlicht: (2023)
Hierarchical Latent Action Model
von: Kim, Hanjung, et al.
Veröffentlicht: (2026)
von: Kim, Hanjung, et al.
Veröffentlicht: (2026)
X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations
von: Pace, Maximus A., et al.
Veröffentlicht: (2025)
von: Pace, Maximus A., et al.
Veröffentlicht: (2025)
OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction
von: Lee, Junyoung, et al.
Veröffentlicht: (2026)
von: Lee, Junyoung, et al.
Veröffentlicht: (2026)
AnySkill: Learning Open-Vocabulary Physical Skill for Interactive Agents
von: Cui, Jieming, et al.
Veröffentlicht: (2024)
von: Cui, Jieming, et al.
Veröffentlicht: (2024)
ModSkill: Physical Character Skill Modularization
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
TraceGen: World Modeling in 3D Trace Space Enables Learning from Cross-Embodiment Videos
von: Lee, Seungjae, et al.
Veröffentlicht: (2025)
von: Lee, Seungjae, et al.
Veröffentlicht: (2025)
EgoMimic: Scaling Imitation Learning via Egocentric Video
von: Kareer, Simar, et al.
Veröffentlicht: (2024)
von: Kareer, Simar, et al.
Veröffentlicht: (2024)
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
von: Li, Puhao, et al.
Veröffentlicht: (2024)
von: Li, Puhao, et al.
Veröffentlicht: (2024)
BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos
von: Jeon, Subin, et al.
Veröffentlicht: (2024)
von: Jeon, Subin, et al.
Veröffentlicht: (2024)
AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
One-shot Video Imitation via Parameterized Symbolic Abstraction Graphs
von: Wang, Jianren, et al.
Veröffentlicht: (2024)
von: Wang, Jianren, et al.
Veröffentlicht: (2024)
Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation
von: Zhang, Xitie, et al.
Veröffentlicht: (2026)
von: Zhang, Xitie, et al.
Veröffentlicht: (2026)
HRDexDB: A Large-Scale Dataset of Dexterous Human and Robotic Hand Grasps
von: Lim, Jongbin, et al.
Veröffentlicht: (2026)
von: Lim, Jongbin, et al.
Veröffentlicht: (2026)
SkillDiffuser: Interpretable Hierarchical Planning via Skill Abstractions in Diffusion-Based Task Execution
von: Liang, Zhixuan, et al.
Veröffentlicht: (2023)
von: Liang, Zhixuan, et al.
Veröffentlicht: (2023)
X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
von: Zheng, Jinliang, et al.
Veröffentlicht: (2025)
von: Zheng, Jinliang, et al.
Veröffentlicht: (2025)
Thermal Chameleon: Task-Adaptive Tone-mapping for Radiometric Thermal-Infrared images
von: Lee, Dong-Guw, et al.
Veröffentlicht: (2024)
von: Lee, Dong-Guw, et al.
Veröffentlicht: (2024)
VILP: Imitation Learning with Latent Video Planning
von: Xu, Zhengtong, et al.
Veröffentlicht: (2025)
von: Xu, Zhengtong, et al.
Veröffentlicht: (2025)
UniDWM: Towards a Unified Driving World Model via Multifaceted Representation Learning
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
Enhancing Reusability of Learned Skills for Robot Manipulation via Gaze Information and Motion Bottlenecks
von: Takizawa, Ryo, et al.
Veröffentlicht: (2025)
von: Takizawa, Ryo, et al.
Veröffentlicht: (2025)
ZeroMimic: Distilling Robotic Manipulation Skills from Web Videos
von: Shi, Junyao, et al.
Veröffentlicht: (2025)
von: Shi, Junyao, et al.
Veröffentlicht: (2025)
Learning to Enhance Aperture Phasor Field for Non-Line-of-Sight Imaging
von: Cho, In, et al.
Veröffentlicht: (2024)
von: Cho, In, et al.
Veröffentlicht: (2024)
3DFlowAction: Learning Cross-Embodiment Manipulation from 3D Flow World Model
von: Zhi, Hongyan, et al.
Veröffentlicht: (2025)
von: Zhi, Hongyan, et al.
Veröffentlicht: (2025)
Affordance Agent Harness: Verification-Gated Skill Orchestration
von: Huang, Haojian, et al.
Veröffentlicht: (2026)
von: Huang, Haojian, et al.
Veröffentlicht: (2026)
HEAT: Heterogeneous End-to-End Autonomous Driving via Trajectory-Guided World Models
von: Cho, Hoonhee, et al.
Veröffentlicht: (2026)
von: Cho, Hoonhee, et al.
Veröffentlicht: (2026)
Towards Fusing Point Cloud and Visual Representations for Imitation Learning
von: Donat, Atalay, et al.
Veröffentlicht: (2025)
von: Donat, Atalay, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VISAGE: Video Instance Segmentation with Appearance-Guided Enhancement
von: Kim, Hanjung, et al.
Veröffentlicht: (2023) -
ActionSwitch: Class-agnostic Detection of Simultaneous Actions in Streaming Videos
von: Kang, Hyolim, et al.
Veröffentlicht: (2024) -
Open-ended Hierarchical Streaming Video Understanding with Vision Language Models
von: Kang, Hyolim, et al.
Veröffentlicht: (2025) -
Object Aware Egocentric Online Action Detection
von: An, Joungbin, et al.
Veröffentlicht: (2024) -
Query-Calibrated Segmental Admission for Descriptor-Agnostic LiDAR Loop Closure in Repetitive Environments
von: Kim, Jaehyun, et al.
Veröffentlicht: (2025)