OmniUMI: Towards Physically Grounded Robot Learning via Human-Aligned Multimodal Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Shaqi, Li, Yuanyuan, Hu, Youhao, Yu, Chenhao, Xu, Chaoran, Zhang, Jiachen, Yao, Guocai, Huang, Tiejun, He, Ran, Wang, Zhongyuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation
by: Yu, Chenhao, et al.
Published: (2026)
by: Yu, Chenhao, et al.
Published: (2026)
Thor: Towards Human-Level Whole-Body Reactions for Intense Contact-Rich Environments
by: Li, Gangyang, et al.
Published: (2025)
by: Li, Gangyang, et al.
Published: (2025)
A Unified Interaction Control Framework for Safe Robotic Ultrasound Scanning with Human-Intention-Aware Compliance
by: Yan, Xiangjie, et al.
Published: (2024)
by: Yan, Xiangjie, et al.
Published: (2024)
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
by: Zeng, Qiyuan, et al.
Published: (2025)
by: Zeng, Qiyuan, et al.
Published: (2025)
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset
by: Liu, Kehui, et al.
Published: (2025)
by: Liu, Kehui, et al.
Published: (2025)
In-the-Wild Compliant Manipulation with UMI-FT
by: Choi, Hojung, et al.
Published: (2026)
by: Choi, Hojung, et al.
Published: (2026)
DexUMI: Using Human Hand as the Universal Manipulation Interface for Dexterous Manipulation
by: Xu, Mengda, et al.
Published: (2025)
by: Xu, Mengda, et al.
Published: (2025)
When Robots Get Chatty: Grounding Multimodal Human-Robot Conversation and Collaboration
by: Allgeuer, Philipp, et al.
Published: (2024)
by: Allgeuer, Philipp, et al.
Published: (2024)
Multimodal Safe Control for Human-Robot Interaction
by: Pandya, Ravi, et al.
Published: (2023)
by: Pandya, Ravi, et al.
Published: (2023)
exUMI: Extensible Robot Teaching System with Action-aware Task-agnostic Tactile Representation
by: Xu, Yue, et al.
Published: (2025)
by: Xu, Yue, et al.
Published: (2025)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
by: Guo, Heyu, et al.
Published: (2025)
by: Guo, Heyu, et al.
Published: (2025)
OmniManip: Towards General Robotic Manipulation via Object-Centric Interaction Primitives as Spatial Constraints
by: Pan, Mingjie, et al.
Published: (2025)
by: Pan, Mingjie, et al.
Published: (2025)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025)
by: Jiang, Shuo, et al.
Published: (2025)
FG-CLTP: Fine-Grained Contrastive Language Tactile Pretraining for Robotic Manipulation
by: Ma, Wenxuan, et al.
Published: (2026)
by: Ma, Wenxuan, et al.
Published: (2026)
Bridging Probabilistic Inference and Behavior Trees: An Interactive Framework for Adaptive Multi-Robot Cooperation
by: Wang, Chaoran, et al.
Published: (2025)
by: Wang, Chaoran, et al.
Published: (2025)
OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction
by: Lee, Junyoung, et al.
Published: (2026)
by: Lee, Junyoung, et al.
Published: (2026)
Ablation Study of Multimodal Perception, Language Grounding, and Control for Human-Robot Interaction in an Object Detection and Grasping Task
by: Tian, Zi, et al.
Published: (2026)
by: Tian, Zi, et al.
Published: (2026)
TacUMI: A Multi-Modal Universal Manipulation Interface for Contact-Rich Tasks
by: Cheng, Tailai, et al.
Published: (2026)
by: Cheng, Tailai, et al.
Published: (2026)
Towards an Omni-Domain Robotics: A Framework for the Standardization of Multi-Environment Robotic Platforms (MERP)
by: Independent Researcher, Cristhian Mauricio
Published: (2025)
by: Independent Researcher, Cristhian Mauricio
Published: (2025)
UMI-on-Air: Embodiment-Aware Guidance for Embodiment-Agnostic Visuomotor Policies
by: Gupta, Harsh, et al.
Published: (2025)
by: Gupta, Harsh, et al.
Published: (2025)
Multimodal Anomaly Detection for Human-Robot Interaction
by: Ribeiro, Guilherme, et al.
Published: (2026)
by: Ribeiro, Guilherme, et al.
Published: (2026)
Towards Affect-Adaptive Human-Robot Interaction: A Protocol for Multimodal Dataset Collection on Social Anxiety
by: Poprcova, Vesna, et al.
Published: (2025)
by: Poprcova, Vesna, et al.
Published: (2025)
Mobile UMI: Cross-View Diffusion Policy with Decoupled Kinematics for Mobile Manipulation
by: Huang, Haoran, et al.
Published: (2026)
by: Huang, Haoran, et al.
Published: (2026)
FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset
by: Zhaxizhuoma, et al.
Published: (2024)
by: Zhaxizhuoma, et al.
Published: (2024)
UMI-Underwater: Learning Underwater Manipulation without Underwater Teleoperation
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Robo-Dopamine: General Process Reward Modeling for High-Precision Robotic Manipulation
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
Aligning Robot and Human Representations
by: Bobu, Andreea, et al.
Published: (2023)
by: Bobu, Andreea, et al.
Published: (2023)
Human Emotion-Mediated Soft Robotic Arts: Exploring the Intersection of Human Emotions, Soft Robotics and Arts
by: Nadipineni, Saitarun, et al.
Published: (2026)
by: Nadipineni, Saitarun, et al.
Published: (2026)
HANDO: Hierarchical Autonomous Navigation and Dexterous Omni-loco-manipulation
by: Sun, Jingyuan, et al.
Published: (2025)
by: Sun, Jingyuan, et al.
Published: (2025)
Bidirectional Human-Robot Communication for Physical Human-Robot Interaction
by: Wang, Junxiang, et al.
Published: (2026)
by: Wang, Junxiang, et al.
Published: (2026)
ERR@HRI 2024 Challenge: Multimodal Detection of Errors and Failures in Human-Robot Interactions
by: Spitale, Micol, et al.
Published: (2024)
by: Spitale, Micol, et al.
Published: (2024)
Development of a Human-Robot Interaction Platform for Dual-Arm Robots Based on ROS and Multimodal Artificial Intelligence
by: Canh, Thanh Nguyen, et al.
Published: (2024)
by: Canh, Thanh Nguyen, et al.
Published: (2024)
STARC: See-Through-Wall Augmented Reality Framework for Human-Robot Collaboration in Emergency Response
by: Yuan, Shenghai, et al.
Published: (2025)
by: Yuan, Shenghai, et al.
Published: (2025)
MERGE: Guided Vision-Language Models for Multi-Actor Event Reasoning and Grounding in Human-Robot Interaction
by: Deigmoeller, Joerg, et al.
Published: (2026)
by: Deigmoeller, Joerg, et al.
Published: (2026)
METIS: Multi-Source Egocentric Training for Integrated Dexterous Vision-Language-Action Model
by: Fu, Yankai, et al.
Published: (2025)
by: Fu, Yankai, et al.
Published: (2025)
A Graph-to-Text Approach to Knowledge-Grounded Response Generation in Human-Robot Interaction
by: Walker, Nicholas Thomas, et al.
Published: (2023)
by: Walker, Nicholas Thomas, et al.
Published: (2023)
ShanghaiTech Mapping Robot is All You Need: Robot System for Collecting Universal Ground Vehicle Datasets
by: Xu, Bowen, et al.
Published: (2024)
by: Xu, Bowen, et al.
Published: (2024)
Learning Multimodal Latent Dynamics for Human-Robot Interaction
by: Prasad, Vignesh, et al.
Published: (2023)
by: Prasad, Vignesh, et al.
Published: (2023)
Towards Unified Interactive Visual Grounding in The Wild
by: Xu, Jie, et al.
Published: (2024)
by: Xu, Jie, et al.
Published: (2024)
Are Large Language Models Aligned with People's Social Intuitions for Human-Robot Interactions?
by: Wachowiak, Lennart, et al.
Published: (2024)
by: Wachowiak, Lennart, et al.
Published: (2024)
Similar Items
-
BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation
by: Yu, Chenhao, et al.
Published: (2026) -
Thor: Towards Human-Level Whole-Body Reactions for Intense Contact-Rich Environments
by: Li, Gangyang, et al.
Published: (2025) -
A Unified Interaction Control Framework for Safe Robotic Ultrasound Scanning with Human-Intention-Aware Compliance
by: Yan, Xiangjie, et al.
Published: (2024) -
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
by: Zeng, Qiyuan, et al.
Published: (2025) -
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset
by: Liu, Kehui, et al.
Published: (2025)