FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhaxizhuoma, Liu, Kehui, Guan, Chuyue, Jia, Zhongjie, Wu, Ziniu, Liu, Xin, Wang, Tianyu, Liang, Shuai, Chen, Pengan, Zhang, Pingrui, Song, Haoming, Qu, Delin, Wang, Dong, Wang, Zhigang, Cao, Nieqing, Ding, Yan, Zhao, Bin, Li, Xuelong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset
por: Liu, Kehui, et al.
Publicado: (2025)
por: Liu, Kehui, et al.
Publicado: (2025)
AlignBot: Aligning VLM-powered Customized Task Planning with User Reminders Through Fine-Tuning for Household Robots
por: Zhaxizhuoma, Zhaxizhuoma, et al.
Publicado: (2024)
por: Zhaxizhuoma, Zhaxizhuoma, et al.
Publicado: (2024)
MLM: Learning Multi-task Loco-Manipulation Whole-Body Control for Quadruped Robot with Arm
por: Liu, Xin, et al.
Publicado: (2025)
por: Liu, Xin, et al.
Publicado: (2025)
MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
por: Zhang, Pingrui, et al.
Publicado: (2025)
por: Zhang, Pingrui, et al.
Publicado: (2025)
Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
por: Yao, Yuanqi, et al.
Publicado: (2025)
por: Yao, Yuanqi, et al.
Publicado: (2025)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
por: Gao, Xianqiang, et al.
Publicado: (2024)
por: Gao, Xianqiang, et al.
Publicado: (2024)
Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning
por: Gao, Xianqiang, et al.
Publicado: (2026)
por: Gao, Xianqiang, et al.
Publicado: (2026)
Closed-Loop Action Chunks with Dynamic Corrections for Training-Free Diffusion Policy
por: Wu, Pengyuan, et al.
Publicado: (2026)
por: Wu, Pengyuan, et al.
Publicado: (2026)
LiveScene: Language Embedding Interactive Radiance Fields for Physical Scene Rendering and Control
por: Qu, Delin, et al.
Publicado: (2024)
por: Qu, Delin, et al.
Publicado: (2024)
COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
por: Liu, Kehui, et al.
Publicado: (2024)
por: Liu, Kehui, et al.
Publicado: (2024)
BestMan: A Modular Mobile Manipulator Platform for Embodied AI with Unified Simulation-Hardware APIs
por: Yang, Kui, et al.
Publicado: (2024)
por: Yang, Kui, et al.
Publicado: (2024)
UMI-3D: Extending Universal Manipulation Interface from Vision-Limited to 3D Spatial Perception
por: Wang, Ziming
Publicado: (2026)
por: Wang, Ziming
Publicado: (2026)
DexUMI: Using Human Hand as the Universal Manipulation Interface for Dexterous Manipulation
por: Xu, Mengda, et al.
Publicado: (2025)
por: Xu, Mengda, et al.
Publicado: (2025)
GS-SLAM: Dense Visual SLAM with 3D Gaussian Splatting
por: Yan, Chi, et al.
Publicado: (2023)
por: Yan, Chi, et al.
Publicado: (2023)
MV-UMI: A Scalable Multi-View Interface for Cross-Embodiment Learning
por: Rayyan, Omar, et al.
Publicado: (2025)
por: Rayyan, Omar, et al.
Publicado: (2025)
In-the-Wild Compliant Manipulation with UMI-FT
por: Choi, Hojung, et al.
Publicado: (2026)
por: Choi, Hojung, et al.
Publicado: (2026)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
por: Qu, Delin, et al.
Publicado: (2025)
por: Qu, Delin, et al.
Publicado: (2025)
TacUMI: A Multi-Modal Universal Manipulation Interface for Contact-Rich Tasks
por: Cheng, Tailai, et al.
Publicado: (2026)
por: Cheng, Tailai, et al.
Publicado: (2026)
BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation
por: Yu, Chenhao, et al.
Publicado: (2026)
por: Yu, Chenhao, et al.
Publicado: (2026)
Exploring the Potential of Encoder-free Architectures in 3D LMMs
por: Tang, Yiwen, et al.
Publicado: (2025)
por: Tang, Yiwen, et al.
Publicado: (2025)
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation
por: Zhang, Pingrui, et al.
Publicado: (2025)
por: Zhang, Pingrui, et al.
Publicado: (2025)
UMI-Underwater: Learning Underwater Manipulation without Underwater Teleoperation
por: Li, Hao, et al.
Publicado: (2026)
por: Li, Hao, et al.
Publicado: (2026)
MoMa-Pos: An Efficient Object-Kinematic-Aware Base Placement Optimization Framework for Mobile Manipulation
por: Shao, Beichen, et al.
Publicado: (2024)
por: Shao, Beichen, et al.
Publicado: (2024)
UMI on Legs: Making Manipulation Policies Mobile with Manipulation-Centric Whole-body Controllers
por: Ha, Huy, et al.
Publicado: (2024)
por: Ha, Huy, et al.
Publicado: (2024)
ORLA*: Mobile Manipulator-Based Object Rearrangement with Lazy A Star
por: Gao, Kai, et al.
Publicado: (2023)
por: Gao, Kai, et al.
Publicado: (2023)
GaN nanowires and nanotubes growth by chemical vapor deposition method at different NH3 flow rate
por: Pengan Li
Publicado: (2016)
por: Pengan Li
Publicado: (2016)
Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs
por: Xia, Wenke, et al.
Publicado: (2023)
por: Xia, Wenke, et al.
Publicado: (2023)
FreeGaussian: Annotation-free Control of Articulated Objects via 3D Gaussian Splats with Flow Derivatives
por: Chen, Qizhi, et al.
Publicado: (2024)
por: Chen, Qizhi, et al.
Publicado: (2024)
UMI--History in the Making.
por: Pack, Thomas, et al.
Publicado: (1994)
por: Pack, Thomas, et al.
Publicado: (1994)
Mobile UMI: Cross-View Diffusion Policy with Decoupled Kinematics for Mobile Manipulation
por: Huang, Haoran, et al.
Publicado: (2026)
por: Huang, Haoran, et al.
Publicado: (2026)
Implicit Event-RGBD Neural SLAM
por: Qu, Delin, et al.
Publicado: (2023)
por: Qu, Delin, et al.
Publicado: (2023)
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
por: Zeng, Qiyuan, et al.
Publicado: (2025)
por: Zeng, Qiyuan, et al.
Publicado: (2025)
SuperSuit: An Isomorphic Bimodal Interface for Scalable Mobile Manipulation
por: Chen, Tongqing, et al.
Publicado: (2026)
por: Chen, Tongqing, et al.
Publicado: (2026)
FoLDTree: A ULDA-Based Decision Tree Framework for Efficient Oblique Splits and Feature Selection
por: Wang, Siyu, et al.
Publicado: (2024)
por: Wang, Siyu, et al.
Publicado: (2024)
A New Forward Discriminant Analysis Framework Based On Pillai's Trace and ULDA
por: Wang, Siyu, et al.
Publicado: (2024)
por: Wang, Siyu, et al.
Publicado: (2024)
Comment on: “Long‐Term Risk of Inflammatory Bowel Disease With Sarcopenia Status: A Large‐Scale Prospective Cohort Study”
por: Ao Wang, et al.
Publicado: (2026)
por: Ao Wang, et al.
Publicado: (2026)
Hume: Introducing System-2 Thinking in Visual-Language-Action Model
por: Song, Haoming, et al.
Publicado: (2025)
por: Song, Haoming, et al.
Publicado: (2025)
SurfPatch: Enabling Patch Matching for Exploratory Stream Surface Visualization
por: An, Delin, et al.
Publicado: (2025)
por: An, Delin, et al.
Publicado: (2025)
Sketch2CT: Multimodal Diffusion for Structure-Aware 3D Medical Volume Generation
por: An, Delin, et al.
Publicado: (2026)
por: An, Delin, et al.
Publicado: (2026)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
por: Zhang, Junjie, et al.
Publicado: (2024)
por: Zhang, Junjie, et al.
Publicado: (2024)
Ejemplares similares
-
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset
por: Liu, Kehui, et al.
Publicado: (2025) -
AlignBot: Aligning VLM-powered Customized Task Planning with User Reminders Through Fine-Tuning for Household Robots
por: Zhaxizhuoma, Zhaxizhuoma, et al.
Publicado: (2024) -
MLM: Learning Multi-task Loco-Manipulation Whole-Body Control for Quadruped Robot with Arm
por: Liu, Xin, et al.
Publicado: (2025) -
MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
por: Zhang, Pingrui, et al.
Publicado: (2025) -
Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
por: Yao, Yuanqi, et al.
Publicado: (2025)