VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Shaofeng, Ze, Yanjie, Yu, Hong-Xing, Liu, C. Karen, Wu, Jiajun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation
by: Dong, Runpei, et al.
Published: (2026)
by: Dong, Runpei, et al.
Published: (2026)
ResMimic: From General Motion Tracking to Humanoid Whole-body Loco-Manipulation via Residual Learning
by: Zhao, Siheng, et al.
Published: (2025)
by: Zhao, Siheng, et al.
Published: (2025)
Generalizable Humanoid Manipulation with 3D Diffusion Policies
by: Ze, Yanjie, et al.
Published: (2024)
by: Ze, Yanjie, et al.
Published: (2024)
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation
by: He, Xialin, et al.
Published: (2026)
by: He, Xialin, et al.
Published: (2026)
Visual Whole-Body Control for Legged Loco-Manipulation
by: Liu, Minghuan, et al.
Published: (2024)
by: Liu, Minghuan, et al.
Published: (2024)
SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework
by: Wu, Tianshu, et al.
Published: (2026)
by: Wu, Tianshu, et al.
Published: (2026)
TWIST2: Scalable, Portable, and Holistic Humanoid Data Collection System
by: Ze, Yanjie, et al.
Published: (2025)
by: Ze, Yanjie, et al.
Published: (2025)
Learning Humanoid Navigation from Human Data
by: Wang, Weizhuo, et al.
Published: (2026)
by: Wang, Weizhuo, et al.
Published: (2026)
Visually-grounded Humanoid Agents
by: Ye, Hang, et al.
Published: (2026)
by: Ye, Hang, et al.
Published: (2026)
Retargeting Matters: General Motion Retargeting for Humanoid Motion Tracking
by: Araujo, Joao Pedro, et al.
Published: (2025)
by: Araujo, Joao Pedro, et al.
Published: (2025)
X-Capture: An Open-Source Portable Device for Multi-Sensory Learning
by: Clarke, Samuel, et al.
Published: (2025)
by: Clarke, Samuel, et al.
Published: (2025)
Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration
by: Ding, Pengxiang, et al.
Published: (2025)
by: Ding, Pengxiang, et al.
Published: (2025)
PhyGile: Physics-Prefix Guided Motion Generation for Agile General Humanoid Motion Tracking
by: Bao, Jiacheng, et al.
Published: (2026)
by: Bao, Jiacheng, et al.
Published: (2026)
TrackVLA: Embodied Visual Tracking in the Wild
by: Wang, Shaoan, et al.
Published: (2025)
by: Wang, Shaoan, et al.
Published: (2025)
Visual Imitation Enables Contextual Humanoid Control
by: Allshire, Arthur, et al.
Published: (2025)
by: Allshire, Arthur, et al.
Published: (2025)
AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
by: Yang, Dejie, et al.
Published: (2025)
by: Yang, Dejie, et al.
Published: (2025)
WholeBodyVLA: Towards Unified Latent VLA for Whole-Body Loco-Manipulation Control
by: Jiang, Haoran, et al.
Published: (2025)
by: Jiang, Haoran, et al.
Published: (2025)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
EgoActor: Grounding Task Planning into Spatial-aware Egocentric Actions for Humanoid Robots via Visual-Language Models
by: Bai, Yu, et al.
Published: (2026)
by: Bai, Yu, et al.
Published: (2026)
Image Generation as a Visual Planner for Robotic Manipulation
by: Pang, Ye
Published: (2025)
by: Pang, Ye
Published: (2025)
MotionHint: Self-Supervised Monocular Visual Odometry with Motion Constraints
by: Wang, Cong, et al.
Published: (2021)
by: Wang, Cong, et al.
Published: (2021)
High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects
by: Xue, Jialong, et al.
Published: (2025)
by: Xue, Jialong, et al.
Published: (2025)
TrajBooster: Boosting Humanoid Whole-Body Manipulation via Trajectory-Centric Learning
by: Liu, Jiacheng, et al.
Published: (2025)
by: Liu, Jiacheng, et al.
Published: (2025)
WildLMa: Long Horizon Loco-Manipulation in the Wild
by: Qiu, Ri-Zhao, et al.
Published: (2024)
by: Qiu, Ri-Zhao, et al.
Published: (2024)
SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy
by: Yao, Wei, et al.
Published: (2026)
by: Yao, Wei, et al.
Published: (2026)
From Motion to Behavior: Hierarchical Modeling of Humanoid Generative Behavior Control
by: Zhang, Jusheng, et al.
Published: (2025)
by: Zhang, Jusheng, et al.
Published: (2025)
Emergent Active Perception and Dexterity of Simulated Humanoids from Visual Reinforcement Learning
by: Luo, Zhengyi, et al.
Published: (2025)
by: Luo, Zhengyi, et al.
Published: (2025)
Iterative Closed-Loop Motion Synthesis for Scaling the Capabilities of Humanoid Control
by: Xu, Weisheng, et al.
Published: (2026)
by: Xu, Weisheng, et al.
Published: (2026)
Diagnose, Correct, and Learn from Manipulation Failures via Visual Symbols
by: Zeng, Xianchao, et al.
Published: (2025)
by: Zeng, Xianchao, et al.
Published: (2025)
MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence
by: Tang, Chao, et al.
Published: (2025)
by: Tang, Chao, et al.
Published: (2025)
TWIST: Teleoperated Whole-Body Imitation System
by: Ze, Yanjie, et al.
Published: (2025)
by: Ze, Yanjie, et al.
Published: (2025)
From Language to Locomotion: Retargeting-free Humanoid Control via Motion Latent Guidance
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
Empowering Embodied Visual Tracking with Visual Foundation Models and Offline RL
by: Zhong, Fangwei, et al.
Published: (2024)
by: Zhong, Fangwei, et al.
Published: (2024)
BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities
by: Jiang, Yunfan, et al.
Published: (2025)
by: Jiang, Yunfan, et al.
Published: (2025)
Mash, Spread, Slice! Learning to Manipulate Object States via Visual Spatial Progress
by: Mandikal, Priyanka, et al.
Published: (2025)
by: Mandikal, Priyanka, et al.
Published: (2025)
Think Proprioceptively: Embodied Visual Reasoning for VLA Manipulation
by: Wang, Fangyuan, et al.
Published: (2026)
by: Wang, Fangyuan, et al.
Published: (2026)
Hierarchical World Models as Visual Whole-Body Humanoid Controllers
by: Hansen, Nicklas, et al.
Published: (2024)
by: Hansen, Nicklas, et al.
Published: (2024)
Mitigating the Human-Robot Domain Discrepancy in Visual Pre-training for Robotic Manipulation
by: Zhou, Jiaming, et al.
Published: (2024)
by: Zhou, Jiaming, et al.
Published: (2024)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
VIMD: Monocular Visual-Inertial Motion and Depth Estimation
by: Katragadda, Saimouli, et al.
Published: (2025)
by: Katragadda, Saimouli, et al.
Published: (2025)
Similar Items
-
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation
by: Dong, Runpei, et al.
Published: (2026) -
ResMimic: From General Motion Tracking to Humanoid Whole-body Loco-Manipulation via Residual Learning
by: Zhao, Siheng, et al.
Published: (2025) -
Generalizable Humanoid Manipulation with 3D Diffusion Policies
by: Ze, Yanjie, et al.
Published: (2024) -
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation
by: He, Xialin, et al.
Published: (2026) -
Visual Whole-Body Control for Legged Loco-Manipulation
by: Liu, Minghuan, et al.
Published: (2024)