OmniMotion-X: Versatile Multimodal Whole-Body Motion Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Guowei, Bian, Yuxuan, Zeng, Ailing, Shi, Mingyi, Huang, Shaoli, Li, Wen, Duan, Lixin, Xu, Qiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls
by: Bian, Yuxuan, et al.
Published: (2024)
by: Bian, Yuxuan, et al.
Published: (2024)
Learning Semantic Latent Directions for Accurate and Controllable Human Motion Prediction
by: Xu, Guowei, et al.
Published: (2024)
by: Xu, Guowei, et al.
Published: (2024)
Motion-X++: A Large-Scale Multimodal 3D Whole-body Human Motion Dataset
by: Zhang, Yuhong, et al.
Published: (2025)
by: Zhang, Yuhong, et al.
Published: (2025)
Motion-X: A Large-scale 3D Expressive Whole-body Human Motion Dataset
by: Lin, Jing, et al.
Published: (2023)
by: Lin, Jing, et al.
Published: (2023)
Realistic Human Motion Generation with Cross-Diffusion Models
by: Ren, Zeping, et al.
Published: (2023)
by: Ren, Zeping, et al.
Published: (2023)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
by: Cheng, Yongkang, et al.
Published: (2025)
by: Cheng, Yongkang, et al.
Published: (2025)
Programmable Motion Generation for Open-Set Motion Control Tasks
by: Liu, Hanchao, et al.
Published: (2024)
by: Liu, Hanchao, et al.
Published: (2024)
Human-Object Interaction via Automatically Designed VLM-Guided Motion Policy
by: Deng, Zekai, et al.
Published: (2025)
by: Deng, Zekai, et al.
Published: (2025)
Dynamic Motion Blending for Versatile Motion Editing
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
Holistic-Motion2D: Scalable Whole-body Human Motion Generation in 2D Space
by: Wang, Yuan, et al.
Published: (2024)
by: Wang, Yuan, et al.
Published: (2024)
S-INF: Towards Realistic Indoor Scene Synthesis via Scene Implicit Neural Field
by: Liang, Zixi, et al.
Published: (2024)
by: Liang, Zixi, et al.
Published: (2024)
Diffgrasp: Whole-Body Grasping Synthesis Guided by Object Motion Using a Diffusion Model
by: Zhang, Yonghao, et al.
Published: (2024)
by: Zhang, Yonghao, et al.
Published: (2024)
MotionLLM: Understanding Human Behaviors from Human Motions and Videos
by: Chen, Ling-Hao, et al.
Published: (2024)
by: Chen, Ling-Hao, et al.
Published: (2024)
A Unified Transformer-Based Framework with Pretraining For Whole Body Grasping Motion Generation
by: Effendy, Edward, et al.
Published: (2025)
by: Effendy, Edward, et al.
Published: (2025)
Lagrangian Motion Fields for Long-term Motion Generation
by: Yang, Yifei, et al.
Published: (2024)
by: Yang, Yifei, et al.
Published: (2024)
Markerless Motion Capture for Biomechanical Whole-Body Kinematic Estimation in Infants
by: Joshi, Divya, et al.
Published: (2026)
by: Joshi, Divya, et al.
Published: (2026)
SignAvatars: A Large-scale 3D Sign Language Holistic Motion Dataset and Benchmark
by: Yu, Zhengdi, et al.
Published: (2023)
by: Yu, Zhengdi, et al.
Published: (2023)
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation
by: He, Xialin, et al.
Published: (2026)
by: He, Xialin, et al.
Published: (2026)
RapVerse: Coherent Vocals and Whole-Body Motions Generations from Text
by: Chen, Jiaben, et al.
Published: (2024)
by: Chen, Jiaben, et al.
Published: (2024)
REWIND: Real-Time Egocentric Whole-Body Motion Diffusion with Exemplar-Based Identity Conditioning
by: Lee, Jihyun, et al.
Published: (2025)
by: Lee, Jihyun, et al.
Published: (2025)
VersatileMotion: A Unified Framework for Motion Synthesis and Comprehension
by: Ling, Zeyu, et al.
Published: (2024)
by: Ling, Zeyu, et al.
Published: (2024)
Make Your Actor Talk: Generalizable and High-Fidelity Lip Sync with Motion and Appearance Disentanglement
by: Yu, Runyi, et al.
Published: (2024)
by: Yu, Runyi, et al.
Published: (2024)
ReinDiffuse: Crafting Physically Plausible Motions with Reinforced Diffusion Model
by: Han, Gaoge, et al.
Published: (2024)
by: Han, Gaoge, et al.
Published: (2024)
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
by: Li, Yuan-Ming, et al.
Published: (2025)
by: Li, Yuan-Ming, et al.
Published: (2025)
The Quest for Generalizable Motion Generation: Data, Model, and Evaluation
by: Lin, Jing, et al.
Published: (2025)
by: Lin, Jing, et al.
Published: (2025)
MOSPA: Human Motion Generation Driven by Spatial Audio
by: Xu, Shuyang, et al.
Published: (2025)
by: Xu, Shuyang, et al.
Published: (2025)
OmniMoGen: Unifying Human Motion Generation via Learning from Interleaved Text-Motion Instructions
by: Bu, Wendong, et al.
Published: (2025)
by: Bu, Wendong, et al.
Published: (2025)
OmniCam: Unified Multimodal Video Generation via Camera Control
by: Yang, Xiaoda, et al.
Published: (2025)
by: Yang, Xiaoda, et al.
Published: (2025)
OmniDrag: Enabling Motion Control for Omnidirectional Image-to-Video Generation
by: Li, Weiqi, et al.
Published: (2024)
by: Li, Weiqi, et al.
Published: (2024)
Progressive Human Motion Generation Based on Text and Few Motion Frames
by: Zeng, Ling-An, et al.
Published: (2025)
by: Zeng, Ling-An, et al.
Published: (2025)
EMHI: A Multimodal Egocentric Human Motion Dataset with HMD and Body-Worn IMUs
by: Fan, Zhen, et al.
Published: (2024)
by: Fan, Zhen, et al.
Published: (2024)
MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding
by: Wang, Yuan, et al.
Published: (2024)
by: Wang, Yuan, et al.
Published: (2024)
OralGPT-Omni: A Versatile Dental Multimodal Large Language Model
by: Hao, Jing, et al.
Published: (2025)
by: Hao, Jing, et al.
Published: (2025)
HandX: Scaling Bimanual Motion and Interaction Generation
by: Zhang, Zimu, et al.
Published: (2026)
by: Zhang, Zimu, et al.
Published: (2026)
MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm
by: Guo, Ziyan, et al.
Published: (2025)
by: Guo, Ziyan, et al.
Published: (2025)
RAGME: Retrieval Augmented Video Generation for Enhanced Motion Realism
by: Peruzzo, Elia, et al.
Published: (2025)
by: Peruzzo, Elia, et al.
Published: (2025)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
by: Shi, Junyu, et al.
Published: (2025)
by: Shi, Junyu, et al.
Published: (2025)
Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game
by: Kim, Jeonghwan, et al.
Published: (2025)
by: Kim, Jeonghwan, et al.
Published: (2025)
Motion-I2V: Consistent and Controllable Image-to-Video Generation with Explicit Motion Modeling
by: Shi, Xiaoyu, et al.
Published: (2024)
by: Shi, Xiaoyu, et al.
Published: (2024)
Similar Items
-
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
by: Li, Zhe, et al.
Published: (2025) -
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls
by: Bian, Yuxuan, et al.
Published: (2024) -
Learning Semantic Latent Directions for Accurate and Controllable Human Motion Prediction
by: Xu, Guowei, et al.
Published: (2024) -
Motion-X++: A Large-Scale Multimodal 3D Whole-body Human Motion Dataset
by: Zhang, Yuhong, et al.
Published: (2025) -
Motion-X: A Large-scale 3D Expressive Whole-body Human Motion Dataset
by: Lin, Jing, et al.
Published: (2023)