The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Changan, Zhang, Juze, Lakshmikanth, Shrinidhi K., Fang, Yusu, Shao, Ruizhi, Wetzstein, Gordon, Fei-Fei, Li, Adeli, Ehsan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body
by: Zhang, Juze, et al.
Published: (2025)
by: Zhang, Juze, et al.
Published: (2025)
SocialGen: Modeling Multi-Human Social Interaction with Language Models
by: Yu, Heng, et al.
Published: (2025)
by: Yu, Heng, et al.
Published: (2025)
HumanScore: Benchmarking Human Motions in Generated Videos
by: Fang, Yusu, et al.
Published: (2026)
by: Fang, Yusu, et al.
Published: (2026)
Interspatial Attention for Efficient 4D Human Video Generation
by: Shao, Ruizhi, et al.
Published: (2025)
by: Shao, Ruizhi, et al.
Published: (2025)
UniEgoMotion: A Unified Model for Egocentric Motion Reconstruction, Forecasting, and Generation
by: Patel, Chaitanya, et al.
Published: (2025)
by: Patel, Chaitanya, et al.
Published: (2025)
UniMo: Unifying 2D Video and 3D Human Motion with an Autoregressive Framework
by: Pang, Youxin, et al.
Published: (2025)
by: Pang, Youxin, et al.
Published: (2025)
OccFusion: Rendering Occluded Humans with Generative Diffusion Priors
by: Sun, Adam, et al.
Published: (2024)
by: Sun, Adam, et al.
Published: (2024)
Latent Forcing: Reordering the Diffusion Trajectory for Pixel-Space Image Generation
by: Baade, Alan, et al.
Published: (2026)
by: Baade, Alan, et al.
Published: (2026)
SignAvatar: Sign Language 3D Motion Reconstruction and Generation
by: Dong, Lu, et al.
Published: (2024)
by: Dong, Lu, et al.
Published: (2024)
Wild2Avatar: Rendering Humans Behind Occlusions
by: Xiang, Tiange, et al.
Published: (2023)
by: Xiang, Tiange, et al.
Published: (2023)
Exploring Vision Transformers for 3D Human Motion-Language Models with Motion Patches
by: Yu, Qing, et al.
Published: (2024)
by: Yu, Qing, et al.
Published: (2024)
BAgger: Backwards Aggregation for Mitigating Drift in Autoregressive Video Diffusion Models
by: Po, Ryan, et al.
Published: (2025)
by: Po, Ryan, et al.
Published: (2025)
Repurposing 2D Diffusion Models with Gaussian Atlas for 3D Generation
by: Xiang, Tiange, et al.
Published: (2025)
by: Xiang, Tiange, et al.
Published: (2025)
Unimotion: Unifying 3D Human Motion Synthesis and Understanding
by: Li, Chuqiao, et al.
Published: (2024)
by: Li, Chuqiao, et al.
Published: (2024)
MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm
by: Guo, Ziyan, et al.
Published: (2025)
by: Guo, Ziyan, et al.
Published: (2025)
MotionScript: Natural Language Descriptions for Expressive 3D Human Motions
by: Yazdian, Payam Jome, et al.
Published: (2023)
by: Yazdian, Payam Jome, et al.
Published: (2023)
Human-in-Context: Unified Cross-Domain 3D Human Motion Modeling via In-Context Learning
by: Liu, Mengyuan, et al.
Published: (2025)
by: Liu, Mengyuan, et al.
Published: (2025)
AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion
by: Li, Hongjie, et al.
Published: (2026)
by: Li, Hongjie, et al.
Published: (2026)
ByteMorph: Benchmarking Instruction-Guided Image Editing with Non-Rigid Motions
by: Chang, Di, et al.
Published: (2025)
by: Chang, Di, et al.
Published: (2025)
VimoRAG: Video-based Retrieval-augmented 3D Motion Generation for Motion Language Models
by: Xu, Haidong, et al.
Published: (2025)
by: Xu, Haidong, et al.
Published: (2025)
AdaVid: Adaptive Video-Language Pretraining
by: Patel, Chaitanya, et al.
Published: (2025)
by: Patel, Chaitanya, et al.
Published: (2025)
MOSS: Motion-based 3D Clothed Human Synthesis from Monocular Video
by: Wang, Hongsheng, et al.
Published: (2024)
by: Wang, Hongsheng, et al.
Published: (2024)
LangPose: Language-Aligned Motion for Robust 3D Human Pose Estimation
by: Liao, Longyun, et al.
Published: (2024)
by: Liao, Longyun, et al.
Published: (2024)
DrawMotion: Generating 3D Human Motions by Freehand Drawing
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation
by: Wang, Xinshun, et al.
Published: (2026)
by: Wang, Xinshun, et al.
Published: (2026)
Human Motion Synthesis in 3D Scenes via Unified Scene Semantic Occupancy
by: Jingyu, Gong, et al.
Published: (2025)
by: Jingyu, Gong, et al.
Published: (2025)
Towards Robust 3D Pose Transfer with Adversarial Learning
by: Chen, Haoyu, et al.
Published: (2024)
by: Chen, Haoyu, et al.
Published: (2024)
Diffusion MRI Transformer with a Diffusion Space Rotary Positional Embedding (D-RoPE)
by: Kung, Gustavo Chau Loo, et al.
Published: (2026)
by: Kung, Gustavo Chau Loo, et al.
Published: (2026)
MotionRFT: Unified Reinforcement Fine-Tuning for Text-to-Motion Generation
by: Tan, Xiaofeng, et al.
Published: (2026)
by: Tan, Xiaofeng, et al.
Published: (2026)
MotionFix: Text-Driven 3D Human Motion Editing
by: Athanasiou, Nikos, et al.
Published: (2024)
by: Athanasiou, Nikos, et al.
Published: (2024)
LaxMotion: Rethinking Supervision Granularity for 3D Human Motion Generation
by: Liu, Sheng, et al.
Published: (2025)
by: Liu, Sheng, et al.
Published: (2025)
Modelling the Distribution of Human Motion for Sign Language Assessment
by: Cory, Oliver, et al.
Published: (2024)
by: Cory, Oliver, et al.
Published: (2024)
Language-Guided Transformer Tokenizer for Human Motion Generation
by: Yan, Sheng, et al.
Published: (2026)
by: Yan, Sheng, et al.
Published: (2026)
Human Motion Video Generation: A Survey
by: Xue, Haiwei, et al.
Published: (2025)
by: Xue, Haiwei, et al.
Published: (2025)
FreeMotion: MoCap-Free Human Motion Synthesis with Multimodal Large Language Models
by: Zhang, Zhikai, et al.
Published: (2024)
by: Zhang, Zhikai, et al.
Published: (2024)
Uni-Inter: Unifying 3D Human Motion Synthesis Across Diverse Interaction Contexts
by: Liu, Sheng, et al.
Published: (2025)
by: Liu, Sheng, et al.
Published: (2025)
MotionCharacter: Fine-Grained Motion Controllable Human Video Generation
by: Fang, Haopeng, et al.
Published: (2024)
by: Fang, Haopeng, et al.
Published: (2024)
Improving 3D Foot Motion Reconstruction in Markerless Monocular Human Motion Capture
by: Wehrbein, Tom, et al.
Published: (2026)
by: Wehrbein, Tom, et al.
Published: (2026)
IoT-Based 3D Pose Estimation and Motion Optimization for Athletes: Application of C3D and OpenPose
by: Ren, Fei, et al.
Published: (2024)
by: Ren, Fei, et al.
Published: (2024)
Repurposing 2D Diffusion Models for 3D Shape Completion
by: He, Yao, et al.
Published: (2025)
by: He, Yao, et al.
Published: (2025)
Similar Items
-
ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body
by: Zhang, Juze, et al.
Published: (2025) -
SocialGen: Modeling Multi-Human Social Interaction with Language Models
by: Yu, Heng, et al.
Published: (2025) -
HumanScore: Benchmarking Human Motions in Generated Videos
by: Fang, Yusu, et al.
Published: (2026) -
Interspatial Attention for Efficient 4D Human Video Generation
by: Shao, Ruizhi, et al.
Published: (2025) -
UniEgoMotion: A Unified Model for Egocentric Motion Reconstruction, Forecasting, and Generation
by: Patel, Chaitanya, et al.
Published: (2025)