MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yi-Yang, Sun, Tengjiao, Fang, Pengcheng, Wang, Deng-Bao, Cai, Xiaohao, Zhang, Min-Ling, Kim, Hansung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AnchorRoute: Human Motion Synthesis with Interval-Routed Sparse Contro
by: Fang, Pengcheng, et al.
Published: (2026)
by: Fang, Pengcheng, et al.
Published: (2026)
UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars
by: Zhan, Xiaoyu, et al.
Published: (2026)
by: Zhan, Xiaoyu, et al.
Published: (2026)
MOGO: Residual Quantized Hierarchical Causal Transformer for High-Quality and Real-Time 3D Human Motion Generation
by: Fu, Dongjie, et al.
Published: (2025)
by: Fu, Dongjie, et al.
Published: (2025)
PartMotionEdit: Fine-Grained Text-Driven 3D Human Motion Editing via Part-Level Modulation
by: Yang, Yujie, et al.
Published: (2025)
by: Yang, Yujie, et al.
Published: (2025)
FlowMotion: Target-Predictive Conditional Flow Matching for Jitter-Reduced Text-Driven Human Motion Generation
by: Cuba, Manolo Canales, et al.
Published: (2025)
by: Cuba, Manolo Canales, et al.
Published: (2025)
MotionFix: Text-Driven 3D Human Motion Editing
by: Athanasiou, Nikos, et al.
Published: (2024)
by: Athanasiou, Nikos, et al.
Published: (2024)
Learning Human Motion from Monocular Videos via Cross-Modal Manifold Alignment
by: Hou, Shuaiying, et al.
Published: (2024)
by: Hou, Shuaiying, et al.
Published: (2024)
Event-T2M: Event-level Conditioning for Complex Text-to-Motion Synthesis
by: Hong, Seong-Eun, et al.
Published: (2026)
by: Hong, Seong-Eun, et al.
Published: (2026)
Generating Human Interaction Motions in Scenes with Text Control
by: Yi, Hongwei, et al.
Published: (2024)
by: Yi, Hongwei, et al.
Published: (2024)
Target Pose Guided Whole-body Grasping Motion Generation for Digital Humans
by: Shao, Quanquan, et al.
Published: (2024)
by: Shao, Quanquan, et al.
Published: (2024)
The Last Mile to Production Readiness: Physics-Based Motion Refinement for Video-Based Capture
by: Tao, Tianxin, et al.
Published: (2026)
by: Tao, Tianxin, et al.
Published: (2026)
TransVDM: Motion-Constrained Video Diffusion Model for Transparent Video Synthesis
by: Li, Menghao, et al.
Published: (2025)
by: Li, Menghao, et al.
Published: (2025)
MDD: A Dataset for Text-and-Music Conditioned Duet Dance Generation
by: Gupta, Prerit, et al.
Published: (2025)
by: Gupta, Prerit, et al.
Published: (2025)
Object-Aware 4D Human Motion Generation
by: Gui, Shurui, et al.
Published: (2025)
by: Gui, Shurui, et al.
Published: (2025)
Improving Global Motion Estimation in Sparse IMU-based Motion Capture with Physics
by: Yi, Xinyu, et al.
Published: (2025)
by: Yi, Xinyu, et al.
Published: (2025)
StableMotion: Training Motion Cleanup Models with Unpaired Corrupted Data
by: Mu, Yuxuan, et al.
Published: (2025)
by: Mu, Yuxuan, et al.
Published: (2025)
LGTM: Local-to-Global Text-Driven Human Motion Diffusion Model
by: Sun, Haowen, et al.
Published: (2024)
by: Sun, Haowen, et al.
Published: (2024)
IKMo: Image-Keyframed Motion Generation with Trajectory-Pose Conditioned Motion Diffusion Model
by: Zhao, Yang, et al.
Published: (2025)
by: Zhao, Yang, et al.
Published: (2025)
Shape Conditioned Human Motion Generation with Diffusion Model
by: Xue, Kebing, et al.
Published: (2024)
by: Xue, Kebing, et al.
Published: (2024)
Beyond Static Scenes: Camera-controllable Background Generation for Human Motion
by: Yao, Mingshuai, et al.
Published: (2025)
by: Yao, Mingshuai, et al.
Published: (2025)
MotionGS: Exploring Explicit Motion Guidance for Deformable 3D Gaussian Splatting
by: Zhu, Ruijie, et al.
Published: (2024)
by: Zhu, Ruijie, et al.
Published: (2024)
MotionWavelet: Human Motion Prediction via Wavelet Manifold Learning
by: Feng, Yuming, et al.
Published: (2024)
by: Feng, Yuming, et al.
Published: (2024)
MotionAgent: Fine-grained Controllable Video Generation via Motion Field Agent
by: Liao, Xinyao, et al.
Published: (2025)
by: Liao, Xinyao, et al.
Published: (2025)
Physical Non-inertial Poser (PNP): Modeling Non-inertial Effects in Sparse-inertial Human Motion Capture
by: Yi, Xinyu, et al.
Published: (2024)
by: Yi, Xinyu, et al.
Published: (2024)
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation
by: Cheng, Shihao, et al.
Published: (2026)
by: Cheng, Shihao, et al.
Published: (2026)
HY-Motion 1.0: Scaling Flow Matching Models for Text-To-Motion Generation
by: Wen, Yuxin, et al.
Published: (2025)
by: Wen, Yuxin, et al.
Published: (2025)
Dynamic Motion Synthesis: Masked Audio-Text Conditioned Spatio-Temporal Transformers
by: Anisetty, Sohan, et al.
Published: (2024)
by: Anisetty, Sohan, et al.
Published: (2024)
MAMM: Motion Control via Metric-Aligning Motion Matching
by: Agata, Naoki, et al.
Published: (2025)
by: Agata, Naoki, et al.
Published: (2025)
Improving Sparse IMU-based Motion Capture with Motion Label Smoothing
by: Meng, Zhaorui, et al.
Published: (2025)
by: Meng, Zhaorui, et al.
Published: (2025)
AnyAct: Towards Human Reenactment of Character Motion From Video
by: Chen, Liuhan, et al.
Published: (2026)
by: Chen, Liuhan, et al.
Published: (2026)
Motion-2-To-3: Leveraging 2D Motion Data for 3D Motion Generations
by: Guo, Ruoxi, et al.
Published: (2024)
by: Guo, Ruoxi, et al.
Published: (2024)
Sound Sparks Motion: Audio and Text Tuning for Video Editing
by: Razlighi, AmirHossein Naghi, et al.
Published: (2026)
by: Razlighi, AmirHossein Naghi, et al.
Published: (2026)
Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation
by: Petrovich, Mathis, et al.
Published: (2024)
by: Petrovich, Mathis, et al.
Published: (2024)
KinMo: Kinematic-aware Human Motion Understanding and Generation
by: Zhang, Pengfei, et al.
Published: (2024)
by: Zhang, Pengfei, et al.
Published: (2024)
Synthesizing Physically Plausible Human Motions in 3D Scenes
by: Pan, Liang, et al.
Published: (2023)
by: Pan, Liang, et al.
Published: (2023)
GaussianPrediction: Dynamic 3D Gaussian Prediction for Motion Extrapolation and Free View Synthesis
by: Zhao, Boming, et al.
Published: (2024)
by: Zhao, Boming, et al.
Published: (2024)
Physics-Based Motion Tracking of Contact-Rich Interacting Characters
by: Zhang, Xiaotang, et al.
Published: (2026)
by: Zhang, Xiaotang, et al.
Published: (2026)
DP-Adapter: Dual-Pathway Adapter for Boosting Fidelity and Text Consistency in Customizable Human Image Generation
by: Wang, Ye, et al.
Published: (2025)
by: Wang, Ye, et al.
Published: (2025)
Real-time and Controllable Reactive Motion Synthesis via Intention Guidance
by: Zhang, Xiaotang, et al.
Published: (2025)
by: Zhang, Xiaotang, et al.
Published: (2025)
MOGRAS: Human Motion with Grasping in 3D Scenes
by: Bhosikar, Kunal, et al.
Published: (2025)
by: Bhosikar, Kunal, et al.
Published: (2025)
Similar Items
-
AnchorRoute: Human Motion Synthesis with Interval-Routed Sparse Contro
by: Fang, Pengcheng, et al.
Published: (2026) -
UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars
by: Zhan, Xiaoyu, et al.
Published: (2026) -
MOGO: Residual Quantized Hierarchical Causal Transformer for High-Quality and Real-Time 3D Human Motion Generation
by: Fu, Dongjie, et al.
Published: (2025) -
PartMotionEdit: Fine-Grained Text-Driven 3D Human Motion Editing via Part-Level Modulation
by: Yang, Yujie, et al.
Published: (2025) -
FlowMotion: Target-Predictive Conditional Flow Matching for Jitter-Reduced Text-Driven Human Motion Generation
by: Cuba, Manolo Canales, et al.
Published: (2025)