MoLingo: Motion-Language Alignment for Text-to-Motion Generation
Fuente:
arXiv
Saved in:
| Main Authors: | He, Yannan, Tiwari, Garvita, Zhang, Xiaohan, Bora, Pankaj, Birdal, Tolga, Lenssen, Jan Eric, Pons-Moll, Gerard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NRDF: Neural Riemannian Distance Fields for Learning Articulated Pose Priors
by: He, Yannan, et al.
Published: (2024)
by: He, Yannan, et al.
Published: (2024)
ActionPlan: Future-Aware Streaming Motion Synthesis via Frame-Level Action Planning
by: Nazarenus, Eric, et al.
Published: (2026)
by: Nazarenus, Eric, et al.
Published: (2026)
InterTrack: Tracking Human Object Interaction without Object Templates
by: Xie, Xianghui, et al.
Published: (2024)
by: Xie, Xianghui, et al.
Published: (2024)
CloSe: A 3D Clothing Segmentation Dataset and Model
by: Antić, Dimitrije, et al.
Published: (2024)
by: Antić, Dimitrije, et al.
Published: (2024)
Template Free Reconstruction of Human-object Interaction with Procedural Interaction Generation
by: Xie, Xianghui, et al.
Published: (2023)
by: Xie, Xianghui, et al.
Published: (2023)
FrankenMotion: Part-level Human Motion Generation and Composition
by: Li, Chuqiao, et al.
Published: (2026)
by: Li, Chuqiao, et al.
Published: (2026)
Interaction Replica: Tracking Human-Object Interaction and Scene Changes From Human Motion
by: Guzov, Vladimir, et al.
Published: (2022)
by: Guzov, Vladimir, et al.
Published: (2022)
Generating Continual Human Motion in Diverse 3D Scenes
by: Mir, Aymen, et al.
Published: (2023)
by: Mir, Aymen, et al.
Published: (2023)
Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera
by: Yu, Zhengdi, et al.
Published: (2024)
by: Yu, Zhengdi, et al.
Published: (2024)
SignAvatars: A Large-scale 3D Sign Language Holistic Motion Dataset and Benchmark
by: Yu, Zhengdi, et al.
Published: (2023)
by: Yu, Zhengdi, et al.
Published: (2023)
Unimotion: Unifying 3D Human Motion Synthesis and Understanding
by: Li, Chuqiao, et al.
Published: (2024)
by: Li, Chuqiao, et al.
Published: (2024)
LingoMotion: An Interpretable and Unambiguous Symbolic Representation for Human Motion
by: Zhang, Yao, et al.
Published: (2026)
by: Zhang, Yao, et al.
Published: (2026)
Exploring Motion-Language Alignment for Text-driven Motion Generation
by: Gu, Ruxi, et al.
Published: (2026)
by: Gu, Ruxi, et al.
Published: (2026)
HyperSDFusion: Bridging Hierarchical Structures in Language and Geometry for Enhanced 3D Text2Shape Generation
by: Leng, Zhiying, et al.
Published: (2024)
by: Leng, Zhiying, et al.
Published: (2024)
GEARS: Local Geometry-aware Hand-object Interaction Synthesis
by: Zhou, Keyang, et al.
Published: (2024)
by: Zhou, Keyang, et al.
Published: (2024)
Geometric Neural Distance Fields for Learning Human Motion Priors
by: Yu, Zhengdi, et al.
Published: (2025)
by: Yu, Zhengdi, et al.
Published: (2025)
Neural Localizer Fields for Continuous 3D Human Pose and Shape Estimation
by: Sárándi, István, et al.
Published: (2024)
by: Sárándi, István, et al.
Published: (2024)
HMD^2: Environment-aware Motion Generation from Single Egocentric Head-Mounted Device
by: Guzov, Vladimir, et al.
Published: (2024)
by: Guzov, Vladimir, et al.
Published: (2024)
UV-free Texture Generation with Denoising and Geodesic Heat Diffusions
by: Foti, Simone, et al.
Published: (2024)
by: Foti, Simone, et al.
Published: (2024)
Recent Trends in 3D Reconstruction of General Non-Rigid Scenes
by: Yunus, Raza, et al.
Published: (2024)
by: Yunus, Raza, et al.
Published: (2024)
CoMo: Compositional Motion Customization for Text-to-Video Generation
by: Xu, Youcan, et al.
Published: (2025)
by: Xu, Youcan, et al.
Published: (2025)
Paint-it: Text-to-Texture Synthesis via Deep Convolutional Texture Map Optimization and Physically-Based Rendering
by: Youwang, Kim, et al.
Published: (2023)
by: Youwang, Kim, et al.
Published: (2023)
NICP: Neural ICP for 3D Human Registration at Scale
by: Marin, Riccardo, et al.
Published: (2023)
by: Marin, Riccardo, et al.
Published: (2023)
MotionFlux: Efficient Text-Guided Motion Generation through Rectified Flow Matching and Preference Alignment
by: Gao, Zhiting, et al.
Published: (2025)
by: Gao, Zhiting, et al.
Published: (2025)
SnapMoGen: Human Motion Generation from Expressive Texts
by: Guo, Chuan, et al.
Published: (2025)
by: Guo, Chuan, et al.
Published: (2025)
SafeMo: Linguistically Grounded Unlearning for Trustworthy Text-to-Motion Generation
by: Wang, Yiling, et al.
Published: (2026)
by: Wang, Yiling, et al.
Published: (2026)
Gen-3Diffusion: Realistic Image-to-3D Generation via 2D & 3D Diffusion Synergy
by: Xue, Yuxuan, et al.
Published: (2024)
by: Xue, Yuxuan, et al.
Published: (2024)
OmniMoGen: Unifying Human Motion Generation via Learning from Interleaved Text-Motion Instructions
by: Bu, Wendong, et al.
Published: (2025)
by: Bu, Wendong, et al.
Published: (2025)
CLUTCH: Contextualized Language model for Unlocking Text-Conditioned Hand motion modelling in the wild
by: Thambiraja, Balamurugan, et al.
Published: (2026)
by: Thambiraja, Balamurugan, et al.
Published: (2026)
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
by: Li, Yuan-Ming, et al.
Published: (2025)
by: Li, Yuan-Ming, et al.
Published: (2025)
Generalization at the Edge of Stability
by: Tuci, Mario, et al.
Published: (2026)
by: Tuci, Mario, et al.
Published: (2026)
MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps
by: Lei, Jiahui, et al.
Published: (2025)
by: Lei, Jiahui, et al.
Published: (2025)
SegMo: Segment-aligned Text to 3D Human Motion Generation
by: Dang, Bowen, et al.
Published: (2025)
by: Dang, Bowen, et al.
Published: (2025)
CrowdMoGen: Zero-Shot Text-Driven Collective Motion Generation
by: Cao, Yukang, et al.
Published: (2024)
by: Cao, Yukang, et al.
Published: (2024)
SCENIC: Scene-aware Semantic Navigation with Instruction-guided Control
by: Zhang, Xiaohan, et al.
Published: (2024)
by: Zhang, Xiaohan, et al.
Published: (2024)
SimLingo: Vision-Only Closed-Loop Autonomous Driving with Language-Action Alignment
by: Renz, Katrin, et al.
Published: (2025)
by: Renz, Katrin, et al.
Published: (2025)
MoTrans: Customized Motion Transfer with Text-driven Video Diffusion Models
by: Li, Xiaomin, et al.
Published: (2024)
by: Li, Xiaomin, et al.
Published: (2024)
Forecasting Continuous Non-Conservative Dynamical Systems in SO(3)
by: Bastian, Lennart, et al.
Published: (2025)
by: Bastian, Lennart, et al.
Published: (2025)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
by: Wu, Jianzong, et al.
Published: (2024)
by: Wu, Jianzong, et al.
Published: (2024)
MoCHA: Denoising Caption Supervision for Motion-Text Retrieval
by: Warner, Nikolai, et al.
Published: (2026)
by: Warner, Nikolai, et al.
Published: (2026)
Similar Items
-
NRDF: Neural Riemannian Distance Fields for Learning Articulated Pose Priors
by: He, Yannan, et al.
Published: (2024) -
ActionPlan: Future-Aware Streaming Motion Synthesis via Frame-Level Action Planning
by: Nazarenus, Eric, et al.
Published: (2026) -
InterTrack: Tracking Human Object Interaction without Object Templates
by: Xie, Xianghui, et al.
Published: (2024) -
CloSe: A 3D Clothing Segmentation Dataset and Model
by: Antić, Dimitrije, et al.
Published: (2024) -
Template Free Reconstruction of Human-object Interaction with Procedural Interaction Generation
by: Xie, Xianghui, et al.
Published: (2023)