Think Before You Move: Latent Motion Reasoning for Text-to-Motion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qian, Yijie, Wang, Juncheng, Feng, Yuxiang, Xu, Chao, Lu, Wang, Liu, Yang, Sun, Baigui, Chen, Yiqiang, Liu, Yong, Wang, Shujun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
NEWTON: Agentic Planning for Physically Grounded Video Generation
von: Feng, Yuxiang, et al.
Veröffentlicht: (2026)
von: Feng, Yuxiang, et al.
Veröffentlicht: (2026)
Guided by the Plan: Enhancing Faithful Autoregressive Text-to-Audio Generation with Guided Decoding
von: Wang, Juncheng, et al.
Veröffentlicht: (2026)
von: Wang, Juncheng, et al.
Veröffentlicht: (2026)
You Think, You ACT: The New Task of Arbitrary Text to Motion Generation
von: Wang, Runqi, et al.
Veröffentlicht: (2024)
von: Wang, Runqi, et al.
Veröffentlicht: (2024)
Beyond Talking -- Generating Holistic 3D Human Dyadic Motion for Communication
von: Sun, Mingze, et al.
Veröffentlicht: (2024)
von: Sun, Mingze, et al.
Veröffentlicht: (2024)
TryOn-Adapter: Efficient Fine-Grained Clothing Identity Adaptation for High-Fidelity Virtual Try-On
von: Xing, Jiazheng, et al.
Veröffentlicht: (2024)
von: Xing, Jiazheng, et al.
Veröffentlicht: (2024)
Wan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance
von: Chu, Ruihang, et al.
Veröffentlicht: (2025)
von: Chu, Ruihang, et al.
Veröffentlicht: (2025)
Move-in-2D: 2D-Conditioned Human Motion Generation
von: Huang, Hsin-Ping, et al.
Veröffentlicht: (2024)
von: Huang, Hsin-Ping, et al.
Veröffentlicht: (2024)
FlowCoMotion: Text-to-Motion Generation via Token-Latent Flow Modeling
von: Guan, Dawei, et al.
Veröffentlicht: (2026)
von: Guan, Dawei, et al.
Veröffentlicht: (2026)
Motion Before Action: Diffusing Object Motion as Manipulation Condition
von: Su, Yue, et al.
Veröffentlicht: (2024)
von: Su, Yue, et al.
Veröffentlicht: (2024)
UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
Think Before You Prune: Self-Reflective Structured Pruning for Reasoning Language Models
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance
von: Wang, Zan, et al.
Veröffentlicht: (2024)
von: Wang, Zan, et al.
Veröffentlicht: (2024)
Think Before You Accept: Semantic Reflective Verification for Faster Speculative Decoding
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents
von: Liu, Dayong, et al.
Veröffentlicht: (2025)
von: Liu, Dayong, et al.
Veröffentlicht: (2025)
Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think
von: Tian, Jie, et al.
Veröffentlicht: (2025)
von: Tian, Jie, et al.
Veröffentlicht: (2025)
Exploring Motion-Language Alignment for Text-driven Motion Generation
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
Think Before Recommend: Unleashing the Latent Reasoning Power for Sequential Recommendation
von: Tang, Jiakai, et al.
Veröffentlicht: (2025)
von: Tang, Jiakai, et al.
Veröffentlicht: (2025)
TriC-Motion: Tri-Domain Causal Modeling Grounded Text-to-Motion Generation
von: Cao, Yiyang, et al.
Veröffentlicht: (2026)
von: Cao, Yiyang, et al.
Veröffentlicht: (2026)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
Generative Latent Kernel Modeling for Blind Motion Deblurring
von: Ding, Chenhao, et al.
Veröffentlicht: (2025)
von: Ding, Chenhao, et al.
Veröffentlicht: (2025)
MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model
von: Dai, Wenxun, et al.
Veröffentlicht: (2024)
von: Dai, Wenxun, et al.
Veröffentlicht: (2024)
Think Before You Prune: Selective Self-Generated Calibration for Pruning Large Reasoning Models
von: Xiang, Yang, et al.
Veröffentlicht: (2025)
von: Xiang, Yang, et al.
Veröffentlicht: (2025)
Language Model Based Text-to-Audio Generation: Anti-Causally Aligned Collaborative Residual Transformers
von: Wang, Juncheng, et al.
Veröffentlicht: (2025)
von: Wang, Juncheng, et al.
Veröffentlicht: (2025)
Think Before You Lie: How Reasoning Leads to Honesty
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
Motion-aware Latent Diffusion Models for Video Frame Interpolation
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
Chain of World: World Model Thinking in Latent Motion
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
Walk Before You Dance: High-fidelity and Editable Dance Synthesis via Generative Masked Motion Prior
von: Shah, Foram N, et al.
Veröffentlicht: (2025)
von: Shah, Foram N, et al.
Veröffentlicht: (2025)
MotionRFT: Unified Reinforcement Fine-Tuning for Text-to-Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
Grasping Motion Generation Through Latent Diffusion Models
von: X. Wang, et al.
Veröffentlicht: (2026)
von: X. Wang, et al.
Veröffentlicht: (2026)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
von: Ma, Minghui, et al.
Veröffentlicht: (2026)
von: Ma, Minghui, et al.
Veröffentlicht: (2026)
DeX-Portrait: Disentangled and Expressive Portrait Animation via Explicit and Latent Motion Representations
von: Shi, Yuxiang, et al.
Veröffentlicht: (2025)
von: Shi, Yuxiang, et al.
Veröffentlicht: (2025)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
Efficient Text-driven Motion Generation via Latent Consistency Training
von: Hu, Mengxian, et al.
Veröffentlicht: (2024)
von: Hu, Mengxian, et al.
Veröffentlicht: (2024)
MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space
von: Xiao, Lixing, et al.
Veröffentlicht: (2025)
von: Xiao, Lixing, et al.
Veröffentlicht: (2025)
Reinforcing Video Reasoning Segmentation to Think Before It Segments
von: Gong, Sitong, et al.
Veröffentlicht: (2025)
von: Gong, Sitong, et al.
Veröffentlicht: (2025)
Think Before You Segment: An Object-aware Reasoning Agent for Referring Audio-Visual Segmentation
von: Zhou, Jinxing, et al.
Veröffentlicht: (2025)
von: Zhou, Jinxing, et al.
Veröffentlicht: (2025)
MotionFlux: Efficient Text-Guided Motion Generation through Rectified Flow Matching and Preference Alignment
von: Gao, Zhiting, et al.
Veröffentlicht: (2025)
von: Gao, Zhiting, et al.
Veröffentlicht: (2025)
An Anatomy of Vision-Language-Action Models: From Modules to Milestones and Challenges
von: Xu, Chao, et al.
Veröffentlicht: (2025)
von: Xu, Chao, et al.
Veröffentlicht: (2025)
Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
Stop Before You Fail: Operational Capability Boundaries for Mitigating Unproductive Reasoning in Large Reasoning Models
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
NEWTON: Agentic Planning for Physically Grounded Video Generation
von: Feng, Yuxiang, et al.
Veröffentlicht: (2026) -
Guided by the Plan: Enhancing Faithful Autoregressive Text-to-Audio Generation with Guided Decoding
von: Wang, Juncheng, et al.
Veröffentlicht: (2026) -
You Think, You ACT: The New Task of Arbitrary Text to Motion Generation
von: Wang, Runqi, et al.
Veröffentlicht: (2024) -
Beyond Talking -- Generating Holistic 3D Human Dyadic Motion for Communication
von: Sun, Mingze, et al.
Veröffentlicht: (2024) -
TryOn-Adapter: Efficient Fine-Grained Clothing Identity Adaptation for High-Fidelity Virtual Try-On
von: Xing, Jiazheng, et al.
Veröffentlicht: (2024)