Motion Prompting: Controlling Video Generation with Motion Trajectories
Fuente:
arXiv
Saved in:
| Main Authors: | Geng, Daniel, Herrmann, Charles, Hur, Junhwa, Cole, Forrester, Zhang, Serena, Pfaff, Tobias, Lopez-Guevara, Tatiana, Doersch, Carl, Aytar, Yusuf, Rubinstein, Michael, Sun, Chen, Wang, Oliver, Owens, Andrew, Sun, Deqing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion
by: Zhang, Junyi, et al.
Published: (2024)
by: Zhang, Junyi, et al.
Published: (2024)
GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure
by: Gu, Leslie, et al.
Published: (2025)
by: Gu, Leslie, et al.
Published: (2025)
LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory
by: Zhang, Junyi, et al.
Published: (2026)
by: Zhang, Junyi, et al.
Published: (2026)
WonderJourney: Going from Anywhere to Everywhere
by: Yu, Hong-Xing, et al.
Published: (2023)
by: Yu, Hong-Xing, et al.
Published: (2023)
MotionV2V: Editing Motion in a Video
by: Burgert, Ryan, et al.
Published: (2025)
by: Burgert, Ryan, et al.
Published: (2025)
Boundary Attention: Learning curves, corners, junctions and grouping
by: Polansky, Mia Gaia, et al.
Published: (2024)
by: Polansky, Mia Gaia, et al.
Published: (2024)
UFO-4D: Unposed Feedforward 4D Reconstruction from Two Images
by: Hur, Junhwa, et al.
Published: (2026)
by: Hur, Junhwa, et al.
Published: (2026)
High-Resolution Frame Interpolation with Patch-based Cascaded Diffusion
by: Hur, Junhwa, et al.
Published: (2024)
by: Hur, Junhwa, et al.
Published: (2024)
Telling Left from Right: Identifying Geometry-Aware Semantic Correspondence
by: Zhang, Junyi, et al.
Published: (2023)
by: Zhang, Junyi, et al.
Published: (2023)
Forecasting Motion in the Wild
by: Thakkar, Neerja, et al.
Published: (2026)
by: Thakkar, Neerja, et al.
Published: (2026)
Repulsive Trajectory Modification and Conflict Resolution for Efficient Multi-Manipulator Motion Planning
by: Hong, Junhwa, et al.
Published: (2025)
by: Hong, Junhwa, et al.
Published: (2025)
Lumiere: A Space-Time Diffusion Model for Video Generation
by: Bar-Tal, Omer, et al.
Published: (2024)
by: Bar-Tal, Omer, et al.
Published: (2024)
Force Prompting: Video Generation Models Can Learn and Generalize Physics-based Control Signals
by: Gillman, Nate, et al.
Published: (2025)
by: Gillman, Nate, et al.
Published: (2025)
Motion Guidance: Diffusion-Based Image Editing with Differentiable Motion Estimators
by: Geng, Daniel, et al.
Published: (2024)
by: Geng, Daniel, et al.
Published: (2024)
LiteFrame: Efficient Vision Encoders Unlock Frame Scaling in Video LLMs
by: Kim, Jihwan, et al.
Published: (2026)
by: Kim, Jihwan, et al.
Published: (2026)
Direct Motion Models for Assessing Generated Videos
by: Allen, Kelsey, et al.
Published: (2025)
by: Allen, Kelsey, et al.
Published: (2025)
Learning from One Continuous Video Stream
by: Carreira, João, et al.
Published: (2023)
by: Carreira, João, et al.
Published: (2023)
Motion meets Attention: Video Motion Prompts
by: Chen, Qixiang, et al.
Published: (2024)
by: Chen, Qixiang, et al.
Published: (2024)
SMooDi: Stylized Motion Diffusion Model
by: Zhong, Lei, et al.
Published: (2024)
by: Zhong, Lei, et al.
Published: (2024)
OVR: A Dataset for Open Vocabulary Temporal Repetition Counting in Videos
by: Dwibedi, Debidatta, et al.
Published: (2024)
by: Dwibedi, Debidatta, et al.
Published: (2024)
MegaSaM: Accurate, Fast, and Robust Structure and Motion from Casual Dynamic Videos
by: Li, Zhengqi, et al.
Published: (2024)
by: Li, Zhengqi, et al.
Published: (2024)
DreamWalk: Style Space Exploration using Diffusion Guidance
by: Shu, Michelle, et al.
Published: (2024)
by: Shu, Michelle, et al.
Published: (2024)
MASIV: Toward Material-Agnostic System Identification from Videos
by: Zhao, Yizhou, et al.
Published: (2025)
by: Zhao, Yizhou, et al.
Published: (2025)
VidPanos: Generative Panoramic Videos from Casual Panning Videos
by: Ma, Jingwei, et al.
Published: (2024)
by: Ma, Jingwei, et al.
Published: (2024)
Motion-o: Trajectory-Grounded Video Reasoning
by: Galoaa, Bishoy, et al.
Published: (2026)
by: Galoaa, Bishoy, et al.
Published: (2026)
OmniControl: Control Any Joint at Any Time for Human Motion Generation
by: Xie, Yiming, et al.
Published: (2023)
by: Xie, Yiming, et al.
Published: (2023)
A Short Note on Evaluating RepNet for Temporal Repetition Counting in Videos
by: Dwibedi, Debidatta, et al.
Published: (2024)
by: Dwibedi, Debidatta, et al.
Published: (2024)
PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion
by: Choi, Jaehyun, et al.
Published: (2025)
by: Choi, Jaehyun, et al.
Published: (2025)
Trajectory Attention for Fine-grained Video Motion Control
by: Xiao, Zeqi, et al.
Published: (2024)
by: Xiao, Zeqi, et al.
Published: (2024)
Mojito: Motion Trajectory and Intensity Control for Video Generation
by: He, Xuehai, et al.
Published: (2024)
by: He, Xuehai, et al.
Published: (2024)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
by: Zhang, Wanyue, et al.
Published: (2025)
by: Zhang, Wanyue, et al.
Published: (2025)
MotionFlow:Learning Implicit Motion Flow for Complex Camera Trajectory Control in Video Generation
by: Lei, Guojun, et al.
Published: (2025)
by: Lei, Guojun, et al.
Published: (2025)
Point Prompting: Counterfactual Tracking with Video Diffusion Models
by: Shrivastava, Ayush, et al.
Published: (2025)
by: Shrivastava, Ayush, et al.
Published: (2025)
MotionShot: Adaptive Motion Transfer across Arbitrary Objects for Text-to-Video Generation
by: Liu, Yanchen, et al.
Published: (2025)
by: Liu, Yanchen, et al.
Published: (2025)
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
by: Liang, Jingyun, et al.
Published: (2025)
by: Liang, Jingyun, et al.
Published: (2025)
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding
by: Kim, Namho, et al.
Published: (2025)
by: Kim, Namho, et al.
Published: (2025)
Emergent Temporal Correspondences from Video Diffusion Transformers
by: Nam, Jisu, et al.
Published: (2025)
by: Nam, Jisu, et al.
Published: (2025)
TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking
by: Nam, Jisu, et al.
Published: (2026)
by: Nam, Jisu, et al.
Published: (2026)
Topology-Agnostic Animal Motion Generation from Text Prompt
by: Chen, Keyi, et al.
Published: (2025)
by: Chen, Keyi, et al.
Published: (2025)
ScaleFlow++: Robust and Accurate Estimation of 3D Motion from Video
by: Ling, Han, et al.
Published: (2024)
by: Ling, Han, et al.
Published: (2024)
Similar Items
-
MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion
by: Zhang, Junyi, et al.
Published: (2024) -
GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure
by: Gu, Leslie, et al.
Published: (2025) -
LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory
by: Zhang, Junyi, et al.
Published: (2026) -
WonderJourney: Going from Anywhere to Everywhere
by: Yu, Hong-Xing, et al.
Published: (2023) -
MotionV2V: Editing Motion in a Video
by: Burgert, Ryan, et al.
Published: (2025)