Motion-I2V: Consistent and Controllable Image-to-Video Generation with Explicit Motion Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Xiaoyu, Huang, Zhaoyang, Wang, Fu-Yun, Bian, Weikang, Li, Dasong, Zhang, Yi, Zhang, Manyuan, Cheung, Ka Chun, See, Simon, Qin, Hongwei, Dai, Jifeng, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking
by: Bian, Weikang, et al.
Published: (2025)
by: Bian, Weikang, et al.
Published: (2025)
AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
M3Net: Multimodal Multi-task Learning for 3D Detection, Segmentation, and Occupancy Prediction in Autonomous Driving
by: Chen, Xuesong, et al.
Published: (2025)
by: Chen, Xuesong, et al.
Published: (2025)
RelightMaster: Precise Video Relighting with Multi-plane Light Images
by: Bian, Weikang, et al.
Published: (2025)
by: Bian, Weikang, et al.
Published: (2025)
GeometrySticker: Enabling Ownership Claim of Recolorized Neural Radiance Fields
by: Huang, Xiufeng, et al.
Published: (2024)
by: Huang, Xiufeng, et al.
Published: (2024)
Phased Consistency Models
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
Explicit Motion Handling and Interactive Prompting for Video Camouflaged Object Detection
by: Zhang, Xin, et al.
Published: (2024)
by: Zhang, Xin, et al.
Published: (2024)
Stable Consistency Tuning: Understanding and Improving Consistency Models
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
GaussianMarker: Uncertainty-Aware Copyright Protection of 3D Gaussian Splatting
by: Huang, Xiufeng, et al.
Published: (2024)
by: Huang, Xiufeng, et al.
Published: (2024)
Resilient Practical Test-Time Adaptation: Soft Batch Normalization Alignment and Entropy-driven Memory Bank
by: Zhou, Xingzhi, et al.
Published: (2024)
by: Zhou, Xingzhi, et al.
Published: (2024)
Delving Deep into Engagement Prediction of Short Videos
by: Li, Dasong, et al.
Published: (2024)
by: Li, Dasong, et al.
Published: (2024)
Training-Free Motion-Guided Video Generation with Enhanced Temporal Consistency Using Motion Consistency Loss
by: Zhang, Xinyu, et al.
Published: (2025)
by: Zhang, Xinyu, et al.
Published: (2025)
BlinkVision: A Benchmark for Optical Flow, Scene Flow and Point Tracking Estimation using RGB Frames and Events
by: Li, Yijin, et al.
Published: (2024)
by: Li, Yijin, et al.
Published: (2024)
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls
by: Bian, Yuxuan, et al.
Published: (2024)
by: Bian, Yuxuan, et al.
Published: (2024)
ReVideo: Remake a Video with Motion and Content Control
by: Mou, Chong, et al.
Published: (2024)
by: Mou, Chong, et al.
Published: (2024)
ImageSentinel: Protecting Visual Datasets from Unauthorized Retrieval-Augmented Image Generation
by: Luo, Ziyuan, et al.
Published: (2025)
by: Luo, Ziyuan, et al.
Published: (2025)
Protecting NeRFs' Copyright via Plug-And-Play Watermarking Base Model
by: Song, Qi, et al.
Published: (2024)
by: Song, Qi, et al.
Published: (2024)
Align 3D Representation and Text Embedding for 3D Content Personalization
by: Song, Qi, et al.
Published: (2025)
by: Song, Qi, et al.
Published: (2025)
Stereo-GS: Multi-View Stereo Vision Model for Generalizable 3D Gaussian Splatting Reconstruction
by: Huang, Xiufeng, et al.
Published: (2025)
by: Huang, Xiufeng, et al.
Published: (2025)
Geometry Cloak: Preventing TGS-based 3D Reconstruction from Copyrighted Images
by: Song, Qi, et al.
Published: (2024)
by: Song, Qi, et al.
Published: (2024)
Proteus-ID: ID-Consistent and Motion-Coherent Video Customization
by: Zhang, Guiyu, et al.
Published: (2025)
by: Zhang, Guiyu, et al.
Published: (2025)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
by: Zhai, Yuanhao, et al.
Published: (2024)
by: Zhai, Yuanhao, et al.
Published: (2024)
Fréchet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos
by: Liu, Jiahe, et al.
Published: (2024)
by: Liu, Jiahe, et al.
Published: (2024)
Ponymation: Learning Articulated 3D Animal Motions from Unlabeled Online Videos
by: Sun, Keqiang, et al.
Published: (2023)
by: Sun, Keqiang, et al.
Published: (2023)
Be-Your-Outpainter: Mastering Video Outpainting through Input-Specific Adaptation
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
Don't Judge by the Look: Towards Motion Coherent Video Representation
by: Zhang, Yitian, et al.
Published: (2024)
by: Zhang, Yitian, et al.
Published: (2024)
MotionV2V: Editing Motion in a Video
by: Burgert, Ryan, et al.
Published: (2025)
by: Burgert, Ryan, et al.
Published: (2025)
A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
by: Dong, Yitong, et al.
Published: (2024)
by: Dong, Yitong, et al.
Published: (2024)
Learning Explicit Continuous Motion Representation for Dynamic Gaussian Splatting from Monocular Videos
by: Zhang, Xuankai, et al.
Published: (2026)
by: Zhang, Xuankai, et al.
Published: (2026)
Deep Reward Supervisions for Tuning Text-to-Image Diffusion Models
by: Wu, Xiaoshi, et al.
Published: (2024)
by: Wu, Xiaoshi, et al.
Published: (2024)
Biological Pathway Informed Models with Graph Attention Networks (GATs)
by: Wong, Gavin, et al.
Published: (2025)
by: Wong, Gavin, et al.
Published: (2025)
TransVDM: Motion-Constrained Video Diffusion Model for Transparent Video Synthesis
by: Li, Menghao, et al.
Published: (2025)
by: Li, Menghao, et al.
Published: (2025)
MoCA-Video: Motion-Aware Concept Alignment for Consistent Video Editing
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
by: Wu, Jianzong, et al.
Published: (2024)
by: Wu, Jianzong, et al.
Published: (2024)
MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model
by: Dai, Wenxun, et al.
Published: (2024)
by: Dai, Wenxun, et al.
Published: (2024)
MotionGS: Exploring Explicit Motion Guidance for Deformable 3D Gaussian Splatting
by: Zhu, Ruijie, et al.
Published: (2024)
by: Zhu, Ruijie, et al.
Published: (2024)
Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency
by: Sun, Shangkun, et al.
Published: (2025)
by: Sun, Shangkun, et al.
Published: (2025)
Denoising Reuse: Exploiting Inter-frame Motion Consistency for Efficient Video Latent Generation
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
Human Motion Video Generation: A Survey
by: Xue, Haiwei, et al.
Published: (2025)
by: Xue, Haiwei, et al.
Published: (2025)
Motion-Aware Video Frame Interpolation
by: Han, Pengfei, et al.
Published: (2024)
by: Han, Pengfei, et al.
Published: (2024)
Similar Items
-
GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking
by: Bian, Weikang, et al.
Published: (2025) -
AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
by: Wang, Fu-Yun, et al.
Published: (2024) -
M3Net: Multimodal Multi-task Learning for 3D Detection, Segmentation, and Occupancy Prediction in Autonomous Driving
by: Chen, Xuesong, et al.
Published: (2025) -
RelightMaster: Precise Video Relighting with Multi-plane Light Images
by: Bian, Weikang, et al.
Published: (2025) -
GeometrySticker: Enabling Ownership Claim of Recolorized Neural Radiance Fields
by: Huang, Xiufeng, et al.
Published: (2024)