Motion-I2V: Consistent and Controllable Image-to-Video Generation with Explicit Motion Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Xiaoyu, Huang, Zhaoyang, Wang, Fu-Yun, Bian, Weikang, Li, Dasong, Zhang, Yi, Zhang, Manyuan, Cheung, Ka Chun, See, Simon, Qin, Hongwei, Dai, Jifeng, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking
von: Bian, Weikang, et al.
Veröffentlicht: (2025)
von: Bian, Weikang, et al.
Veröffentlicht: (2025)
AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
M3Net: Multimodal Multi-task Learning for 3D Detection, Segmentation, and Occupancy Prediction in Autonomous Driving
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
RelightMaster: Precise Video Relighting with Multi-plane Light Images
von: Bian, Weikang, et al.
Veröffentlicht: (2025)
von: Bian, Weikang, et al.
Veröffentlicht: (2025)
GeometrySticker: Enabling Ownership Claim of Recolorized Neural Radiance Fields
von: Huang, Xiufeng, et al.
Veröffentlicht: (2024)
von: Huang, Xiufeng, et al.
Veröffentlicht: (2024)
Phased Consistency Models
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
Explicit Motion Handling and Interactive Prompting for Video Camouflaged Object Detection
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
Stable Consistency Tuning: Understanding and Improving Consistency Models
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
GaussianMarker: Uncertainty-Aware Copyright Protection of 3D Gaussian Splatting
von: Huang, Xiufeng, et al.
Veröffentlicht: (2024)
von: Huang, Xiufeng, et al.
Veröffentlicht: (2024)
Resilient Practical Test-Time Adaptation: Soft Batch Normalization Alignment and Entropy-driven Memory Bank
von: Zhou, Xingzhi, et al.
Veröffentlicht: (2024)
von: Zhou, Xingzhi, et al.
Veröffentlicht: (2024)
Delving Deep into Engagement Prediction of Short Videos
von: Li, Dasong, et al.
Veröffentlicht: (2024)
von: Li, Dasong, et al.
Veröffentlicht: (2024)
Training-Free Motion-Guided Video Generation with Enhanced Temporal Consistency Using Motion Consistency Loss
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
BlinkVision: A Benchmark for Optical Flow, Scene Flow and Point Tracking Estimation using RGB Frames and Events
von: Li, Yijin, et al.
Veröffentlicht: (2024)
von: Li, Yijin, et al.
Veröffentlicht: (2024)
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls
von: Bian, Yuxuan, et al.
Veröffentlicht: (2024)
von: Bian, Yuxuan, et al.
Veröffentlicht: (2024)
ReVideo: Remake a Video with Motion and Content Control
von: Mou, Chong, et al.
Veröffentlicht: (2024)
von: Mou, Chong, et al.
Veröffentlicht: (2024)
ImageSentinel: Protecting Visual Datasets from Unauthorized Retrieval-Augmented Image Generation
von: Luo, Ziyuan, et al.
Veröffentlicht: (2025)
von: Luo, Ziyuan, et al.
Veröffentlicht: (2025)
Protecting NeRFs' Copyright via Plug-And-Play Watermarking Base Model
von: Song, Qi, et al.
Veröffentlicht: (2024)
von: Song, Qi, et al.
Veröffentlicht: (2024)
Align 3D Representation and Text Embedding for 3D Content Personalization
von: Song, Qi, et al.
Veröffentlicht: (2025)
von: Song, Qi, et al.
Veröffentlicht: (2025)
Stereo-GS: Multi-View Stereo Vision Model for Generalizable 3D Gaussian Splatting Reconstruction
von: Huang, Xiufeng, et al.
Veröffentlicht: (2025)
von: Huang, Xiufeng, et al.
Veröffentlicht: (2025)
Geometry Cloak: Preventing TGS-based 3D Reconstruction from Copyrighted Images
von: Song, Qi, et al.
Veröffentlicht: (2024)
von: Song, Qi, et al.
Veröffentlicht: (2024)
Proteus-ID: ID-Consistent and Motion-Coherent Video Customization
von: Zhang, Guiyu, et al.
Veröffentlicht: (2025)
von: Zhang, Guiyu, et al.
Veröffentlicht: (2025)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
Fréchet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos
von: Liu, Jiahe, et al.
Veröffentlicht: (2024)
von: Liu, Jiahe, et al.
Veröffentlicht: (2024)
Ponymation: Learning Articulated 3D Animal Motions from Unlabeled Online Videos
von: Sun, Keqiang, et al.
Veröffentlicht: (2023)
von: Sun, Keqiang, et al.
Veröffentlicht: (2023)
Be-Your-Outpainter: Mastering Video Outpainting through Input-Specific Adaptation
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
Don't Judge by the Look: Towards Motion Coherent Video Representation
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
MotionV2V: Editing Motion in a Video
von: Burgert, Ryan, et al.
Veröffentlicht: (2025)
von: Burgert, Ryan, et al.
Veröffentlicht: (2025)
A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
von: Dong, Yitong, et al.
Veröffentlicht: (2024)
von: Dong, Yitong, et al.
Veröffentlicht: (2024)
Learning Explicit Continuous Motion Representation for Dynamic Gaussian Splatting from Monocular Videos
von: Zhang, Xuankai, et al.
Veröffentlicht: (2026)
von: Zhang, Xuankai, et al.
Veröffentlicht: (2026)
Deep Reward Supervisions for Tuning Text-to-Image Diffusion Models
von: Wu, Xiaoshi, et al.
Veröffentlicht: (2024)
von: Wu, Xiaoshi, et al.
Veröffentlicht: (2024)
Biological Pathway Informed Models with Graph Attention Networks (GATs)
von: Wong, Gavin, et al.
Veröffentlicht: (2025)
von: Wong, Gavin, et al.
Veröffentlicht: (2025)
TransVDM: Motion-Constrained Video Diffusion Model for Transparent Video Synthesis
von: Li, Menghao, et al.
Veröffentlicht: (2025)
von: Li, Menghao, et al.
Veröffentlicht: (2025)
MoCA-Video: Motion-Aware Concept Alignment for Consistent Video Editing
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model
von: Dai, Wenxun, et al.
Veröffentlicht: (2024)
von: Dai, Wenxun, et al.
Veröffentlicht: (2024)
MotionGS: Exploring Explicit Motion Guidance for Deformable 3D Gaussian Splatting
von: Zhu, Ruijie, et al.
Veröffentlicht: (2024)
von: Zhu, Ruijie, et al.
Veröffentlicht: (2024)
Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency
von: Sun, Shangkun, et al.
Veröffentlicht: (2025)
von: Sun, Shangkun, et al.
Veröffentlicht: (2025)
Denoising Reuse: Exploiting Inter-frame Motion Consistency for Efficient Video Latent Generation
von: Wang, Chenyu, et al.
Veröffentlicht: (2024)
von: Wang, Chenyu, et al.
Veröffentlicht: (2024)
Human Motion Video Generation: A Survey
von: Xue, Haiwei, et al.
Veröffentlicht: (2025)
von: Xue, Haiwei, et al.
Veröffentlicht: (2025)
Motion-Aware Video Frame Interpolation
von: Han, Pengfei, et al.
Veröffentlicht: (2024)
von: Han, Pengfei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking
von: Bian, Weikang, et al.
Veröffentlicht: (2025) -
AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024) -
M3Net: Multimodal Multi-task Learning for 3D Detection, Segmentation, and Occupancy Prediction in Autonomous Driving
von: Chen, Xuesong, et al.
Veröffentlicht: (2025) -
RelightMaster: Precise Video Relighting with Multi-plane Light Images
von: Bian, Weikang, et al.
Veröffentlicht: (2025) -
GeometrySticker: Enabling Ownership Claim of Recolorized Neural Radiance Fields
von: Huang, Xiufeng, et al.
Veröffentlicht: (2024)