MotionCanvas: Cinematic Shot Design with Controllable Image-to-Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Xing, Jinbo, Mai, Long, Ham, Cusuh, Huang, Jiahui, Mahapatra, Aniruddha, Fu, Chi-Wing, Wong, Tien-Tsin, Liu, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DreamLoop: Controllable Cinemagraph Generation from a Single Photograph
by: Mahapatra, Aniruddha, et al.
Published: (2026)
by: Mahapatra, Aniruddha, et al.
Published: (2026)
TRACE: Object Motion Editing in Videos with First-Frame Trajectory Guidance
by: Phung, Quynh, et al.
Published: (2026)
by: Phung, Quynh, et al.
Published: (2026)
TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets
by: Liu, Zhixuan, et al.
Published: (2026)
by: Liu, Zhixuan, et al.
Published: (2026)
CineVerse: Consistent Keyframe Synthesis for Cinematic Scene Composition
by: Phung, Quynh, et al.
Published: (2025)
by: Phung, Quynh, et al.
Published: (2025)
Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers
by: Ma, Xin, et al.
Published: (2025)
by: Ma, Xin, et al.
Published: (2025)
Physics-Grounded Motion Forecasting via Equation Discovery for Trajectory-Guided Image-to-Video Generation
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
Learning to Control Physically-simulated 3D Characters via Generating and Mimicking 2D Motions
by: Li, Jianan, et al.
Published: (2025)
by: Li, Jianan, et al.
Published: (2025)
Progressive Growing of Video Tokenizers for Temporally Compact Latent Spaces
by: Mahapatra, Aniruddha, et al.
Published: (2025)
by: Mahapatra, Aniruddha, et al.
Published: (2025)
Physics-based Scene Layout Generation from Human Motion
by: Li, Jianan, et al.
Published: (2024)
by: Li, Jianan, et al.
Published: (2024)
Synchronized Multi‐Frame Diffusion for Temporally Consistent Video Stylization
by: Minshan Xie, et al.
Published: (2025)
by: Minshan Xie, et al.
Published: (2025)
ToonCrafter: Generative Cartoon Interpolation
by: Xing, Jinbo, et al.
Published: (2024)
by: Xing, Jinbo, et al.
Published: (2024)
REGEN: Learning Compact Video Embedding with (Re-)Generative Decoder
by: Zhang, Yitian, et al.
Published: (2025)
by: Zhang, Yitian, et al.
Published: (2025)
FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models
by: Lim, Youngsun, et al.
Published: (2026)
by: Lim, Youngsun, et al.
Published: (2026)
ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis
by: Yu, Wangbo, et al.
Published: (2024)
by: Yu, Wangbo, et al.
Published: (2024)
Taming Reversible Halftoning via Predictive Luminance
by: Lau, Cheuk-Kit, et al.
Published: (2023)
by: Lau, Cheuk-Kit, et al.
Published: (2023)
MITO: A Millimeter-Wave Dataset and Simulator for Non-Line-of-Sight Perception
by: Dodds, Laura, et al.
Published: (2025)
by: Dodds, Laura, et al.
Published: (2025)
Training-free Stylized Text-to-Image Generation with Fast Inference
by: Ma, Xin, et al.
Published: (2025)
by: Ma, Xin, et al.
Published: (2025)
ShotVerse: Advancing Cinematic Camera Control for Text-Driven Multi-Shot Video Creation
by: Yang, Songlin, et al.
Published: (2026)
by: Yang, Songlin, et al.
Published: (2026)
V2M-Zero: Zero-Pair Time-Aligned Video-to-Music Generation
by: Lin, Yan-Bo, et al.
Published: (2026)
by: Lin, Yan-Bo, et al.
Published: (2026)
Text-Guided Texturing by Synchronized Multi-View Diffusion
by: Liu, Yuxin, et al.
Published: (2023)
by: Liu, Yuxin, et al.
Published: (2023)
Personalized Residuals for Concept-Driven Text-to-Image Generation
by: Ham, Cusuh, et al.
Published: (2024)
by: Ham, Cusuh, et al.
Published: (2024)
PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion
by: Gao, Heyuan, et al.
Published: (2026)
by: Gao, Heyuan, et al.
Published: (2026)
Cinematic Images
Published: (2024)
Published: (2024)
Sketch2Manga: Shaded Manga Screening from Sketch with Diffusion Models
by: Lin, Jian, et al.
Published: (2024)
by: Lin, Jian, et al.
Published: (2024)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
by: Meng, Yihao, et al.
Published: (2025)
by: Meng, Yihao, et al.
Published: (2025)
Geopolitics and the Planet: Does Geopolitical Risk Drive Carbon Emissions?
by: Meng Tsin; Chi Wei Su
Published: (2026)
by: Meng Tsin; Chi Wei Su
Published: (2026)
Geopolitics and the Planet: Does Geopolitical Risk Drive Carbon Emissions?
by: Meng Tsin; Chi Wei Su
Published: (2026)
by: Meng Tsin; Chi Wei Su
Published: (2026)
Zero-Shot Personalized Camera Motion Control for Image-to-Video Synthesis
by: Guhan, Pooja, et al.
Published: (2025)
by: Guhan, Pooja, et al.
Published: (2025)
CanvasMAR: Improving Masked Autoregressive Video Prediction With Canvas
by: Li, Zian, et al.
Published: (2025)
by: Li, Zian, et al.
Published: (2025)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
by: Dalva, Yusuf, et al.
Published: (2025)
by: Dalva, Yusuf, et al.
Published: (2025)
Unveiling Deep Shadows: A Survey and Benchmark on Image and Video Shadow Detection, Removal, and Generation in the Deep Learning Era
by: Hu, Xiaowei, et al.
Published: (2024)
by: Hu, Xiaowei, et al.
Published: (2024)
D2Diff : A Dual Domain Diffusion Model for Accurate Multi-Contrast MRI Synthesis
by: Dayarathna, Sanuwani, et al.
Published: (2025)
by: Dayarathna, Sanuwani, et al.
Published: (2025)
Generative Photographic Control for Scene-Consistent Video Cinematic Editing
by: Sun, Huiqiang, et al.
Published: (2025)
by: Sun, Huiqiang, et al.
Published: (2025)
CoT-RVS: Zero-Shot Chain-of-Thought Reasoning Segmentation for Videos
by: Kao, Shiu-hong, et al.
Published: (2025)
by: Kao, Shiu-hong, et al.
Published: (2025)
ID-Sim: An Identity-Focused Similarity Metric
by: Chae, Julia, et al.
Published: (2026)
by: Chae, Julia, et al.
Published: (2026)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
by: Wei, Yujie, et al.
Published: (2024)
by: Wei, Yujie, et al.
Published: (2024)
MotionPro: A Precise Motion Controller for Image-to-Video Generation
by: Zhang, Zhongwei, et al.
Published: (2025)
by: Zhang, Zhongwei, et al.
Published: (2025)
DynamicEval: Rethinking Evaluation for Dynamic Text-to-Video Synthesis
by: Babu, Nithin C., et al.
Published: (2025)
by: Babu, Nithin C., et al.
Published: (2025)
On the Content Bias in Fréchet Video Distance
by: Ge, Songwei, et al.
Published: (2024)
by: Ge, Songwei, et al.
Published: (2024)
I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
Similar Items
-
DreamLoop: Controllable Cinemagraph Generation from a Single Photograph
by: Mahapatra, Aniruddha, et al.
Published: (2026) -
TRACE: Object Motion Editing in Videos with First-Frame Trajectory Guidance
by: Phung, Quynh, et al.
Published: (2026) -
TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets
by: Liu, Zhixuan, et al.
Published: (2026) -
CineVerse: Consistent Keyframe Synthesis for Cinematic Scene Composition
by: Phung, Quynh, et al.
Published: (2025) -
Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers
by: Ma, Xin, et al.
Published: (2025)