ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, David Junhao, Paiss, Roni, Zada, Shiran, Karnad, Nikhil, Jacobs, David E., Pritch, Yael, Mosseri, Inbar, Shou, Mike Zheng, Wadhwa, Neal, Ruiz, Nataniel |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Still-Moving: Customized Video Generation without Customized Video Data
by: Chefer, Hila, et al.
Published: (2024)
by: Chefer, Hila, et al.
Published: (2024)
Versatile Editing of Video Content, Actions, and Dynamics without Training
by: Kulikov, Vladimir, et al.
Published: (2026)
by: Kulikov, Vladimir, et al.
Published: (2026)
TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space
by: Garibi, Daniel, et al.
Published: (2025)
by: Garibi, Daniel, et al.
Published: (2025)
VidPanos: Generative Panoramic Videos from Casual Panning Videos
by: Ma, Jingwei, et al.
Published: (2024)
by: Ma, Jingwei, et al.
Published: (2024)
Magic Insert: Style-Aware Drag-and-Drop
by: Ruiz, Nataniel, et al.
Published: (2024)
by: Ruiz, Nataniel, et al.
Published: (2024)
Unbounded: A Generative Infinite Game of Character Life Simulation
by: Li, Jialu, et al.
Published: (2024)
by: Li, Jialu, et al.
Published: (2024)
Lumiere: A Space-Time Diffusion Model for Video Generation
by: Bar-Tal, Omer, et al.
Published: (2024)
by: Bar-Tal, Omer, et al.
Published: (2024)
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
by: Hong, Yeobin, et al.
Published: (2025)
by: Hong, Yeobin, et al.
Published: (2025)
RealFill: Reference-Driven Generation for Authentic Image Completion
by: Tang, Luming, et al.
Published: (2023)
by: Tang, Luming, et al.
Published: (2023)
MotionV2V: Editing Motion in a Video
by: Burgert, Ryan, et al.
Published: (2025)
by: Burgert, Ryan, et al.
Published: (2025)
HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models
by: Ruiz, Nataniel, et al.
Published: (2023)
by: Ruiz, Nataniel, et al.
Published: (2023)
What Are You Doing? A Closer Look at Controllable Human Video Generation
by: Bugliarello, Emanuele, et al.
Published: (2025)
by: Bugliarello, Emanuele, et al.
Published: (2025)
MultiGen: Level-Design for Editable Multiplayer Worlds in Diffusion Game Engines
by: Po, Ryan, et al.
Published: (2026)
by: Po, Ryan, et al.
Published: (2026)
Eye2Eye: A Simple Approach for Monocular-to-Stereo Video Synthesis
by: Geyer, Michal, et al.
Published: (2025)
by: Geyer, Michal, et al.
Published: (2025)
Impossible Videos
by: Bai, Zechen, et al.
Published: (2025)
by: Bai, Zechen, et al.
Published: (2025)
Moonshot: Towards Controllable Video Generation and Editing with Multimodal Conditions
by: Zhang, David Junhao, et al.
Published: (2024)
by: Zhang, David Junhao, et al.
Published: (2024)
TPDiff: Temporal Pyramid Video Diffusion Model
by: Ran, Lingmin, et al.
Published: (2025)
by: Ran, Lingmin, et al.
Published: (2025)
Paper2Video: Automatic Video Generation from Scientific Papers
by: Zhu, Zeyu, et al.
Published: (2025)
by: Zhu, Zeyu, et al.
Published: (2025)
Code2Video: A Code-centric Paradigm for Educational Video Generation
by: Chen, Yanzhe, et al.
Published: (2025)
by: Chen, Yanzhe, et al.
Published: (2025)
PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion
by: Gao, Heyuan, et al.
Published: (2026)
by: Gao, Heyuan, et al.
Published: (2026)
X-Humanoid: Robotize Human Videos to Generate Humanoid Videos at Scale
by: Yang, Pei, et al.
Published: (2025)
by: Yang, Pei, et al.
Published: (2025)
Recurrent Video Masked Autoencoders
by: Zoran, Daniel, et al.
Published: (2025)
by: Zoran, Daniel, et al.
Published: (2025)
VLog: Video-Language Models by Generative Retrieval of Narration Vocabulary
by: Lin, Kevin Qinghong, et al.
Published: (2025)
by: Lin, Kevin Qinghong, et al.
Published: (2025)
LiveSVG: Zero-Shot SVG Animation via Video Generation
by: Levy, Matan, et al.
Published: (2026)
by: Levy, Matan, et al.
Published: (2026)
GLOW: Global Illumination-Aware Inverse Rendering of Indoor Scenes Captured with Dynamic Co-Located Light & Camera
by: Wu, Jiaye, et al.
Published: (2025)
by: Wu, Jiaye, et al.
Published: (2025)
VideoMind: A Chain-of-LoRA Agent for Temporal-Grounded Video Reasoning
by: Liu, Ye, et al.
Published: (2025)
by: Liu, Ye, et al.
Published: (2025)
Long-Context Autoregressive Video Modeling with Next-Frame Prediction
by: Gu, Yuchao, et al.
Published: (2025)
by: Gu, Yuchao, et al.
Published: (2025)
Leveraging Near-Field Lighting for Monocular Depth Estimation from Endoscopy Videos
by: Paruchuri, Akshay, et al.
Published: (2024)
by: Paruchuri, Akshay, et al.
Published: (2024)
Capturing Stable HDR Videos Using a Dual-Camera System
by: Zhang, Qianyu, et al.
Published: (2025)
by: Zhang, Qianyu, et al.
Published: (2025)
Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation
by: Zhang, David Junhao, et al.
Published: (2023)
by: Zhang, David Junhao, et al.
Published: (2023)
AssistPDA: An Online Video Surveillance Assistant for Video Anomaly Prediction, Detection, and Analysis
by: Yang, Zhiwei, et al.
Published: (2025)
by: Yang, Zhiwei, et al.
Published: (2025)
Mitty: Diffusion-based Human-to-Robot Video Generation
by: Song, Yiren, et al.
Published: (2025)
by: Song, Yiren, et al.
Published: (2025)
PANDA: Towards Generalist Video Anomaly Detection via Agentic AI Engineer
by: Yang, Zhiwei, et al.
Published: (2025)
by: Yang, Zhiwei, et al.
Published: (2025)
Text-Guided Video Masked Autoencoder
by: Fan, David, et al.
Published: (2024)
by: Fan, David, et al.
Published: (2024)
VideoLLM-online: Online Video Large Language Model for Streaming Video
by: Chen, Joya, et al.
Published: (2024)
by: Chen, Joya, et al.
Published: (2024)
VISTA: Triplet-Supervised Video Style Transfer with Diffusion Transformers
by: Song, Yiren, et al.
Published: (2026)
by: Song, Yiren, et al.
Published: (2026)
Factorized Learning for Temporally Grounded Video-Language Models
by: Zeng, Wenzheng, et al.
Published: (2025)
by: Zeng, Wenzheng, et al.
Published: (2025)
VideoGUI: A Benchmark for GUI Automation from Instructional Videos
by: Lin, Kevin Qinghong, et al.
Published: (2024)
by: Lin, Kevin Qinghong, et al.
Published: (2024)
H2R-Grounder: A Paired-Data-Free Paradigm for Translating Human Interaction Videos into Physically Grounded Robot Videos
by: Ci, Hai, et al.
Published: (2025)
by: Ci, Hai, et al.
Published: (2025)
RobotSeg: A Model and Dataset for Segmenting Robots in Image and Video
by: Mei, Haiyang, et al.
Published: (2025)
by: Mei, Haiyang, et al.
Published: (2025)
Similar Items
-
Still-Moving: Customized Video Generation without Customized Video Data
by: Chefer, Hila, et al.
Published: (2024) -
Versatile Editing of Video Content, Actions, and Dynamics without Training
by: Kulikov, Vladimir, et al.
Published: (2026) -
TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space
by: Garibi, Daniel, et al.
Published: (2025) -
VidPanos: Generative Panoramic Videos from Casual Panning Videos
by: Ma, Jingwei, et al.
Published: (2024) -
Magic Insert: Style-Aware Drag-and-Drop
by: Ruiz, Nataniel, et al.
Published: (2024)