Versatile Editing of Video Content, Actions, and Dynamics without Training
Fuente:
arXiv
Saved in:
| Main Authors: | Kulikov, Vladimir, Paiss, Roni, Voynov, Andrey, Mosseri, Inbar, Dekel, Tali, Michaeli, Tomer |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space
by: Garibi, Daniel, et al.
Published: (2025)
by: Garibi, Daniel, et al.
Published: (2025)
Still-Moving: Customized Video Generation without Customized Video Data
by: Chefer, Hila, et al.
Published: (2024)
by: Chefer, Hila, et al.
Published: (2024)
FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models
by: Kulikov, Vladimir, et al.
Published: (2024)
by: Kulikov, Vladimir, et al.
Published: (2024)
Slicedit: Zero-Shot Video Editing With Text-to-Image Diffusion Models Using Spatio-Temporal Slices
by: Cohen, Nathaniel, et al.
Published: (2024)
by: Cohen, Nathaniel, et al.
Published: (2024)
An Edit Friendly DDPM Noise Space: Inversion and Manipulations
by: Huberman-Spiegelglas, Inbar, et al.
Published: (2023)
by: Huberman-Spiegelglas, Inbar, et al.
Published: (2023)
FlowOpt: Fast Optimization Through Whole Flow Processes for Training-Free Editing
by: Ronai, Or, et al.
Published: (2025)
by: Ronai, Or, et al.
Published: (2025)
Lumiere: A Space-Time Diffusion Model for Video Generation
by: Bar-Tal, Omer, et al.
Published: (2024)
by: Bar-Tal, Omer, et al.
Published: (2024)
Eye2Eye: A Simple Approach for Monocular-to-Stereo Video Synthesis
by: Geyer, Michal, et al.
Published: (2025)
by: Geyer, Michal, et al.
Published: (2025)
DynVFX: Augmenting Real Videos with Dynamic Content
by: Yatim, Danah, et al.
Published: (2025)
by: Yatim, Danah, et al.
Published: (2025)
VidPanos: Generative Panoramic Videos from Casual Panning Videos
by: Ma, Jingwei, et al.
Published: (2024)
by: Ma, Jingwei, et al.
Published: (2024)
ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning
by: Zhang, David Junhao, et al.
Published: (2024)
by: Zhang, David Junhao, et al.
Published: (2024)
MineTheGap: Automatic Mining of Biases in Text-to-Image Models
by: Cohen, Noa, et al.
Published: (2025)
by: Cohen, Noa, et al.
Published: (2025)
What Are You Doing? A Closer Look at Controllable Human Video Generation
by: Bugliarello, Emanuele, et al.
Published: (2025)
by: Bugliarello, Emanuele, et al.
Published: (2025)
DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video
by: Tumanyan, Narek, et al.
Published: (2024)
by: Tumanyan, Narek, et al.
Published: (2024)
Imitating the Functionality of Image-to-Image Models Using a Single Example
by: Spingarn-Eliezer, Nurit, et al.
Published: (2024)
by: Spingarn-Eliezer, Nurit, et al.
Published: (2024)
OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation
by: Liang, Sen, et al.
Published: (2025)
by: Liang, Sen, et al.
Published: (2025)
Match-and-Fuse: Consistent Generation from Unstructured Image Sets
by: Feingold, Kate, et al.
Published: (2025)
by: Feingold, Kate, et al.
Published: (2025)
Illumination Angular Spectrum Encoding for Controlling the Functionality of Diffractive Networks
by: Kleiner, Matan, et al.
Published: (2026)
by: Kleiner, Matan, et al.
Published: (2026)
Alias-Free Convnets: Fractional Shift Invariance via Polynomial Activations
by: Michaeli, Hagay, et al.
Published: (2023)
by: Michaeli, Hagay, et al.
Published: (2023)
Coherence Awareness in Diffractive Neural Networks
by: Kleiner, Matan, et al.
Published: (2024)
by: Kleiner, Matan, et al.
Published: (2024)
On the Posterior Distribution in Denoising: Application to Uncertainty Quantification
by: Manor, Hila, et al.
Published: (2023)
by: Manor, Hila, et al.
Published: (2023)
Uncertainty Visualization via Low-Dimensional Posterior Projections
by: Yair, Omer, et al.
Published: (2023)
by: Yair, Omer, et al.
Published: (2023)
Discovering Interpretable Directions in the Semantic Latent Space of Diffusion Models
by: Haas, René, et al.
Published: (2023)
by: Haas, René, et al.
Published: (2023)
What's in the Image? A Deep-Dive into the Vision of Vision Language Models
by: Kaduri, Omri, et al.
Published: (2024)
by: Kaduri, Omri, et al.
Published: (2024)
MotionV2V: Editing Motion in a Video
by: Burgert, Ryan, et al.
Published: (2025)
by: Burgert, Ryan, et al.
Published: (2025)
DreamColour: Controllable Video Colour Editing without Training
by: Utintu, Chaitat, et al.
Published: (2024)
by: Utintu, Chaitat, et al.
Published: (2024)
Action Reimagined: Text-to-Pose Video Editing for Dynamic Human Actions
by: Wang, Lan, et al.
Published: (2024)
by: Wang, Lan, et al.
Published: (2024)
Dynamic Motion Blending for Versatile Motion Editing
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
Generative Omnimatte: Learning to Decompose Video into Layers
by: Lee, Yao-Chih, et al.
Published: (2024)
by: Lee, Yao-Chih, et al.
Published: (2024)
GSURE-Based Diffusion Model Training with Corrupted Data
by: Kawar, Bahjat, et al.
Published: (2023)
by: Kawar, Bahjat, et al.
Published: (2023)
PSC: Posterior Sampling-Based Compression
by: Elata, Noam, et al.
Published: (2024)
by: Elata, Noam, et al.
Published: (2024)
Leveraging Near-Field Lighting for Monocular Depth Estimation from Endoscopy Videos
by: Paruchuri, Akshay, et al.
Published: (2024)
by: Paruchuri, Akshay, et al.
Published: (2024)
SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder
by: Kamenetsky, Ronen, et al.
Published: (2025)
by: Kamenetsky, Ronen, et al.
Published: (2025)
Style Aligned Image Generation via Shared Attention
by: Hertz, Amir, et al.
Published: (2023)
by: Hertz, Amir, et al.
Published: (2023)
DRoPS: Dynamic 3D Reconstruction of Pre-Scanned Objects
by: Tumanyan, Narek, et al.
Published: (2026)
by: Tumanyan, Narek, et al.
Published: (2026)
InvFusion: Bridging Supervised and Zero-shot Diffusion for Inverse Problems
by: Elata, Noam, et al.
Published: (2025)
by: Elata, Noam, et al.
Published: (2025)
Adaptive Compressed Sensing with Diffusion-Based Posterior Sampling
by: Elata, Noam, et al.
Published: (2024)
by: Elata, Noam, et al.
Published: (2024)
Hierarchical Uncertainty Exploration via Feedforward Posterior Trees
by: Nehme, Elias, et al.
Published: (2024)
by: Nehme, Elias, et al.
Published: (2024)
Vamos: Versatile Action Models for Video Understanding
by: Wang, Shijie, et al.
Published: (2023)
by: Wang, Shijie, et al.
Published: (2023)
TalkPhoto: A Versatile Training-Free Conversational Assistant for Intelligent Image Editing
by: Hu, Yujie, et al.
Published: (2026)
by: Hu, Yujie, et al.
Published: (2026)
Similar Items
-
TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space
by: Garibi, Daniel, et al.
Published: (2025) -
Still-Moving: Customized Video Generation without Customized Video Data
by: Chefer, Hila, et al.
Published: (2024) -
FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models
by: Kulikov, Vladimir, et al.
Published: (2024) -
Slicedit: Zero-Shot Video Editing With Text-to-Image Diffusion Models Using Spatio-Temporal Slices
by: Cohen, Nathaniel, et al.
Published: (2024) -
An Edit Friendly DDPM Noise Space: Inversion and Manipulations
by: Huberman-Spiegelglas, Inbar, et al.
Published: (2023)