VideoSketcher: Video Models Prior Enable Versatile Sequential Sketch Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Hui, Alaluf, Yuval, Tal, Omer Bar, Schwing, Alexander, Torralba, Antonio, Vinker, Yael |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SketchAgent: Language-Driven Sequential Sketch Generation
by: Vinker, Yael, et al.
Published: (2024)
by: Vinker, Yael, et al.
Published: (2024)
NeuralSVG: An Implicit Representation for Text-to-Vector Generation
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
Generative Visual Communication in the Era of Vision-Language Models
by: Vinker, Yael
Published: (2024)
by: Vinker, Yael
Published: (2024)
SwiftSketch: A Diffusion Model for Image-to-Vector Sketch Generation
by: Arar, Ellie, et al.
Published: (2025)
by: Arar, Ellie, et al.
Published: (2025)
Instance Segmentation of Scene Sketches Using Natural Image Priors
by: Tang, Mia, et al.
Published: (2025)
by: Tang, Mia, et al.
Published: (2025)
Piece it Together: Part-Based Concepting with IP-Priors
by: Richardson, Elad, et al.
Published: (2025)
by: Richardson, Elad, et al.
Published: (2025)
DynVFX: Augmenting Real Videos with Dynamic Content
by: Yatim, Danah, et al.
Published: (2025)
by: Yatim, Danah, et al.
Published: (2025)
MultiModal Action Conditioned Video Generation
by: Li, Yichen, et al.
Published: (2025)
by: Li, Yichen, et al.
Published: (2025)
Teaching an Agent to Sketch One Part at a Time
by: Du, Xiaodan, et al.
Published: (2026)
by: Du, Xiaodan, et al.
Published: (2026)
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
by: Zhang, Yabo, et al.
Published: (2024)
by: Zhang, Yabo, et al.
Published: (2024)
SketchVideo: Sketch-based Video Generation and Editing
by: Liu, Feng-Lin, et al.
Published: (2025)
by: Liu, Feng-Lin, et al.
Published: (2025)
DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models
by: Xing, Ximing, et al.
Published: (2023)
by: Xing, Ximing, et al.
Published: (2023)
GoMAvatar: Efficient Animatable Human Modeling from Monocular Video Using Gaussians-on-Mesh
by: Wen, Jing, et al.
Published: (2024)
by: Wen, Jing, et al.
Published: (2024)
StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback
by: Park, Jiho, et al.
Published: (2025)
by: Park, Jiho, et al.
Published: (2025)
Enabling Versatile Controls for Video Diffusion Models
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
Implicit Style-Content Separation using B-LoRA
by: Frenkel, Yarden, et al.
Published: (2024)
by: Frenkel, Yarden, et al.
Published: (2024)
Lumiere: A Space-Time Diffusion Model for Video Generation
by: Bar-Tal, Omer, et al.
Published: (2024)
by: Bar-Tal, Omer, et al.
Published: (2024)
OW-VISCapTor: Abstractors for Open-World Video Instance Segmentation and Captioning
by: Choudhuri, Anwesa, et al.
Published: (2024)
by: Choudhuri, Anwesa, et al.
Published: (2024)
V-Bridge: Bridging Video Generative Priors to Versatile Few-shot Image Restoration
by: Zheng, Shenghe, et al.
Published: (2026)
by: Zheng, Shenghe, et al.
Published: (2026)
On Inductive Biases That Enable Generalization of Diffusion Transformers
by: An, Jie, et al.
Published: (2024)
by: An, Jie, et al.
Published: (2024)
pOps: Photo-Inspired Diffusion Operators
by: Richardson, Elad, et al.
Published: (2024)
by: Richardson, Elad, et al.
Published: (2024)
Pseudo-Generalized Dynamic View Synthesis from a Video
by: Zhao, Xiaoming, et al.
Published: (2023)
by: Zhao, Xiaoming, et al.
Published: (2023)
Putting the Object Back into Video Object Segmentation
by: Cheng, Ho Kei, et al.
Published: (2023)
by: Cheng, Ho Kei, et al.
Published: (2023)
UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors
by: Chen, Houyuan, et al.
Published: (2026)
by: Chen, Houyuan, et al.
Published: (2026)
MyVLM: Personalizing VLMs for User-Specific Queries
by: Alaluf, Yuval, et al.
Published: (2024)
by: Alaluf, Yuval, et al.
Published: (2024)
VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models
by: Huang, Ziqi, et al.
Published: (2024)
by: Huang, Ziqi, et al.
Published: (2024)
Versatile Transition Generation with Image-to-Video Diffusion
by: Yang, Zuhao, et al.
Published: (2025)
by: Yang, Zuhao, et al.
Published: (2025)
AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
by: Chai, Wenhao, et al.
Published: (2024)
by: Chai, Wenhao, et al.
Published: (2024)
DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving
by: Shi, Chen, et al.
Published: (2026)
by: Shi, Chen, et al.
Published: (2026)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
by: Yang, Ruolin, et al.
Published: (2025)
by: Yang, Ruolin, et al.
Published: (2025)
VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control
by: Jiang, Lifan, et al.
Published: (2025)
by: Jiang, Lifan, et al.
Published: (2025)
PriorNet: Prior-Guided Engagement Estimation from Face Video
by: Vedernikov, Alexander
Published: (2026)
by: Vedernikov, Alexander
Published: (2026)
LiveSVG: Zero-Shot SVG Animation via Video Generation
by: Levy, Matan, et al.
Published: (2026)
by: Levy, Matan, et al.
Published: (2026)
Generative Neural Video Compression via Video Diffusion Prior
by: Mao, Qi, et al.
Published: (2025)
by: Mao, Qi, et al.
Published: (2025)
Image-Guided Geometric Stylization of 3D Meshes
by: Choi, Changwoon, et al.
Published: (2026)
by: Choi, Changwoon, et al.
Published: (2026)
DeepSketcher: Internalizing Visual Manipulation for Multimodal Reasoning
by: Zhang, Chi, et al.
Published: (2025)
by: Zhang, Chi, et al.
Published: (2025)
DVD: Deterministic Video Depth Estimation with Generative Priors
by: Zhang, Hongfei, et al.
Published: (2026)
by: Zhang, Hongfei, et al.
Published: (2026)
VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
by: Chefer, Hila, et al.
Published: (2025)
by: Chefer, Hila, et al.
Published: (2025)
VIRES: Video Instance Repainting via Sketch and Text Guided Generation
by: Weng, Shuchen, et al.
Published: (2024)
by: Weng, Shuchen, et al.
Published: (2024)
The Curse of Conditions: Analyzing and Improving Optimal Transport for Conditional Flow-Based Generation
by: Cheng, Ho Kei, et al.
Published: (2025)
by: Cheng, Ho Kei, et al.
Published: (2025)
Similar Items
-
SketchAgent: Language-Driven Sequential Sketch Generation
by: Vinker, Yael, et al.
Published: (2024) -
NeuralSVG: An Implicit Representation for Text-to-Vector Generation
by: Polaczek, Sagi, et al.
Published: (2025) -
Generative Visual Communication in the Era of Vision-Language Models
by: Vinker, Yael
Published: (2024) -
SwiftSketch: A Diffusion Model for Image-to-Vector Sketch Generation
by: Arar, Ellie, et al.
Published: (2025) -
Instance Segmentation of Scene Sketches Using Natural Image Priors
by: Tang, Mia, et al.
Published: (2025)