VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Lifan, Chen, Shuang, Wu, Boxi, Guan, Xiaotong, Zhang, Jiahui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HuViDPO:Enhancing Video Generation through Direct Preference Optimization for Human-Centric Alignment
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
von: Zhan, Ruohao, et al.
Veröffentlicht: (2025)
von: Zhan, Ruohao, et al.
Veröffentlicht: (2025)
SketchJudge: A Diagnostic Benchmark for Grading Hand-drawn Diagrams with Multimodal Large Language Models
von: Su, Yuhang, et al.
Veröffentlicht: (2026)
von: Su, Yuhang, et al.
Veröffentlicht: (2026)
AirSketch: Generative Motion to Sketch
von: Lim, Hui Xian Grace, et al.
Veröffentlicht: (2024)
von: Lim, Hui Xian Grace, et al.
Veröffentlicht: (2024)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
von: Brioschi, Riccardo, et al.
Veröffentlicht: (2025)
von: Brioschi, Riccardo, et al.
Veröffentlicht: (2025)
U-Sketch: An Efficient Approach for Sketch to Image Diffusion Models
von: Mitsouras, Ilias, et al.
Veröffentlicht: (2024)
von: Mitsouras, Ilias, et al.
Veröffentlicht: (2024)
UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
S3D: Sketch-Driven 3D Model Generation
von: Song, Hail, et al.
Veröffentlicht: (2025)
von: Song, Hail, et al.
Veröffentlicht: (2025)
Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation
von: Chen, Minglin, et al.
Veröffentlicht: (2024)
von: Chen, Minglin, et al.
Veröffentlicht: (2024)
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
von: Cheng, Shenggan, et al.
Veröffentlicht: (2025)
von: Cheng, Shenggan, et al.
Veröffentlicht: (2025)
ViSketch-GPT: Collaborative Multi-Scale Feature Extraction for Sketch Recognition and Generation
von: Federico, Giulio, et al.
Veröffentlicht: (2025)
von: Federico, Giulio, et al.
Veröffentlicht: (2025)
KnobGen: Controlling the Sophistication of Artwork in Sketch-Based Diffusion Models
von: Navard, Pouyan, et al.
Veröffentlicht: (2024)
von: Navard, Pouyan, et al.
Veröffentlicht: (2024)
Gen-AI Police Sketches with Stable Diffusion
von: Fidalgo, Nicholas, et al.
Veröffentlicht: (2025)
von: Fidalgo, Nicholas, et al.
Veröffentlicht: (2025)
SketchRef: a Multi-Task Evaluation Benchmark for Sketch Synthesis
von: Lin, Xingyue, et al.
Veröffentlicht: (2024)
von: Lin, Xingyue, et al.
Veröffentlicht: (2024)
SketchINR: A First Look into Sketches as Implicit Neural Representations
von: Bandyopadhyay, Hmrishav, et al.
Veröffentlicht: (2024)
von: Bandyopadhyay, Hmrishav, et al.
Veröffentlicht: (2024)
Generating Sketches in a Hierarchical Auto-Regressive Process for Flexible Sketch Drawing Manipulation at Stroke-Level
von: Zang, Sicong, et al.
Veröffentlicht: (2025)
von: Zang, Sicong, et al.
Veröffentlicht: (2025)
InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward
von: Ning, Zhiwei, et al.
Veröffentlicht: (2026)
von: Ning, Zhiwei, et al.
Veröffentlicht: (2026)
Planning with Sketch-Guided Verification for Physics-Aware Video Generation
von: Huang, Yidong, et al.
Veröffentlicht: (2025)
von: Huang, Yidong, et al.
Veröffentlicht: (2025)
Equipping Sketch Patches with Context-Aware Positional Encoding for Graphic Sketch Representation
von: Zang, Sicong, et al.
Veröffentlicht: (2024)
von: Zang, Sicong, et al.
Veröffentlicht: (2024)
RelightVid: Temporal-Consistent Diffusion Model for Video Relighting
von: Fang, Ye, et al.
Veröffentlicht: (2025)
von: Fang, Ye, et al.
Veröffentlicht: (2025)
Content-Conditioned Generation of Stylized Free hand Sketches
von: Liu, Jiajun, et al.
Veröffentlicht: (2024)
von: Liu, Jiajun, et al.
Veröffentlicht: (2024)
EDT: An Efficient Diffusion Transformer Framework Inspired by Human-like Sketching
von: Chen, Xinwang, et al.
Veröffentlicht: (2024)
von: Chen, Xinwang, et al.
Veröffentlicht: (2024)
SketchGraphNet: A Memory-Efficient Hybrid Graph Transformer for Large-Scale Sketch Corpora Recognition
von: Chen, Shilong, et al.
Veröffentlicht: (2026)
von: Chen, Shilong, et al.
Veröffentlicht: (2026)
On the Temporality for Sketch Representation Learning
von: Junior, Marcelo Isaias de Moraes, et al.
Veröffentlicht: (2025)
von: Junior, Marcelo Isaias de Moraes, et al.
Veröffentlicht: (2025)
Terrain Diffusion Network: Climatic-Aware Terrain Generation with Geological Sketch Guidance
von: Hu, Zexin, et al.
Veröffentlicht: (2023)
von: Hu, Zexin, et al.
Veröffentlicht: (2023)
CadVLM: Bridging Language and Vision in the Generation of Parametric CAD Sketches
von: Wu, Sifan, et al.
Veröffentlicht: (2024)
von: Wu, Sifan, et al.
Veröffentlicht: (2024)
DrawVideo: Generating Long Video from Storyboard Keyframe Sketches
von: Xu, Chuanzhi, et al.
Veröffentlicht: (2026)
von: Xu, Chuanzhi, et al.
Veröffentlicht: (2026)
DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
InstructVid2Vid: Controllable Video Editing with Natural Language Instructions
von: Qin, Bosheng, et al.
Veröffentlicht: (2023)
von: Qin, Bosheng, et al.
Veröffentlicht: (2023)
UniVid: Pyramid Diffusion Model for High Quality Video Generation
von: Xiao, Xinyu, et al.
Veröffentlicht: (2026)
von: Xiao, Xinyu, et al.
Veröffentlicht: (2026)
VidLaDA: Bidirectional Diffusion Large Language Models for Efficient Video Understanding
von: He, Zhihao, et al.
Veröffentlicht: (2026)
von: He, Zhihao, et al.
Veröffentlicht: (2026)
Freehand Sketch Generation from Mechanical Components
von: Liao, Zhichao, et al.
Veröffentlicht: (2024)
von: Liao, Zhichao, et al.
Veröffentlicht: (2024)
Multi-Style Facial Sketch Synthesis through Masked Generative Modeling
von: Sun, Bowen, et al.
Veröffentlicht: (2024)
von: Sun, Bowen, et al.
Veröffentlicht: (2024)
StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback
von: Park, Jiho, et al.
Veröffentlicht: (2025)
von: Park, Jiho, et al.
Veröffentlicht: (2025)
SNR-Edit: Structure-Aware Noise Rectification for Inversion-Free Flow-Based Editing
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval
von: Wang, Siyuan, et al.
Veröffentlicht: (2026)
von: Wang, Siyuan, et al.
Veröffentlicht: (2026)
Any-to-Bokeh: Arbitrary-Subject Video Refocusing with Video Diffusion Model
von: Yang, Yang, et al.
Veröffentlicht: (2025)
von: Yang, Yang, et al.
Veröffentlicht: (2025)
Sketch-guided Image Inpainting with Partial Discrete Diffusion Process
von: Sharma, Nakul, et al.
Veröffentlicht: (2024)
von: Sharma, Nakul, et al.
Veröffentlicht: (2024)
LOTS of Fashion! Multi-Conditioning for Image Generation via Sketch-Text Pairing
von: Girella, Federico, et al.
Veröffentlicht: (2025)
von: Girella, Federico, et al.
Veröffentlicht: (2025)
SketchVideo: Sketch-based Video Generation and Editing
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2025)
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HuViDPO:Enhancing Video Generation through Direct Preference Optimization for Human-Centric Alignment
von: Jiang, Lifan, et al.
Veröffentlicht: (2025) -
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
von: Zhan, Ruohao, et al.
Veröffentlicht: (2025) -
SketchJudge: A Diagnostic Benchmark for Grading Hand-drawn Diagrams with Multimodal Large Language Models
von: Su, Yuhang, et al.
Veröffentlicht: (2026) -
AirSketch: Generative Motion to Sketch
von: Lim, Hui Xian Grace, et al.
Veröffentlicht: (2024) -
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
von: Brioschi, Riccardo, et al.
Veröffentlicht: (2025)