ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kara, Ozgur, Singh, Krishna Kumar, Liu, Feng, Ceylan, Duygu, Rehg, James M., Hinz, Tobias |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiffEye: Diffusion-Based Continuous Eye-Tracking Data Generation Conditioned on Natural Images
by: Kara, Ozgur, et al.
Published: (2025)
by: Kara, Ozgur, et al.
Published: (2025)
Immune2V: Image Immunization Against Dual-Stream Image-to-Video Generation
by: Long, Zeqian, et al.
Published: (2026)
by: Long, Zeqian, et al.
Published: (2026)
Few-Shot-Based Modular Image-to-Video Adapter for Diffusion Models
by: Li, Zhenhao, et al.
Published: (2025)
by: Li, Zhenhao, et al.
Published: (2025)
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
by: Ren, Yixuan, et al.
Published: (2024)
by: Ren, Yixuan, et al.
Published: (2024)
EchoShot: Multi-Shot Portrait Video Generation
by: Wang, Jiahao, et al.
Published: (2025)
by: Wang, Jiahao, et al.
Published: (2025)
CapS-Adapter: Caption-based MultiModal Adapter in Zero-Shot Classification
by: Wang, Qijie, et al.
Published: (2024)
by: Wang, Qijie, et al.
Published: (2024)
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework
by: Wang, Qinghe, et al.
Published: (2025)
by: Wang, Qinghe, et al.
Published: (2025)
Layer-Aware Video Composition via Split-then-Merge
by: Kara, Ozgur, et al.
Published: (2025)
by: Kara, Ozgur, et al.
Published: (2025)
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
by: Luo, Yawen, et al.
Published: (2026)
by: Luo, Yawen, et al.
Published: (2026)
ShotVerse: Advancing Cinematic Camera Control for Text-Driven Multi-Shot Video Creation
by: Yang, Songlin, et al.
Published: (2026)
by: Yang, Songlin, et al.
Published: (2026)
Zero-Shot Video Deraining with Video Diffusion Models
by: Varanka, Tuomas, et al.
Published: (2025)
by: Varanka, Tuomas, et al.
Published: (2025)
One-Shot Diffusion Mimicker for Handwritten Text Generation
by: Dai, Gang, et al.
Published: (2024)
by: Dai, Gang, et al.
Published: (2024)
Efficient Text-Guided Convolutional Adapter for the Diffusion Model
by: Das, Aryan, et al.
Published: (2026)
by: Das, Aryan, et al.
Published: (2026)
ShotDirector: Directorially Controllable Multi-Shot Video Generation with Cinematographic Transitions
by: Wu, Xiaoxue, et al.
Published: (2025)
by: Wu, Xiaoxue, et al.
Published: (2025)
Zero-Shot Video Restoration and Enhancement with Assistance of Video Diffusion Models
by: Cao, Cong, et al.
Published: (2026)
by: Cao, Cong, et al.
Published: (2026)
DiffVax: Optimization-Free Image Immunization Against Diffusion-Based Editing
by: Ozden, Tarik Can, et al.
Published: (2024)
by: Ozden, Tarik Can, et al.
Published: (2024)
I2V-Adapter: A General Image-to-Video Adapter for Diffusion Models
by: Guo, Xun, et al.
Published: (2023)
by: Guo, Xun, et al.
Published: (2023)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
VideoHandles: Editing 3D Object Compositions in Videos Using Video Generative Priors
by: Koo, Juil, et al.
Published: (2025)
by: Koo, Juil, et al.
Published: (2025)
MD-ProjTex: Texturing 3D Shapes with Multi-Diffusion Projection
by: Yildirim, Ahmet Burak, et al.
Published: (2025)
by: Yildirim, Ahmet Burak, et al.
Published: (2025)
Improving Video Diffusion Transformer Training by Multi-Feature Fusion and Alignment from Self-Supervised Vision Encoders
by: Lee, Dohun, et al.
Published: (2025)
by: Lee, Dohun, et al.
Published: (2025)
Motion-Adapter: A Diffusion Model Adapter for Text-to-Motion Generation of Compound Actions
by: Jiang, Yue, et al.
Published: (2026)
by: Jiang, Yue, et al.
Published: (2026)
TI2V-Zero: Zero-Shot Image Conditioning for Text-to-Video Diffusion Models
by: Ni, Haomiao, et al.
Published: (2024)
by: Ni, Haomiao, et al.
Published: (2024)
3D Stylization via Large Reconstruction Model
by: Oztas, Ipek, et al.
Published: (2025)
by: Oztas, Ipek, et al.
Published: (2025)
Hold-One-Shot-Out (HOSO) for Validation-Free Few-Shot CLIP Adapters
by: Vorster, Chris, et al.
Published: (2026)
by: Vorster, Chris, et al.
Published: (2026)
SonicDiffusion: Audio-Driven Image Generation and Editing with Pretrained Diffusion Models
by: Biner, Burak Can, et al.
Published: (2024)
by: Biner, Burak Can, et al.
Published: (2024)
SNED: Superposition Network Architecture Search for Efficient Video Diffusion Model
by: Li, Zhengang, et al.
Published: (2024)
by: Li, Zhengang, et al.
Published: (2024)
GANFusion: Feed-Forward Text-to-3D with Diffusion in GAN Space
by: Attaiki, Souhaib, et al.
Published: (2024)
by: Attaiki, Souhaib, et al.
Published: (2024)
One-Shot Learning Meets Depth Diffusion in Multi-Object Videos
by: Jain, Anisha
Published: (2024)
by: Jain, Anisha
Published: (2024)
Slicedit: Zero-Shot Video Editing With Text-to-Image Diffusion Models Using Spatio-Temporal Slices
by: Cohen, Nathaniel, et al.
Published: (2024)
by: Cohen, Nathaniel, et al.
Published: (2024)
MotionShot: Adaptive Motion Transfer across Arbitrary Objects for Text-to-Video Generation
by: Liu, Yanchen, et al.
Published: (2025)
by: Liu, Yanchen, et al.
Published: (2025)
OneStory: Coherent Multi-Shot Video Generation with Adaptive Memory
by: An, Zhaochong, et al.
Published: (2025)
by: An, Zhaochong, et al.
Published: (2025)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
by: Hu, Teng, et al.
Published: (2023)
by: Hu, Teng, et al.
Published: (2023)
NVS-Solver: Video Diffusion Model as Zero-Shot Novel View Synthesizer
by: You, Meng, et al.
Published: (2024)
by: You, Meng, et al.
Published: (2024)
Scaling Zero-Shot Reference-to-Video Generation
by: Zhou, Zijian, et al.
Published: (2025)
by: Zhou, Zijian, et al.
Published: (2025)
From Generation to Generalization: Emergent Few-Shot Learning in Video Diffusion Models
by: Acuaviva, Pablo, et al.
Published: (2025)
by: Acuaviva, Pablo, et al.
Published: (2025)
DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter
by: Dong, Ziyi, et al.
Published: (2022)
by: Dong, Ziyi, et al.
Published: (2022)
Liberating Seen Classes: Boosting Few-Shot and Zero-Shot Text Classification via Anchor Generation and Classification Reframing
by: Liu, Han, et al.
Published: (2024)
by: Liu, Han, et al.
Published: (2024)
FilmWeaver: Weaving Consistent Multi-Shot Videos with Cache-Guided Autoregressive Diffusion
by: Luo, Xiangyang, et al.
Published: (2025)
by: Luo, Xiangyang, et al.
Published: (2025)
Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models
by: Wang, Wen, et al.
Published: (2023)
by: Wang, Wen, et al.
Published: (2023)
Similar Items
-
DiffEye: Diffusion-Based Continuous Eye-Tracking Data Generation Conditioned on Natural Images
by: Kara, Ozgur, et al.
Published: (2025) -
Immune2V: Image Immunization Against Dual-Stream Image-to-Video Generation
by: Long, Zeqian, et al.
Published: (2026) -
Few-Shot-Based Modular Image-to-Video Adapter for Diffusion Models
by: Li, Zhenhao, et al.
Published: (2025) -
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
by: Ren, Yixuan, et al.
Published: (2024) -
EchoShot: Multi-Shot Portrait Video Generation
by: Wang, Jiahao, et al.
Published: (2025)