Decoupled Video Generation with Chain of Training-free Diffusion Model Experts
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Wenhao, Cao, Yichao, Su, Xiu, Lin, Xi, You, Shan, Zheng, Mingkai, Chen, Yi, Xu, Chang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Adaptive Training Meets Progressive Scaling: Elevating Efficiency in Diffusion Models
di: Li, Wenhao, et al.
Pubblicazione: (2023)
di: Li, Wenhao, et al.
Pubblicazione: (2023)
ZeroSmooth: Training-free Diffuser Adaptation for High Frame Rate Video Generation
di: Yang, Shaoshu, et al.
Pubblicazione: (2024)
di: Yang, Shaoshu, et al.
Pubblicazione: (2024)
MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation
di: Shi, Shuwei, et al.
Pubblicazione: (2024)
di: Shi, Shuwei, et al.
Pubblicazione: (2024)
GVDIFF: Grounded Text-to-Video Generation with Diffusion Models
di: Dou, Huanzhang, et al.
Pubblicazione: (2024)
di: Dou, Huanzhang, et al.
Pubblicazione: (2024)
Identify, Isolate, and Purge: Mitigating Hallucinations in LVLMs via Self-Evolving Distillation
di: Li, Wenhao, et al.
Pubblicazione: (2025)
di: Li, Wenhao, et al.
Pubblicazione: (2025)
TVG: A Training-free Transition Video Generation Method with Diffusion Models
di: Zhang, Rui, et al.
Pubblicazione: (2024)
di: Zhang, Rui, et al.
Pubblicazione: (2024)
VideoMaker: Zero-shot Customized Video Generation with the Inherent Force of Video Diffusion Models
di: Wu, Tao, et al.
Pubblicazione: (2024)
di: Wu, Tao, et al.
Pubblicazione: (2024)
Cut to the Chase: Training-free Multimodal Summarization via Chain-of-Events
di: You, Xiaoxing, et al.
Pubblicazione: (2026)
di: You, Xiaoxing, et al.
Pubblicazione: (2026)
Video Diffusion Models are Training-free Motion Interpreter and Controller
di: Xiao, Zeqi, et al.
Pubblicazione: (2024)
di: Xiao, Zeqi, et al.
Pubblicazione: (2024)
CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
di: Yang, Zhuoyi, et al.
Pubblicazione: (2024)
di: Yang, Zhuoyi, et al.
Pubblicazione: (2024)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
di: He, Xu, et al.
Pubblicazione: (2024)
di: He, Xu, et al.
Pubblicazione: (2024)
Weak Augmentation Guided Relational Self-Supervised Learning
di: Zheng, Mingkai, et al.
Pubblicazione: (2022)
di: Zheng, Mingkai, et al.
Pubblicazione: (2022)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
di: Zheng, Guangcong, et al.
Pubblicazione: (2023)
di: Zheng, Guangcong, et al.
Pubblicazione: (2023)
MFTF: Mask-free Training-free Object Level Layout Control Diffusion Model
di: Yang, Shan
Pubblicazione: (2024)
di: Yang, Shan
Pubblicazione: (2024)
VideoUFO: A Million-Scale User-Focused Dataset for Text-to-Video Generation
di: Wang, Wenhao, et al.
Pubblicazione: (2025)
di: Wang, Wenhao, et al.
Pubblicazione: (2025)
Consistent Human Image and Video Generation with Spatially Conditioned Diffusion
di: Cao, Mingdeng, et al.
Pubblicazione: (2024)
di: Cao, Mingdeng, et al.
Pubblicazione: (2024)
Anchored Video Generation: Decoupling Scene Construction and Temporal Synthesis in Text-to-Video Diffusion Models
di: Hassan, Mariam, et al.
Pubblicazione: (2025)
di: Hassan, Mariam, et al.
Pubblicazione: (2025)
SAU: A Dual-Branch Network to Enhance Long-Tailed Recognition via Generative Models
di: Li, Guangxi, et al.
Pubblicazione: (2024)
di: Li, Guangxi, et al.
Pubblicazione: (2024)
Audio-visual Controlled Video Diffusion with Masked Selective State Spaces Modeling for Natural Talking Head Generation
di: Hong, Fa-Ting, et al.
Pubblicazione: (2025)
di: Hong, Fa-Ting, et al.
Pubblicazione: (2025)
MetaDD: Boosting Dataset Distillation with Neural Network Architecture-Invariant Generalization
di: Zhao, Yunlong, et al.
Pubblicazione: (2024)
di: Zhao, Yunlong, et al.
Pubblicazione: (2024)
Unpaired Deblurring via Decoupled Diffusion Model
di: Cheng, Junhao, et al.
Pubblicazione: (2025)
di: Cheng, Junhao, et al.
Pubblicazione: (2025)
ByTheWay: Boost Your Text-to-Video Generation Model to Higher Quality in a Training-free Way
di: Bu, Jiazi, et al.
Pubblicazione: (2024)
di: Bu, Jiazi, et al.
Pubblicazione: (2024)
VidProM: A Million-scale Real Prompt-Gallery Dataset for Text-to-Video Diffusion Models
di: Wang, Wenhao, et al.
Pubblicazione: (2024)
di: Wang, Wenhao, et al.
Pubblicazione: (2024)
Training-free Regional Prompting for Diffusion Transformers
di: Chen, Anthony, et al.
Pubblicazione: (2024)
di: Chen, Anthony, et al.
Pubblicazione: (2024)
Image Anything: Towards Reasoning-coherent and Training-free Multi-modal Image Generation
di: Lyu, Yuanhuiyi, et al.
Pubblicazione: (2024)
di: Lyu, Yuanhuiyi, et al.
Pubblicazione: (2024)
Decoupling Dynamic Monocular Videos for Dynamic View Synthesis
di: You, Meng, et al.
Pubblicazione: (2023)
di: You, Meng, et al.
Pubblicazione: (2023)
AGLLDiff: Guiding Diffusion Models Towards Unsupervised Training-free Real-world Low-light Image Enhancement
di: Lin, Yunlong, et al.
Pubblicazione: (2024)
di: Lin, Yunlong, et al.
Pubblicazione: (2024)
Training-free Camera Control for Video Generation
di: Hou, Chen, et al.
Pubblicazione: (2024)
di: Hou, Chen, et al.
Pubblicazione: (2024)
MOVi: Training-free Text-conditioned Multi-Object Video Generation
di: Rahman, Aimon, et al.
Pubblicazione: (2025)
di: Rahman, Aimon, et al.
Pubblicazione: (2025)
Learning Long-form Video Prior via Generative Pre-Training
di: Xie, Jinheng, et al.
Pubblicazione: (2024)
di: Xie, Jinheng, et al.
Pubblicazione: (2024)
VideoMerge: Towards Training-free Long Video Generation
di: Zhang, Siyang, et al.
Pubblicazione: (2025)
di: Zhang, Siyang, et al.
Pubblicazione: (2025)
MotionMaster: Training-free Camera Motion Transfer For Video Generation
di: Hu, Teng, et al.
Pubblicazione: (2024)
di: Hu, Teng, et al.
Pubblicazione: (2024)
OmnimatteZero: Fast Training-free Omnimatte with Pre-trained Video Diffusion Models
di: Samuel, Dvir, et al.
Pubblicazione: (2025)
di: Samuel, Dvir, et al.
Pubblicazione: (2025)
Train Short, Inference Long: Training-free Horizon Extension for Autoregressive Video Generation
di: Li, Jia, et al.
Pubblicazione: (2026)
di: Li, Jia, et al.
Pubblicazione: (2026)
Edit-Your-Motion: Space-Time Diffusion Decoupling Learning for Video Motion Editing
di: Zuo, Yi, et al.
Pubblicazione: (2024)
di: Zuo, Yi, et al.
Pubblicazione: (2024)
Decoupling Training-Free Guided Diffusion by ADMM
di: Zhang, Youyuan, et al.
Pubblicazione: (2024)
di: Zhang, Youyuan, et al.
Pubblicazione: (2024)
ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
di: Fang, Zixun, et al.
Pubblicazione: (2025)
di: Fang, Zixun, et al.
Pubblicazione: (2025)
Training-free Motion Factorization for Compositional Video Generation
di: Wang, Zixuan, et al.
Pubblicazione: (2026)
di: Wang, Zixuan, et al.
Pubblicazione: (2026)
UniVST: A Unified Framework for Training-free Localized Video Style Transfer
di: Song, Quanjian, et al.
Pubblicazione: (2024)
di: Song, Quanjian, et al.
Pubblicazione: (2024)
Hierarchical Codec Diffusion for Video-to-Speech Generation
di: Ye, Jiaxin, et al.
Pubblicazione: (2026)
di: Ye, Jiaxin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Adaptive Training Meets Progressive Scaling: Elevating Efficiency in Diffusion Models
di: Li, Wenhao, et al.
Pubblicazione: (2023) -
ZeroSmooth: Training-free Diffuser Adaptation for High Frame Rate Video Generation
di: Yang, Shaoshu, et al.
Pubblicazione: (2024) -
MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation
di: Shi, Shuwei, et al.
Pubblicazione: (2024) -
GVDIFF: Grounded Text-to-Video Generation with Diffusion Models
di: Dou, Huanzhang, et al.
Pubblicazione: (2024) -
Identify, Isolate, and Purge: Mitigating Hallucinations in LVLMs via Self-Evolving Distillation
di: Li, Wenhao, et al.
Pubblicazione: (2025)