FreeInit: Bridging Initialization Gap in Video Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Tianxing, Si, Chenyang, Jiang, Yuming, Huang, Ziqi, Liu, Ziwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ReVersion: Diffusion-Based Relation Inversion from Images
von: Huang, Ziqi, et al.
Veröffentlicht: (2023)
von: Huang, Ziqi, et al.
Veröffentlicht: (2023)
InitNO: Boosting Text-to-Image Diffusion Models via Initial Noise Optimization
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
FastInit: Fast Noise Initialization for Temporally Consistent Video Generation
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
RepVideo: Rethinking Cross-Layer Representation for Video Generation
von: Si, Chenyang, et al.
Veröffentlicht: (2025)
von: Si, Chenyang, et al.
Veröffentlicht: (2025)
FasterCache: Training-Free Video Diffusion Model Acceleration with High Quality
von: Lv, Zhengyao, et al.
Veröffentlicht: (2024)
von: Lv, Zhengyao, et al.
Veröffentlicht: (2024)
RealDPO: Real or Not Real, that is the Preference
von: Cheng, Guo, et al.
Veröffentlicht: (2025)
von: Cheng, Guo, et al.
Veröffentlicht: (2025)
Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation
von: Chen, Gordon, et al.
Veröffentlicht: (2026)
von: Chen, Gordon, et al.
Veröffentlicht: (2026)
Robust Sequential DeepFake Detection
von: Shao, Rui, et al.
Veröffentlicht: (2023)
von: Shao, Rui, et al.
Veröffentlicht: (2023)
FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models
von: Qiu, Haonan, et al.
Veröffentlicht: (2024)
von: Qiu, Haonan, et al.
Veröffentlicht: (2024)
VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models
von: Huang, Ziqi, et al.
Veröffentlicht: (2024)
von: Huang, Ziqi, et al.
Veröffentlicht: (2024)
Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models
von: Fan, Weichen, et al.
Veröffentlicht: (2025)
von: Fan, Weichen, et al.
Veröffentlicht: (2025)
Bridging Semantic and Kinematic Conditions with Diffusion-based Discrete Motion Tokenizer
von: Gu, Chenyang, et al.
Veröffentlicht: (2026)
von: Gu, Chenyang, et al.
Veröffentlicht: (2026)
PoInit-of-View: Poisoning Initialization of Views Transfers Across Multiple 3D Reconstruction Systems
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
DeepFake-Adapter: Dual-Level Adapter for DeepFake Detection
von: Shao, Rui, et al.
Veröffentlicht: (2023)
von: Shao, Rui, et al.
Veröffentlicht: (2023)
Dual-Expert Consistency Model for Efficient and High-Quality Video Generation
von: Lv, Zhengyao, et al.
Veröffentlicht: (2025)
von: Lv, Zhengyao, et al.
Veröffentlicht: (2025)
CineScale: Free Lunch in High-Resolution Cinematic Visual Generation
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling
von: Qiu, Haonan, et al.
Veröffentlicht: (2023)
von: Qiu, Haonan, et al.
Veröffentlicht: (2023)
LongVie: Multimodal-Guided Controllable Ultra-Long Video Generation
von: Gao, Jianxiong, et al.
Veröffentlicht: (2025)
von: Gao, Jianxiong, et al.
Veröffentlicht: (2025)
ActiveInitSplat: How Active Image Selection Helps Gaussian Splatting
von: Polyzos, Konstantinos D., et al.
Veröffentlicht: (2025)
von: Polyzos, Konstantinos D., et al.
Veröffentlicht: (2025)
LongVie 2: Multimodal Controllable Ultra-Long Video World Model
von: Gao, Jianxiong, et al.
Veröffentlicht: (2025)
von: Gao, Jianxiong, et al.
Veröffentlicht: (2025)
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
Rethinking Cross-Modal Interaction in Multimodal Diffusion Transformers
von: Lv, Zhengyao, et al.
Veröffentlicht: (2025)
von: Lv, Zhengyao, et al.
Veröffentlicht: (2025)
Towards Language-Driven Video Inpainting via Multimodal Large Language Models
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
GOOD: Training-Free Guided Diffusion Sampling for Out-of-Distribution Detection
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
GroupDiff: Diffusion-based Group Portrait Editing
von: Jiang, Yuming, et al.
Veröffentlicht: (2024)
von: Jiang, Yuming, et al.
Veröffentlicht: (2024)
Dino-Diffusion Modular Designs Bridge the Cross-Domain Gap in Autonomous Parking
von: Wu, Zixuan, et al.
Veröffentlicht: (2025)
von: Wu, Zixuan, et al.
Veröffentlicht: (2025)
Bridging the Micro--Macro Gap: Frequency-Aware Semantic Alignment for Image Manipulation Localization
von: Liang, Xiaojie, et al.
Veröffentlicht: (2026)
von: Liang, Xiaojie, et al.
Veröffentlicht: (2026)
StableWorld: Towards Stable and Consistent Long Interactive Video Generation
von: Yang, Ying, et al.
Veröffentlicht: (2026)
von: Yang, Ying, et al.
Veröffentlicht: (2026)
Bridging the Gap: Aligning Text-to-Image Diffusion Models with Specific Feedback
von: Niu, Xuexiang, et al.
Veröffentlicht: (2024)
von: Niu, Xuexiang, et al.
Veröffentlicht: (2024)
Can Diffusion Models Bridge the Domain Gap in Cardiac MR Imaging?
von: Wong, Xin Ci, et al.
Veröffentlicht: (2025)
von: Wong, Xin Ci, et al.
Veröffentlicht: (2025)
Stencil: Subject-Driven Generation with Context Guidance
von: Chen, Gordon, et al.
Veröffentlicht: (2025)
von: Chen, Gordon, et al.
Veröffentlicht: (2025)
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
von: Huang, Xun, et al.
Veröffentlicht: (2025)
von: Huang, Xun, et al.
Veröffentlicht: (2025)
Bridging the Intent Gap: Knowledge-Enhanced Visual Generation
von: Cheng, Yi, et al.
Veröffentlicht: (2024)
von: Cheng, Yi, et al.
Veröffentlicht: (2024)
DBINDS -- Can Initial Noise from Diffusion Model Inversion Help Reveal AI-Generated Videos?
von: Wu, Yanlin, et al.
Veröffentlicht: (2025)
von: Wu, Yanlin, et al.
Veröffentlicht: (2025)
Multi-clue Consistency Learning to Bridge Gaps Between General and Oriented Object in Semi-supervised Detection
von: Wang, Chenxu, et al.
Veröffentlicht: (2024)
von: Wang, Chenxu, et al.
Veröffentlicht: (2024)
FreeScale: Unleashing the Resolution of Diffusion Models via Tuning-Free Scale Fusion
von: Qiu, Haonan, et al.
Veröffentlicht: (2024)
von: Qiu, Haonan, et al.
Veröffentlicht: (2024)
V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
von: Cheng, Zixu, et al.
Veröffentlicht: (2025)
von: Cheng, Zixu, et al.
Veröffentlicht: (2025)
LSGQuant: Layer-Sensitivity Guided Quantization for One-Step Diffusion Real-World Video Super-Resolution
von: Wu, Tianxing, et al.
Veröffentlicht: (2026)
von: Wu, Tianxing, et al.
Veröffentlicht: (2026)
Layered Rendering Diffusion Model for Controllable Zero-Shot Image Synthesis
von: Qi, Zipeng, et al.
Veröffentlicht: (2023)
von: Qi, Zipeng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
ReVersion: Diffusion-Based Relation Inversion from Images
von: Huang, Ziqi, et al.
Veröffentlicht: (2023) -
InitNO: Boosting Text-to-Image Diffusion Models via Initial Noise Optimization
von: Guo, Xiefan, et al.
Veröffentlicht: (2024) -
FastInit: Fast Noise Initialization for Temporally Consistent Video Generation
von: Bai, Chengyu, et al.
Veröffentlicht: (2025) -
FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model
von: Cao, Yukang, et al.
Veröffentlicht: (2025) -
RepVideo: Rethinking Cross-Layer Representation for Video Generation
von: Si, Chenyang, et al.
Veröffentlicht: (2025)