Diffusion-DRF: Free, Rich, and Differentiable Reward for Video Diffusion Fine-Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yifan, Li, Yanyu, Qian, Gordon Guocheng, Tulyakov, Sergey, Fu, Yun, Kag, Anil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EFlow: Fast Few-Step Video Generator Training from Scratch via Efficient Solution Flow
von: Park, Dogyun, et al.
Veröffentlicht: (2026)
von: Park, Dogyun, et al.
Veröffentlicht: (2026)
DenseDPO: Fine-Grained Temporal Preference Optimization for Video Diffusion Models
von: Wu, Ziyi, et al.
Veröffentlicht: (2025)
von: Wu, Ziyi, et al.
Veröffentlicht: (2025)
Sprint: Sparse-Dense Residual Fusion for Efficient Diffusion Transformers
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
H3AE: High Compression, High Speed, and High Quality AutoEncoder for Video Diffusion Models
von: Wu, Yushu, et al.
Veröffentlicht: (2025)
von: Wu, Yushu, et al.
Veröffentlicht: (2025)
BitsFusion: 1.99 bits Weight Quantization of Diffusion Model
von: Sui, Yang, et al.
Veröffentlicht: (2024)
von: Sui, Yang, et al.
Veröffentlicht: (2024)
Taming Diffusion Transformer for Efficient Mobile Video Generation in Seconds
von: Wu, Yushu, et al.
Veröffentlicht: (2025)
von: Wu, Yushu, et al.
Veröffentlicht: (2025)
S2DiT: Sandwich Diffusion Transformer for Mobile Streaming Video Generation
von: Zhao, Lin, et al.
Veröffentlicht: (2026)
von: Zhao, Lin, et al.
Veröffentlicht: (2026)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
One Model, Many Budgets: Elastic Latent Interfaces for Diffusion Transformers
von: Haji-Ali, Moayed, et al.
Veröffentlicht: (2026)
von: Haji-Ali, Moayed, et al.
Veröffentlicht: (2026)
Scalable Ranked Preference Optimization for Text-to-Image Generation
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
Directly Fine-Tuning Diffusion Models on Differentiable Rewards
von: Clark, Kevin, et al.
Veröffentlicht: (2023)
von: Clark, Kevin, et al.
Veröffentlicht: (2023)
Self-NPO: Data-Free Diffusion Model Enhancement via Truncated Diffusion Fine-Tuning
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2025)
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2025)
Towards Physical Understanding in Video Generation: A 3D Point Regularization Approach
von: Chen, Yunuo, et al.
Veröffentlicht: (2025)
von: Chen, Yunuo, et al.
Veröffentlicht: (2025)
SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices
von: Hu, Dongting, et al.
Veröffentlicht: (2026)
von: Hu, Dongting, et al.
Veröffentlicht: (2026)
Hierarchical Patch Diffusion Models for High-Resolution Video Generation
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2024)
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2024)
TextCraftor: Your Text Encoder Can be Image Quality Controller
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
SF-V: Single Forward Video Generation Model
von: Zhang, Zhixing, et al.
Veröffentlicht: (2024)
von: Zhang, Zhixing, et al.
Veröffentlicht: (2024)
SPAD : Spatially Aware Multiview Diffusers
von: Kant, Yash, et al.
Veröffentlicht: (2024)
von: Kant, Yash, et al.
Veröffentlicht: (2024)
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
Improving the Diffusability of Autoencoders
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion
von: Liu, Xian, et al.
Veröffentlicht: (2023)
von: Liu, Xian, et al.
Veröffentlicht: (2023)
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
Video Motion Transfer with Diffusion Transformers
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
Temporal-Conditional Referring Video Object Segmentation with Noise-Free Text-to-Video Diffusion Model
von: Zhang, Ruixin, et al.
Veröffentlicht: (2025)
von: Zhang, Ruixin, et al.
Veröffentlicht: (2025)
CountDiffusion: Text-to-Image Synthesis with Training-Free Counting-Guidance Diffusion
von: Li, Yanyu, et al.
Veröffentlicht: (2025)
von: Li, Yanyu, et al.
Veröffentlicht: (2025)
Snap Video: Scaled Spatiotemporal Transformers for Text-to-Video Synthesis
von: Menapace, Willi, et al.
Veröffentlicht: (2024)
von: Menapace, Willi, et al.
Veröffentlicht: (2024)
AsCAN: Asymmetric Convolution-Attention Networks for Efficient Recognition and Generation
von: Kag, Anil, et al.
Veröffentlicht: (2024)
von: Kag, Anil, et al.
Veröffentlicht: (2024)
FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models
von: Qiu, Haonan, et al.
Veröffentlicht: (2024)
von: Qiu, Haonan, et al.
Veröffentlicht: (2024)
Differentially Private Fine-Tuning of Diffusion Models
von: Tsai, Yu-Lin, et al.
Veröffentlicht: (2024)
von: Tsai, Yu-Lin, et al.
Veröffentlicht: (2024)
FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling
von: Qiu, Haonan, et al.
Veröffentlicht: (2023)
von: Qiu, Haonan, et al.
Veröffentlicht: (2023)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device
von: Wu, Yushu, et al.
Veröffentlicht: (2024)
von: Wu, Yushu, et al.
Veröffentlicht: (2024)
4Real-Video: Learning Generalizable Photo-Realistic 4D Video Diffusion
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
Omni-ID: Holistic Identity Representation Designed for Generative Tasks
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
Tuning-free Visual Effect Transfer across Videos
von: Jones, Maxwell, et al.
Veröffentlicht: (2026)
von: Jones, Maxwell, et al.
Veröffentlicht: (2026)
Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models
von: Hwang, Sungwon, et al.
Veröffentlicht: (2025)
von: Hwang, Sungwon, et al.
Veröffentlicht: (2025)
AToM: Amortized Text-to-Mesh using 2D Diffusion
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EFlow: Fast Few-Step Video Generator Training from Scratch via Efficient Solution Flow
von: Park, Dogyun, et al.
Veröffentlicht: (2026) -
DenseDPO: Fine-Grained Temporal Preference Optimization for Video Diffusion Models
von: Wu, Ziyi, et al.
Veröffentlicht: (2025) -
Sprint: Sparse-Dense Residual Fusion for Efficient Diffusion Transformers
von: Park, Dogyun, et al.
Veröffentlicht: (2025) -
H3AE: High Compression, High Speed, and High Quality AutoEncoder for Video Diffusion Models
von: Wu, Yushu, et al.
Veröffentlicht: (2025) -
BitsFusion: 1.99 bits Weight Quantization of Diffusion Model
von: Sui, Yang, et al.
Veröffentlicht: (2024)