Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Luo, Weijian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
von: Wu, Junyi, et al.
Veröffentlicht: (2026)
von: Wu, Junyi, et al.
Veröffentlicht: (2026)
David and Goliath: Small One-step Model Beats Large Diffusion with Score Post-training
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
Uni-Instruct: One-step Diffusion Model through Unified Diffusion Divergence Instruction
von: Wang, Yifei, et al.
Veröffentlicht: (2025)
von: Wang, Yifei, et al.
Veröffentlicht: (2025)
Denoising Fisher Training For Neural Implicit Samplers
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
Exploring Text-to-Motion Generation with Human Preference
von: Sheng, Jenny, et al.
Veröffentlicht: (2024)
von: Sheng, Jenny, et al.
Veröffentlicht: (2024)
Diff-Instruct: A Universal Approach for Transferring Knowledge From Pre-trained Diffusion Models
von: Luo, Weijian, et al.
Veröffentlicht: (2023)
von: Luo, Weijian, et al.
Veröffentlicht: (2023)
Di$\mathtt{[M]}$O: Distilling Masked Diffusion Models into One-step Generator
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
Multi-student Diffusion Distillation for Better One-step Generators
von: Song, Yanke, et al.
Veröffentlicht: (2024)
von: Song, Yanke, et al.
Veröffentlicht: (2024)
One-Step Diffusion Distillation through Score Implicit Matching
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models
von: Kim, Sungnyun, et al.
Veröffentlicht: (2023)
von: Kim, Sungnyun, et al.
Veröffentlicht: (2023)
One-step Diffusion Models with $f$-Divergence Distribution Matching
von: Xu, Yilun, et al.
Veröffentlicht: (2025)
von: Xu, Yilun, et al.
Veröffentlicht: (2025)
One-step Diffusion Models with Bregman Density Ratio Matching
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
von: Kim, Beomsu, et al.
Veröffentlicht: (2025)
von: Kim, Beomsu, et al.
Veröffentlicht: (2025)
Boost Your Human Image Generation Model via Direct Preference Optimization
von: Na, Sanghyeon, et al.
Veröffentlicht: (2024)
von: Na, Sanghyeon, et al.
Veröffentlicht: (2024)
TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
Text-to-image Diffusion Models in Generative AI: A Survey
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2023)
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2023)
Curriculum-DPO++: Direct Preference Optimization via Data and Model Curricula for Text-to-Image Generation
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2026)
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2026)
MoLE: Enhancing Human-centric Text-to-image Diffusion via Mixture of Low-rank Experts
von: Zhu, Jie, et al.
Veröffentlicht: (2024)
von: Zhu, Jie, et al.
Veröffentlicht: (2024)
Taming the Entropy Cliff: Variable Codebook Size Quantization for Autoregressive Visual Generation
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
MM-Instruct: Generated Visual Instructions for Large Multimodal Model Alignment
von: Liu, Jihao, et al.
Veröffentlicht: (2024)
von: Liu, Jihao, et al.
Veröffentlicht: (2024)
Aligning Text to Image in Diffusion Models is Easier Than You Think
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2025)
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2025)
Flow Generator Matching
von: Huang, Zemin, et al.
Veröffentlicht: (2024)
von: Huang, Zemin, et al.
Veröffentlicht: (2024)
Learning to Instruct for Visual Instruction Tuning
von: Zhou, Zhihan, et al.
Veröffentlicht: (2025)
von: Zhou, Zhihan, et al.
Veröffentlicht: (2025)
Subject-driven Text-to-Image Generation via Preference-based Reinforcement Learning
von: Miao, Yanting, et al.
Veröffentlicht: (2024)
von: Miao, Yanting, et al.
Veröffentlicht: (2024)
Fine Structure-Aware Sampling: A New Sampling Training Scheme for Pixel-Aligned Implicit Models in Single-View Human Reconstruction
von: Chan, Kennard Yanting, et al.
Veröffentlicht: (2024)
von: Chan, Kennard Yanting, et al.
Veröffentlicht: (2024)
Towards Effective Usage of Human-Centric Priors in Diffusion Models for Text-based Human Image Generation
von: Wang, Junyan, et al.
Veröffentlicht: (2024)
von: Wang, Junyan, et al.
Veröffentlicht: (2024)
Fine-tuning CLIP Text Encoders with Two-step Paraphrasing
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
Diffusion Adversarial Post-Training for One-Step Video Generation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
von: Prabhudesai, Mihir, et al.
Veröffentlicht: (2023)
von: Prabhudesai, Mihir, et al.
Veröffentlicht: (2023)
DiffFinger: Advancing Synthetic Fingerprint Generation through Denoising Diffusion Probabilistic Models
von: Grabovski, Freddie, et al.
Veröffentlicht: (2024)
von: Grabovski, Freddie, et al.
Veröffentlicht: (2024)
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
von: Liu, Tao, et al.
Veröffentlicht: (2025)
von: Liu, Tao, et al.
Veröffentlicht: (2025)
Preference-Aligned LoRA Merging: Preserving Subspace Coverage and Addressing Directional Anisotropy
von: Jeong, Wooseong, et al.
Veröffentlicht: (2026)
von: Jeong, Wooseong, et al.
Veröffentlicht: (2026)
Latent Guard: a Safety Framework for Text-to-image Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
One-Step is Enough: Sparse Autoencoders for Text-to-Image Diffusion Models
von: Surkov, Viacheslav, et al.
Veröffentlicht: (2024)
von: Surkov, Viacheslav, et al.
Veröffentlicht: (2024)
Self-Correcting Self-Consuming Loops for Generative Model Training
von: Gillman, Nate, et al.
Veröffentlicht: (2024)
von: Gillman, Nate, et al.
Veröffentlicht: (2024)
Training-Free Consistent Text-to-Image Generation
von: Tewel, Yoad, et al.
Veröffentlicht: (2024)
von: Tewel, Yoad, et al.
Veröffentlicht: (2024)
How to Train your Text-to-Image Model: Evaluating Design Choices for Synthetic Training Captions
von: Brack, Manuel, et al.
Veröffentlicht: (2025)
von: Brack, Manuel, et al.
Veröffentlicht: (2025)
PCPO: Proportionate Credit Policy Optimization for Aligning Image Generation Models
von: Lee, Jeongjae, et al.
Veröffentlicht: (2025)
von: Lee, Jeongjae, et al.
Veröffentlicht: (2025)
SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation
von: Yoon, Jaehong, et al.
Veröffentlicht: (2024)
von: Yoon, Jaehong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
von: Wu, Junyi, et al.
Veröffentlicht: (2026) -
David and Goliath: Small One-step Model Beats Large Diffusion with Score Post-training
von: Luo, Weijian, et al.
Veröffentlicht: (2024) -
Uni-Instruct: One-step Diffusion Model through Unified Diffusion Divergence Instruction
von: Wang, Yifei, et al.
Veröffentlicht: (2025) -
Denoising Fisher Training For Neural Implicit Samplers
von: Luo, Weijian, et al.
Veröffentlicht: (2024) -
Exploring Text-to-Motion Generation with Human Preference
von: Sheng, Jenny, et al.
Veröffentlicht: (2024)