Preference Alignment on Diffusion Model: A Comprehensive Survey for Image Generation and Editing
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Sihao, Si, Xiaonan, Xing, Chi, Wang, Jianhong, Jin, Gaojie, Cheng, Guangliang, Zhang, Lijun, Huang, Xiaowei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Activation Steering Meets Preference Optimization: Defense Against Jailbreaks in Vision Language Models
di: Wu, Sihao, et al.
Pubblicazione: (2025)
di: Wu, Sihao, et al.
Pubblicazione: (2025)
Diffusion Model-Based Image Editing: A Survey
di: Huang, Yi, et al.
Pubblicazione: (2024)
di: Huang, Yi, et al.
Pubblicazione: (2024)
Optimising Event-Driven Spiking Neural Network with Regularisation and Cutoff
di: Wu, Dengyu, et al.
Pubblicazione: (2023)
di: Wu, Dengyu, et al.
Pubblicazione: (2023)
Diffusion Models for Image Restoration and Enhancement: A Comprehensive Survey
di: Li, Xin, et al.
Pubblicazione: (2023)
di: Li, Xin, et al.
Pubblicazione: (2023)
Instant Preference Alignment for Text-to-Image Diffusion Models
di: Li, Yang, et al.
Pubblicazione: (2025)
di: Li, Yang, et al.
Pubblicazione: (2025)
Free Lunch Alignment of Text-to-Image Diffusion Models without Preference Image Pairs
di: Xian, Jia Jun Cheng, et al.
Pubblicazione: (2025)
di: Xian, Jia Jun Cheng, et al.
Pubblicazione: (2025)
MagDiff: Multi-Alignment Diffusion for High-Fidelity Video Generation and Editing
di: Zhao, Haoyu, et al.
Pubblicazione: (2023)
di: Zhao, Haoyu, et al.
Pubblicazione: (2023)
SIDA: Social Media Image Deepfake Detection, Localization and Explanation with Large Multimodal Model
di: Huang, Zhenglin, et al.
Pubblicazione: (2024)
di: Huang, Zhenglin, et al.
Pubblicazione: (2024)
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
di: Shuai, Xincheng, et al.
Pubblicazione: (2024)
di: Shuai, Xincheng, et al.
Pubblicazione: (2024)
Instruction-Oriented Preference Alignment for Enhancing Multi-Modal Comprehension Capability of MLLMs
di: Wang, Zitian, et al.
Pubblicazione: (2025)
di: Wang, Zitian, et al.
Pubblicazione: (2025)
Rethinking Cross-Generator Image Forgery Detection through DINOv3
di: Huang, Zhenglin, et al.
Pubblicazione: (2025)
di: Huang, Zhenglin, et al.
Pubblicazione: (2025)
LayerDiffusion: Layered Controlled Image Editing with Diffusion Models
di: Li, Pengzhi, et al.
Pubblicazione: (2023)
di: Li, Pengzhi, et al.
Pubblicazione: (2023)
Diffusion Blend: Inference-Time Multi-Preference Alignment for Diffusion Models
di: Cheng, Min, et al.
Pubblicazione: (2025)
di: Cheng, Min, et al.
Pubblicazione: (2025)
Unveiling Deep Shadows: A Survey and Benchmark on Image and Video Shadow Detection, Removal, and Generation in the Deep Learning Era
di: Hu, Xiaowei, et al.
Pubblicazione: (2024)
di: Hu, Xiaowei, et al.
Pubblicazione: (2024)
Towards General Preference Alignment: Diffusion Models at Nash Equilibrium
di: Hu, Jiaming, et al.
Pubblicazione: (2026)
di: Hu, Jiaming, et al.
Pubblicazione: (2026)
Spatial-DISE: A Unified Benchmark for Evaluating Spatial Reasoning in Vision-Language Models
di: Huang, Xinmiao, et al.
Pubblicazione: (2025)
di: Huang, Xinmiao, et al.
Pubblicazione: (2025)
Personalized Image Editing in Text-to-Image Diffusion Models via Collaborative Direct Preference Optimization
di: Dunlop, Connor, et al.
Pubblicazione: (2025)
di: Dunlop, Connor, et al.
Pubblicazione: (2025)
Vision Mamba in Remote Sensing: A Comprehensive Survey of Techniques, Applications and Outlook
di: Bao, Muyi, et al.
Pubblicazione: (2025)
di: Bao, Muyi, et al.
Pubblicazione: (2025)
Text to Image Generation and Editing: A Survey
di: Yang, Pengfei, et al.
Pubblicazione: (2025)
di: Yang, Pengfei, et al.
Pubblicazione: (2025)
BEARD: Benchmarking the Adversarial Robustness for Dataset Distillation
di: Zhou, Zheng, et al.
Pubblicazione: (2024)
di: Zhou, Zheng, et al.
Pubblicazione: (2024)
Integrating Object Detection Modality into Visual Language Model for Enhanced Autonomous Driving Agent
di: He, Linfeng, et al.
Pubblicazione: (2024)
di: He, Linfeng, et al.
Pubblicazione: (2024)
SonicDiffusion: Audio-Driven Image Generation and Editing with Pretrained Diffusion Models
di: Biner, Burak Can, et al.
Pubblicazione: (2024)
di: Biner, Burak Can, et al.
Pubblicazione: (2024)
RPiAE: A Representation-Pivoted Autoencoder Enhancing Both Image Generation and Editing
di: Gong, Yue, et al.
Pubblicazione: (2026)
di: Gong, Yue, et al.
Pubblicazione: (2026)
Out-of-Bounding-Box Triggers: A Stealthy Approach to Cheat Object Detectors
di: Lin, Tao, et al.
Pubblicazione: (2024)
di: Lin, Tao, et al.
Pubblicazione: (2024)
Light-T2M: A Lightweight and Fast Model for Text-to-motion Generation
di: Zeng, Ling-An, et al.
Pubblicazione: (2024)
di: Zeng, Ling-An, et al.
Pubblicazione: (2024)
Image Editing As Programs with Diffusion Models
di: Hu, Yujia, et al.
Pubblicazione: (2025)
di: Hu, Yujia, et al.
Pubblicazione: (2025)
SUDO: Enhancing Text-to-Image Diffusion Models with Self-Supervised Direct Preference Optimization
di: Peng, Liang, et al.
Pubblicazione: (2025)
di: Peng, Liang, et al.
Pubblicazione: (2025)
Re-Align: Structured Reasoning-guided Alignment for In-Context Image Generation and Editing
di: He, Runze, et al.
Pubblicazione: (2026)
di: He, Runze, et al.
Pubblicazione: (2026)
I2VEdit: First-Frame-Guided Video Editing via Image-to-Video Diffusion Models
di: Ouyang, Wenqi, et al.
Pubblicazione: (2024)
di: Ouyang, Wenqi, et al.
Pubblicazione: (2024)
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers
di: Zhu, Jingyuan, et al.
Pubblicazione: (2026)
di: Zhu, Jingyuan, et al.
Pubblicazione: (2026)
Deep Learning Based Domain Adaptation Methods in Remote Sensing: A Comprehensive Survey
di: Lyu, Shuchang, et al.
Pubblicazione: (2025)
di: Lyu, Shuchang, et al.
Pubblicazione: (2025)
Rethinking Preference Alignment for Diffusion Models with Classifier-Free Guidance
di: Jiang, Zhou, et al.
Pubblicazione: (2026)
di: Jiang, Zhou, et al.
Pubblicazione: (2026)
Diffusion Model-Based Video Editing: A Survey
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
LLMs Meet Multimodal Generation and Editing: A Survey
di: He, Yingqing, et al.
Pubblicazione: (2024)
di: He, Yingqing, et al.
Pubblicazione: (2024)
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
di: Liu, Runtao, et al.
Pubblicazione: (2024)
di: Liu, Runtao, et al.
Pubblicazione: (2024)
Consistent Image Layout Editing with Diffusion Models
di: Xia, Tao, et al.
Pubblicazione: (2025)
di: Xia, Tao, et al.
Pubblicazione: (2025)
Taming Preference Mode Collapse via Directional Decoupling Alignment in Diffusion Reinforcement Learning
di: Chen, Chubin, et al.
Pubblicazione: (2025)
di: Chen, Chubin, et al.
Pubblicazione: (2025)
A Comprehensive Survey on Diffusion Models and Their Applications
di: Ahsan, Md Manjurul, et al.
Pubblicazione: (2024)
di: Ahsan, Md Manjurul, et al.
Pubblicazione: (2024)
GeoDiffuser: Geometry-Based Image Editing with Diffusion Models
di: Sajnani, Rahul, et al.
Pubblicazione: (2024)
di: Sajnani, Rahul, et al.
Pubblicazione: (2024)
Divergence Minimization Preference Optimization for Diffusion Model Alignment
di: Li, Binxu, et al.
Pubblicazione: (2025)
di: Li, Binxu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Activation Steering Meets Preference Optimization: Defense Against Jailbreaks in Vision Language Models
di: Wu, Sihao, et al.
Pubblicazione: (2025) -
Diffusion Model-Based Image Editing: A Survey
di: Huang, Yi, et al.
Pubblicazione: (2024) -
Optimising Event-Driven Spiking Neural Network with Regularisation and Cutoff
di: Wu, Dengyu, et al.
Pubblicazione: (2023) -
Diffusion Models for Image Restoration and Enhancement: A Comprehensive Survey
di: Li, Xin, et al.
Pubblicazione: (2023) -
Instant Preference Alignment for Text-to-Image Diffusion Models
di: Li, Yang, et al.
Pubblicazione: (2025)