Di3PO - Diptych Diffusion DPO for Targeted Improvements in Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Reddy, Sanjana, Malhi, Ishaan, Ma, Sally, Dutta, Praneet |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Preserving Product Fidelity in Large Scale Image Recontextualization with Diffusion Models
von: Malhi, Ishaan, et al.
Veröffentlicht: (2025)
von: Malhi, Ishaan, et al.
Veröffentlicht: (2025)
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models
von: Li, Zejian, et al.
Veröffentlicht: (2025)
von: Li, Zejian, et al.
Veröffentlicht: (2025)
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
DiMSUM: Diffusion Mamba -- A Scalable and Unified Spatial-Frequency Method for Image Generation
von: Phung, Hao, et al.
Veröffentlicht: (2024)
von: Phung, Hao, et al.
Veröffentlicht: (2024)
Remix-DiT: Mixing Diffusion Transformers for Multi-Expert Denoising
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Interactive Video Generation via Domain Adaptation
von: Rawal, Ishaan, et al.
Veröffentlicht: (2025)
von: Rawal, Ishaan, et al.
Veröffentlicht: (2025)
When Preferences Diverge: Aligning Diffusion Models with Minority-Aware Adaptive DPO
von: Zhang, Lingfan, et al.
Veröffentlicht: (2025)
von: Zhang, Lingfan, et al.
Veröffentlicht: (2025)
FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus
von: Jin, Qiaoqiao, et al.
Veröffentlicht: (2025)
von: Jin, Qiaoqiao, et al.
Veröffentlicht: (2025)
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers
von: Chen, Lei, et al.
Veröffentlicht: (2024)
von: Chen, Lei, et al.
Veröffentlicht: (2024)
DiVa-360: The Dynamic Visual Dataset for Immersive Neural Fields
von: Lu, Cheng-You, et al.
Veröffentlicht: (2023)
von: Lu, Cheng-You, et al.
Veröffentlicht: (2023)
LaDiC: Are Diffusion Models Really Inferior to Autoregressive Counterparts for Image-to-Text Generation?
von: Wang, Yuchi, et al.
Veröffentlicht: (2024)
von: Wang, Yuchi, et al.
Veröffentlicht: (2024)
RealDPO: Real or Not Real, that is the Preference
von: Cheng, Guo, et al.
Veröffentlicht: (2025)
von: Cheng, Guo, et al.
Veröffentlicht: (2025)
CoDi: Conditional Diffusion Distillation for Higher-Fidelity and Faster Image Generation
von: Mei, Kangfu, et al.
Veröffentlicht: (2023)
von: Mei, Kangfu, et al.
Veröffentlicht: (2023)
MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
BalancedDPO: Adaptive Multi-Metric Alignment
von: Tamboli, Dipesh, et al.
Veröffentlicht: (2025)
von: Tamboli, Dipesh, et al.
Veröffentlicht: (2025)
Longitudinal Boundary Sharpness Coefficient Slopes Predict Time to Alzheimer's Disease Conversion in Mild Cognitive Impairment: A Survival Analysis Using the ADNI Cohort
von: Cherukuri, Ishaan
Veröffentlicht: (2026)
von: Cherukuri, Ishaan
Veröffentlicht: (2026)
MMFace-DiT: A Dual-Stream Diffusion Transformer for High-Fidelity Multimodal Face Generation
von: Krishnamurthy, Bharath, et al.
Veröffentlicht: (2026)
von: Krishnamurthy, Bharath, et al.
Veröffentlicht: (2026)
Generation of Heterogeneous PET Images from Uniform Organ Activity Maps Using a Pretrained Domain-Adapted Diffusion Model
von: Li, Suya, et al.
Veröffentlicht: (2026)
von: Li, Suya, et al.
Veröffentlicht: (2026)
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2025)
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2025)
Reg-DPO: SFT-Regularized Direct Preference Optimization with GT-Pair for Improving Video Generation
von: Du, Jie, et al.
Veröffentlicht: (2025)
von: Du, Jie, et al.
Veröffentlicht: (2025)
DPO: Dual-Perturbation Optimization for Test-time Adaptation in 3D Object Detection
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2024)
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2024)
Curriculum-DPO++: Direct Preference Optimization via Data and Model Curricula for Text-to-Image Generation
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2026)
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2026)
OmniDiT: Extending Diffusion Transformer to Omni-VTON Framework
von: Zeng, Weixuan, et al.
Veröffentlicht: (2026)
von: Zeng, Weixuan, et al.
Veröffentlicht: (2026)
Delving into RL for Image Generation with CoT: A Study on DPO vs. GRPO
von: Tong, Chengzhuo, et al.
Veröffentlicht: (2025)
von: Tong, Chengzhuo, et al.
Veröffentlicht: (2025)
DPO Learning with LLMs-Judge Signal for Computer Use Agents
von: Luo, Man, et al.
Veröffentlicht: (2025)
von: Luo, Man, et al.
Veröffentlicht: (2025)
HQ-DiT: Efficient Diffusion Transformer with FP4 Hybrid Quantization
von: Liu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Liu, Wenxuan, et al.
Veröffentlicht: (2024)
DiG: Scalable and Efficient Diffusion Models with Gated Linear Attention
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion Transformers
von: Byung-Ki, Kwon, et al.
Veröffentlicht: (2025)
von: Byung-Ki, Kwon, et al.
Veröffentlicht: (2025)
Unveiling Redundancy in Diffusion Transformers (DiTs): A Systematic Study
von: Sun, Xibo, et al.
Veröffentlicht: (2024)
von: Sun, Xibo, et al.
Veröffentlicht: (2024)
DiT-VTON: Diffusion Transformer Framework for Unified Multi-Category Virtual Try-On and Virtual Try-All with Integrated Image Editing
von: Li, Qi, et al.
Veröffentlicht: (2025)
von: Li, Qi, et al.
Veröffentlicht: (2025)
Paired Image Generation with Diffusion-Guided Diffusion Models
von: Zhang, Haoxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Haoxuan, et al.
Veröffentlicht: (2025)
DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
Wonder3D++: Cross-domain Diffusion for High-fidelity 3D Generation from a Single Image
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
City-Mesh3R: Simulation-Ready City-Scale 3D Mesh Reconstruction from Multi-View Images
von: Paul, Sayan, et al.
Veröffentlicht: (2026)
von: Paul, Sayan, et al.
Veröffentlicht: (2026)
LuxDiT: Lighting Estimation with Video Diffusion Transformer
von: Liang, Ruofan, et al.
Veröffentlicht: (2025)
von: Liang, Ruofan, et al.
Veröffentlicht: (2025)
VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving
von: Xu, Zhefan, et al.
Veröffentlicht: (2026)
von: Xu, Zhefan, et al.
Veröffentlicht: (2026)
Advanced Underwater Image Quality Enhancement via Hybrid Super-Resolution Convolutional Neural Networks and Multi-Scale Retinex-Based Defogging Techniques
von: Gogireddy, Yugandhar Reddy, et al.
Veröffentlicht: (2024)
von: Gogireddy, Yugandhar Reddy, et al.
Veröffentlicht: (2024)
Unified Text-Image Generation with Weakness-Targeted Post-Training
von: Chen, Jiahui, et al.
Veröffentlicht: (2026)
von: Chen, Jiahui, et al.
Veröffentlicht: (2026)
CCIS-Diff: A Generative Model with Stable Diffusion Prior for Controlled Colonoscopy Image Synthesis
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Preserving Product Fidelity in Large Scale Image Recontextualization with Diffusion Models
von: Malhi, Ishaan, et al.
Veröffentlicht: (2025) -
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026) -
Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models
von: Li, Zejian, et al.
Veröffentlicht: (2025) -
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024) -
DiMSUM: Diffusion Mamba -- A Scalable and Unified Spatial-Frequency Method for Image Generation
von: Phung, Hao, et al.
Veröffentlicht: (2024)