Bridging SFT and DPO for Diffusion Model Alignment with Self-Sampling Preference Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Daoan, Lan, Guangchen, Han, Dong-Jun, Yao, Wenlin, Pan, Xiaoman, Zhang, Hongming, Li, Mingxiao, Chen, Pengcheng, Dong, Yu, Brinton, Christopher, Luo, Jiebo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Don't Forget your Inverse DDIM for Image Editing
von: Gomez-Trenado, Guillermo, et al.
Veröffentlicht: (2025)
von: Gomez-Trenado, Guillermo, et al.
Veröffentlicht: (2025)
Scalable Face Security Vision Foundation Model for Deepfake, Diffusion, and Spoofing Detection
von: Wang, Gaojian, et al.
Veröffentlicht: (2025)
von: Wang, Gaojian, et al.
Veröffentlicht: (2025)
Efficient Neural Network Encoding for 3D Color Lookup Tables
von: Zehtab, Vahid, et al.
Veröffentlicht: (2024)
von: Zehtab, Vahid, et al.
Veröffentlicht: (2024)
Once-For-All: A Train-Once and Select-Anytime Framework for Multimodal Instruction Tuning
von: Dong, Mingkang, et al.
Veröffentlicht: (2026)
von: Dong, Mingkang, et al.
Veröffentlicht: (2026)
NAAQA: A Neural Architecture for Acoustic Question Answering
von: Abdelnour, Jerome, et al.
Veröffentlicht: (2021)
von: Abdelnour, Jerome, et al.
Veröffentlicht: (2021)
Efficiently Scanning and Resampling Spatio-Temporal Tasks with Irregular Observations
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
PCA- and SVM-Grad-CAM for Convolutional Neural Networks: Closed-form Jacobian Expression
von: Omae, Yuto
Veröffentlicht: (2025)
von: Omae, Yuto
Veröffentlicht: (2025)
Carefully Structured Compression: Efficiently Managing StarCraft II Data
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
Better Schedules for Low Precision Training of Deep Neural Networks
von: Wolfe, Cameron R., et al.
Veröffentlicht: (2024)
von: Wolfe, Cameron R., et al.
Veröffentlicht: (2024)
Sample as You Infer: Predictive Coding With Langevin Dynamics
von: Zahid, Umais, et al.
Veröffentlicht: (2023)
von: Zahid, Umais, et al.
Veröffentlicht: (2023)
Explainable Classifier for Malignant Lymphoma Subtyping via Cell Graph and Image Fusion
von: Nishiyama, Daiki, et al.
Veröffentlicht: (2025)
von: Nishiyama, Daiki, et al.
Veröffentlicht: (2025)
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
Equip Pre-ranking with Target Attention by Residual Quantization
von: Li, Yutong, et al.
Veröffentlicht: (2025)
von: Li, Yutong, et al.
Veröffentlicht: (2025)
TSFool: Crafting Highly-Imperceptible Adversarial Time Series through Multi-Objective Attack
von: Wang, Yanyun, et al.
Veröffentlicht: (2022)
von: Wang, Yanyun, et al.
Veröffentlicht: (2022)
PlantDreamer: Achieving Realistic 3D Plant Models with Diffusion-Guided Gaussian Splatting
von: Hartley, Zane K J, et al.
Veröffentlicht: (2025)
von: Hartley, Zane K J, et al.
Veröffentlicht: (2025)
BlanketGen2-Fit3D: Synthetic Blanket Augmentation Towards Improving Real-World In-Bed Blanket Occluded Human Pose Estimation
von: Karácsony, Tamás, et al.
Veröffentlicht: (2025)
von: Karácsony, Tamás, et al.
Veröffentlicht: (2025)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
von: Chen, Pei-Chi, et al.
Veröffentlicht: (2025)
von: Chen, Pei-Chi, et al.
Veröffentlicht: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
Cost-Effective Attention Mechanisms for Low Resource Settings: Necessity & Sufficiency of Linear Transformations
von: Hosseini, Peyman, et al.
Veröffentlicht: (2024)
von: Hosseini, Peyman, et al.
Veröffentlicht: (2024)
MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
Adapting Multimodal Foundation Models for Few-Shot Learning: A Comprehensive Study on Contrastive Captioners
von: Narasinghe, N. K. B. M. P. K. B., et al.
Veröffentlicht: (2025)
von: Narasinghe, N. K. B. M. P. K. B., et al.
Veröffentlicht: (2025)
Tiny models from tiny data: Textual and null-text inversion for few-shot distillation
von: Landolsi, Erik, et al.
Veröffentlicht: (2024)
von: Landolsi, Erik, et al.
Veröffentlicht: (2024)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts
von: Tangtartharakul, Gene, et al.
Veröffentlicht: (2026)
von: Tangtartharakul, Gene, et al.
Veröffentlicht: (2026)
Statistical Analysis of the Impact of Quaternion Components in Convolutional Neural Networks
von: Altamirano-Gómez, Gerardo, et al.
Veröffentlicht: (2024)
von: Altamirano-Gómez, Gerardo, et al.
Veröffentlicht: (2024)
Developing Acoustic Models for Automatic Speech Recognition in Swedish
von: Salvi, Giampiero
Veröffentlicht: (2024)
von: Salvi, Giampiero
Veröffentlicht: (2024)
Interpreting Structured Perturbations in Image Protection Methods for Diffusion Models
von: Martin, Michael R., et al.
Veröffentlicht: (2025)
von: Martin, Michael R., et al.
Veröffentlicht: (2025)
Tensor Generalized Approximate Message Passing
von: Li, Yinchuan, et al.
Veröffentlicht: (2025)
von: Li, Yinchuan, et al.
Veröffentlicht: (2025)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
Mechanisms of Prompt-Induced Hallucination in Vision-Language Models
von: Rudman, William, et al.
Veröffentlicht: (2026)
von: Rudman, William, et al.
Veröffentlicht: (2026)
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
von: Bian, Zhipeng, et al.
Veröffentlicht: (2026)
von: Bian, Zhipeng, et al.
Veröffentlicht: (2026)
T-Norm Operators for EU AI Act Compliance Classification: An Empirical Comparison of Lukasiewicz, Product, and Gödel Semantics in a Neuro-Symbolic Reasoning System
von: Laabs, Adam
Veröffentlicht: (2026)
von: Laabs, Adam
Veröffentlicht: (2026)
KGTN-ens: Few-Shot Image Classification with Knowledge Graph Ensembles
von: Filipiak, Dominik, et al.
Veröffentlicht: (2022)
von: Filipiak, Dominik, et al.
Veröffentlicht: (2022)
A Survey on Dynamic Neural Networks: from Computer Vision to Multi-modal Sensor Fusion
von: Montello, Fabio, et al.
Veröffentlicht: (2025)
von: Montello, Fabio, et al.
Veröffentlicht: (2025)
WaveMix: A Resource-efficient Neural Network for Image Analysis
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Proto-FG3D: Prototype-based Interpretable Fine-Grained 3D Shape Classification
von: Ma, Shuxian, et al.
Veröffentlicht: (2025)
von: Ma, Shuxian, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for Causal Modeling
von: Razouk, Houssam, et al.
Veröffentlicht: (2024)
von: Razouk, Houssam, et al.
Veröffentlicht: (2024)
Augmenting Replay in World Models for Continual Reinforcement Learning
von: Yang, Luke, et al.
Veröffentlicht: (2024)
von: Yang, Luke, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Don't Forget your Inverse DDIM for Image Editing
von: Gomez-Trenado, Guillermo, et al.
Veröffentlicht: (2025) -
Scalable Face Security Vision Foundation Model for Deepfake, Diffusion, and Spoofing Detection
von: Wang, Gaojian, et al.
Veröffentlicht: (2025) -
Efficient Neural Network Encoding for 3D Color Lookup Tables
von: Zehtab, Vahid, et al.
Veröffentlicht: (2024) -
Once-For-All: A Train-Once and Select-Anytime Framework for Multimodal Instruction Tuning
von: Dong, Mingkang, et al.
Veröffentlicht: (2026) -
NAAQA: A Neural Architecture for Acoustic Question Answering
von: Abdelnour, Jerome, et al.
Veröffentlicht: (2021)