Curriculum Direct Preference Optimization for Diffusion and Consistency Models
Fuente:
arXiv
Saved in:
| Main Authors: | Croitoru, Florinel-Alin, Hondru, Vlad, Ionescu, Radu Tudor, Sebe, Nicu, Shah, Mubarak |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Curriculum-DPO++: Direct Preference Optimization via Data and Model Curricula for Text-to-Image Generation
by: Croitoru, Florinel-Alin, et al.
Published: (2026)
by: Croitoru, Florinel-Alin, et al.
Published: (2026)
Diffusion Models in Vision: A Survey
by: Croitoru, Florinel-Alin, et al.
Published: (2022)
by: Croitoru, Florinel-Alin, et al.
Published: (2022)
Masked Image Modeling: A Survey
by: Hondru, Vlad, et al.
Published: (2024)
by: Hondru, Vlad, et al.
Published: (2024)
Reverse Stable Diffusion: What prompt was used to generate this image?
by: Croitoru, Florinel-Alin, et al.
Published: (2023)
by: Croitoru, Florinel-Alin, et al.
Published: (2023)
PRNU-Bench: A Novel Benchmark and Model for PRNU-Based Camera Identification
by: Croitoru, Florinel Alin, et al.
Published: (2025)
by: Croitoru, Florinel Alin, et al.
Published: (2025)
MAVOS-DD: Multilingual Audio-Video Open-Set Deepfake Detection Benchmark
by: Croitoru, Florinel-Alin, et al.
Published: (2025)
by: Croitoru, Florinel-Alin, et al.
Published: (2025)
CBM: Curriculum by Masking
by: Jarca, Andrei, et al.
Published: (2024)
by: Jarca, Andrei, et al.
Published: (2024)
Learning Rate Curriculum
by: Croitoru, Florinel-Alin, et al.
Published: (2022)
by: Croitoru, Florinel-Alin, et al.
Published: (2022)
MTL-MAD: Multi-Task Learners are Effective Medical Anomaly Detectors
by: Bercean, Bogdan Alexandru, et al.
Published: (2026)
by: Bercean, Bogdan Alexandru, et al.
Published: (2026)
Lightning Fast Video Anomaly Detection via Adversarial Knowledge Distillation
by: Croitoru, Florinel-Alin, et al.
Published: (2022)
by: Croitoru, Florinel-Alin, et al.
Published: (2022)
Towards Few-Call Model Stealing via Active Self-Paced Knowledge Distillation and Diffusion-Based Image Generation
by: Hondru, Vlad, et al.
Published: (2023)
by: Hondru, Vlad, et al.
Published: (2023)
Task-Informed Anti-Curriculum by Masking Improves Downstream Performance on Text
by: Jarca, Andrei, et al.
Published: (2025)
by: Jarca, Andrei, et al.
Published: (2025)
Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook
by: Croitoru, Florinel-Alin, et al.
Published: (2024)
by: Croitoru, Florinel-Alin, et al.
Published: (2024)
DLCR: A Generative Data Expansion Framework via Diffusion for Clothes-Changing Person Re-ID
by: Siddiqui, Nyle, et al.
Published: (2024)
by: Siddiqui, Nyle, et al.
Published: (2024)
ExDDV: A New Dataset for Explainable Deepfake Detection in Video
by: Hondru, Vlad, et al.
Published: (2025)
by: Hondru, Vlad, et al.
Published: (2025)
Self-Distilled Masked Auto-Encoders are Efficient Video Anomaly Detectors
by: Ristea, Nicolae-Catalin, et al.
Published: (2023)
by: Ristea, Nicolae-Catalin, et al.
Published: (2023)
Curriculum Multi-Task Self-Supervision Improves Lightweight Architectures for Onboard Satellite Hyperspectral Image Segmentation
by: Carlesso, Hugo, et al.
Published: (2025)
by: Carlesso, Hugo, et al.
Published: (2025)
Machine Unlearning in the Era of Quantum Machine Learning: An Empirical Study
by: Crivoi, Carla, et al.
Published: (2025)
by: Crivoi, Carla, et al.
Published: (2025)
CL-MAE: Curriculum-Learned Masked Autoencoders
by: Madan, Neelu, et al.
Published: (2023)
by: Madan, Neelu, et al.
Published: (2023)
Multi-Level Feature Distillation of Joint Teachers Trained on Distinct Image Datasets
by: Iordache, Adrian, et al.
Published: (2024)
by: Iordache, Adrian, et al.
Published: (2024)
Hyperbolic Busemann Neural Networks
by: Chen, Ziheng, et al.
Published: (2026)
by: Chen, Ziheng, et al.
Published: (2026)
GeMix: Conditional GAN-Based Mixup for Improved Medical Image Augmentation
by: Carlesso, Hugo, et al.
Published: (2025)
by: Carlesso, Hugo, et al.
Published: (2025)
Weight Copy and Low-Rank Adaptation for Few-Shot Distillation of Vision Transformers
by: Grigore, Diana-Nicoleta, et al.
Published: (2024)
by: Grigore, Diana-Nicoleta, et al.
Published: (2024)
POCI-Diff: Position Objects Consistently and Interactively with 3D-Layout Guided Diffusion
by: Rigo, Andrea, et al.
Published: (2026)
by: Rigo, Andrea, et al.
Published: (2026)
ESPLoRA: Enhanced Spatial Precision with Low-Rank Adaption in Text-to-Image Diffusion Models for High-Definition Synthesis
by: Rigo, Andrea, et al.
Published: (2025)
by: Rigo, Andrea, et al.
Published: (2025)
PQPP: A Joint Benchmark for Text-to-Image Prompt and Query Performance Prediction
by: Poesina, Eduard, et al.
Published: (2024)
by: Poesina, Eduard, et al.
Published: (2024)
CLewR: Curriculum Learning with Restarts for Machine Translation Preference Learning
by: Dragomir, Alexandra, et al.
Published: (2026)
by: Dragomir, Alexandra, et al.
Published: (2026)
Uncertainty-Aware Testing-Time Optimization for 3D Human Pose Estimation
by: Wang, Ti, et al.
Published: (2024)
by: Wang, Ti, et al.
Published: (2024)
PAIR-Diffusion: A Comprehensive Multimodal Object-Level Image Editor
by: Goel, Vidit, et al.
Published: (2023)
by: Goel, Vidit, et al.
Published: (2023)
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
by: Grigore, Diana-Nicoleta, et al.
Published: (2025)
by: Grigore, Diana-Nicoleta, et al.
Published: (2025)
H$_{2}$OT: Hierarchical Hourglass Tokenizer for Efficient Video Pose Transformers
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
by: Li, Wenhao, et al.
Published: (2023)
by: Li, Wenhao, et al.
Published: (2023)
Surgical Triplet Recognition via Diffusion Model
by: Liu, Daochang, et al.
Published: (2024)
by: Liu, Daochang, et al.
Published: (2024)
CE-SDWV: Effective and Efficient Concept Erasure for Text-to-Image Diffusion Models via a Semantic-Driven Word Vocabulary
by: Tu, Jiahang, et al.
Published: (2025)
by: Tu, Jiahang, et al.
Published: (2025)
GraphMLP: A Graph MLP-Like Architecture for 3D Human Pose Estimation
by: Li, Wenhao, et al.
Published: (2022)
by: Li, Wenhao, et al.
Published: (2022)
Boost Your Human Image Generation Model via Direct Preference Optimization
by: Na, Sanghyeon, et al.
Published: (2024)
by: Na, Sanghyeon, et al.
Published: (2024)
Diffexplainer: Towards Cross-modal Global Explanations with Diffusion Models
by: Pennisi, Matteo, et al.
Published: (2024)
by: Pennisi, Matteo, et al.
Published: (2024)
Rip Current Segmentation: A Novel Benchmark and YOLOv8 Baseline Results
by: Dumitriu, Andrei, et al.
Published: (2025)
by: Dumitriu, Andrei, et al.
Published: (2025)
PoInit-of-View: Poisoning Initialization of Views Transfers Across Multiple 3D Reconstruction Systems
by: Wang, Weijie, et al.
Published: (2026)
by: Wang, Weijie, et al.
Published: (2026)
SafeR-CLIP: Mitigating NSFW Content in Vision-Language Models While Preserving Pre-Trained Knowledge
by: Yousaf, Adeel, et al.
Published: (2025)
by: Yousaf, Adeel, et al.
Published: (2025)
Similar Items
-
Curriculum-DPO++: Direct Preference Optimization via Data and Model Curricula for Text-to-Image Generation
by: Croitoru, Florinel-Alin, et al.
Published: (2026) -
Diffusion Models in Vision: A Survey
by: Croitoru, Florinel-Alin, et al.
Published: (2022) -
Masked Image Modeling: A Survey
by: Hondru, Vlad, et al.
Published: (2024) -
Reverse Stable Diffusion: What prompt was used to generate this image?
by: Croitoru, Florinel-Alin, et al.
Published: (2023) -
PRNU-Bench: A Novel Benchmark and Model for PRNU-Based Camera Identification
by: Croitoru, Florinel Alin, et al.
Published: (2025)