Preference-Guided Diffusion for Multi-Objective Offline Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Annadani, Yashas, Belakaria, Syrine, Ermon, Stefano, Bauer, Stefan, Engelhardt, Barbara E |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Non-Myopic Multi-Objective Bayesian Optimization
by: Belakaria, Syrine, et al.
Published: (2024)
by: Belakaria, Syrine, et al.
Published: (2024)
Sharpe Ratio-Guided Active Learning for Preference Optimization in RLHF
by: Belakaria, Syrine, et al.
Published: (2025)
by: Belakaria, Syrine, et al.
Published: (2025)
Active Learning for Derivative-Based Global Sensitivity Analysis with Gaussian Processes
by: Belakaria, Syrine, et al.
Published: (2024)
by: Belakaria, Syrine, et al.
Published: (2024)
Pareto Front-Diverse Batch Multi-Objective Bayesian Optimization
by: Ahmadianshalchi, Alaleh, et al.
Published: (2024)
by: Ahmadianshalchi, Alaleh, et al.
Published: (2024)
Sample Efficient Preference Alignment in LLMs via Active Exploration
by: Mehta, Viraj, et al.
Published: (2023)
by: Mehta, Viraj, et al.
Published: (2023)
Amortized Active Causal Induction with Deep Reinforcement Learning
by: Annadani, Yashas, et al.
Published: (2024)
by: Annadani, Yashas, et al.
Published: (2024)
Challenges and Considerations in the Evaluation of Bayesian Causal Discovery
by: Mamaghan, Amir Mohammad Karimi, et al.
Published: (2024)
by: Mamaghan, Amir Mohammad Karimi, et al.
Published: (2024)
Causal Bayesian Optimization with Unknown Graphs
by: Durand, Jean, et al.
Published: (2025)
by: Durand, Jean, et al.
Published: (2025)
Trust Your $\nabla$: Gradient-based Intervention Targeting for Causal Discovery
by: Olko, Mateusz, et al.
Published: (2022)
by: Olko, Mateusz, et al.
Published: (2022)
MADiff: Offline Multi-agent Learning with Diffusion Models
by: Zhu, Zhengbang, et al.
Published: (2023)
by: Zhu, Zhengbang, et al.
Published: (2023)
Pareto-Conditioned Diffusion Models for Offline Multi-Objective Optimization
by: Shrestha, Jatan, et al.
Published: (2026)
by: Shrestha, Jatan, et al.
Published: (2026)
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
Offline Multi-Objective Optimization
by: Xue, Ke, et al.
Published: (2024)
by: Xue, Ke, et al.
Published: (2024)
Preferred-Action-Optimized Diffusion Policies for Offline Reinforcement Learning
by: Zhang, Tianle, et al.
Published: (2024)
by: Zhang, Tianle, et al.
Published: (2024)
Robust Guided Diffusion for Offline Black-Box Optimization
by: Chen, Can Sam, et al.
Published: (2024)
by: Chen, Can Sam, et al.
Published: (2024)
Self-Refining Diffusion Samplers: Enabling Parallelization via Parareal Iterations
by: Selvam, Nikil Roashan, et al.
Published: (2024)
by: Selvam, Nikil Roashan, et al.
Published: (2024)
Beyond One-Preference-Fits-All Alignment: Multi-Objective Direct Preference Optimization
by: Zhou, Zhanhui, et al.
Published: (2023)
by: Zhou, Zhanhui, et al.
Published: (2023)
SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking
by: Cundy, Chris, et al.
Published: (2023)
by: Cundy, Chris, et al.
Published: (2023)
Latent Adversarial Regularization for Offline Preference Optimization
by: Jiang, Enyi, et al.
Published: (2026)
by: Jiang, Enyi, et al.
Published: (2026)
Learning Design-Score Manifold to Guide Diffusion Models for Offline Optimization
by: Zhou, Tailin, et al.
Published: (2025)
by: Zhou, Tailin, et al.
Published: (2025)
Guided Trajectory Generation with Diffusion Models for Offline Model-based Optimization
by: Yun, Taeyoung, et al.
Published: (2024)
by: Yun, Taeyoung, et al.
Published: (2024)
Meta-Learning Objectives for Preference Optimization
by: Alfano, Carlo, et al.
Published: (2024)
by: Alfano, Carlo, et al.
Published: (2024)
Interactive Hyperparameter Optimization in Multi-Objective Problems via Preference Learning
by: Giovanelli, Joseph, et al.
Published: (2023)
by: Giovanelli, Joseph, et al.
Published: (2023)
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
by: Rafailov, Rafael, et al.
Published: (2023)
by: Rafailov, Rafael, et al.
Published: (2023)
Meta-Aligner: Bidirectional Preference-Policy Optimization for Multi-Objective LLMs Alignment
by: Xu, Wenzhe, et al.
Published: (2026)
by: Xu, Wenzhe, et al.
Published: (2026)
An Offline Adaptation Framework for Constrained Multi-Objective Reinforcement Learning
by: Lin, Qian, et al.
Published: (2024)
by: Lin, Qian, et al.
Published: (2024)
Adversarial Policy Optimization for Offline Preference-based Reinforcement Learning
by: Kang, Hyungkyu, et al.
Published: (2025)
by: Kang, Hyungkyu, et al.
Published: (2025)
Generalized Preference Optimization: A Unified Approach to Offline Alignment
by: Tang, Yunhao, et al.
Published: (2024)
by: Tang, Yunhao, et al.
Published: (2024)
Preference Conditioned Multi-Objective Reinforcement Learning: Decomposed, Diversity-Driven Policy Optimization
by: Ambadkar, Tanmay, et al.
Published: (2026)
by: Ambadkar, Tanmay, et al.
Published: (2026)
Expensive Multi-Objective Bayesian Optimization Based on Diffusion Models
by: Li, Bingdong, et al.
Published: (2024)
by: Li, Bingdong, et al.
Published: (2024)
FairDICE: Fairness-Driven Offline Multi-Objective Reinforcement Learning
by: Kim, Woosung, et al.
Published: (2025)
by: Kim, Woosung, et al.
Published: (2025)
Inductive Moment Matching
by: Zhou, Linqi, et al.
Published: (2025)
by: Zhou, Linqi, et al.
Published: (2025)
Hindsight Preference Replay Improves Preference-Conditioned Multi-Objective Reinforcement Learning
by: Shianifar, Jonaid, et al.
Published: (2026)
by: Shianifar, Jonaid, et al.
Published: (2026)
The Principles of Diffusion Models
by: Lai, Chieh-Hsin, et al.
Published: (2025)
by: Lai, Chieh-Hsin, et al.
Published: (2025)
Collaborative Expert LLMs Guided Multi-Objective Molecular Optimization
by: Yu, Jiajun, et al.
Published: (2025)
by: Yu, Jiajun, et al.
Published: (2025)
GeoAda: Efficiently Finetune Geometric Diffusion Models with Equivariant Adapters
by: Zhao, Wanjia, et al.
Published: (2025)
by: Zhao, Wanjia, et al.
Published: (2025)
POLO: Preference-Guided Multi-Turn Reinforcement Learning for Lead Optimization
by: Wang, Ziqing, et al.
Published: (2025)
by: Wang, Ziqing, et al.
Published: (2025)
Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF
by: Cen, Shicong, et al.
Published: (2024)
by: Cen, Shicong, et al.
Published: (2024)
Preference Elicitation for Offline Reinforcement Learning
by: Pace, Alizée, et al.
Published: (2024)
by: Pace, Alizée, et al.
Published: (2024)
Hindsight Preference Learning for Offline Preference-based Reinforcement Learning
by: Gao, Chen-Xiao, et al.
Published: (2024)
by: Gao, Chen-Xiao, et al.
Published: (2024)
Similar Items
-
Non-Myopic Multi-Objective Bayesian Optimization
by: Belakaria, Syrine, et al.
Published: (2024) -
Sharpe Ratio-Guided Active Learning for Preference Optimization in RLHF
by: Belakaria, Syrine, et al.
Published: (2025) -
Active Learning for Derivative-Based Global Sensitivity Analysis with Gaussian Processes
by: Belakaria, Syrine, et al.
Published: (2024) -
Pareto Front-Diverse Batch Multi-Objective Bayesian Optimization
by: Ahmadianshalchi, Alaleh, et al.
Published: (2024) -
Sample Efficient Preference Alignment in LLMs via Active Exploration
by: Mehta, Viraj, et al.
Published: (2023)