COS-DPO: Conditioned One-Shot Multi-Objective Fine-Tuning Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Yinuo, Xiao, Tesi, Shavlovsky, Michael, Ying, Lexing, Rahmanian, Holakou |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Objective Optimization via Wasserstein-Fisher-Rao Gradient Flow
by: Ren, Yinuo, et al.
Published: (2023)
by: Ren, Yinuo, et al.
Published: (2023)
A Sinkhorn-type Algorithm for Constrained Optimal Transport
by: Tang, Xun, et al.
Published: (2024)
by: Tang, Xun, et al.
Published: (2024)
Accelerating Sinkhorn Algorithm with Sparse Newton Iterations
by: Tang, Xun, et al.
Published: (2024)
by: Tang, Xun, et al.
Published: (2024)
An efficient algorithm for entropic optimal transport under martingale-type constraints
by: Tang, Xun, et al.
Published: (2025)
by: Tang, Xun, et al.
Published: (2025)
A note on continuous-time online learning
by: Ying, Lexing
Published: (2024)
by: Ying, Lexing
Published: (2024)
One-Shot Safety Alignment for Large Language Models via Optimal Dualization
by: Huang, Xinmeng, et al.
Published: (2024)
by: Huang, Xinmeng, et al.
Published: (2024)
Variance-reduced Zeroth-Order Methods for Fine-Tuning Language Models
by: Gautam, Tanmay, et al.
Published: (2024)
by: Gautam, Tanmay, et al.
Published: (2024)
FERERO: A Flexible Framework for Preference-Guided Multi-Objective Learning
by: Chen, Lisha, et al.
Published: (2024)
by: Chen, Lisha, et al.
Published: (2024)
LISA: Layerwise Importance Sampling for Memory-Efficient Large Language Model Fine-Tuning
by: Pan, Rui, et al.
Published: (2024)
by: Pan, Rui, et al.
Published: (2024)
Understanding Forgetting in LLM Supervised Fine-Tuning and Preference Learning -- A Convex Optimization Perspective
by: Fernando, Heshan, et al.
Published: (2024)
by: Fernando, Heshan, et al.
Published: (2024)
Aligned Multi Objective Optimization
by: Efroni, Yonathan, et al.
Published: (2025)
by: Efroni, Yonathan, et al.
Published: (2025)
Locally Adaptive Multi-Objective Learning
by: Kaur, Jivat Neet, et al.
Published: (2026)
by: Kaur, Jivat Neet, et al.
Published: (2026)
Distributionally Robust Multi-Objective Optimization
by: Yang, Yufeng, et al.
Published: (2026)
by: Yang, Yufeng, et al.
Published: (2026)
Preferential Multi-Objective Bayesian Optimization
by: Astudillo, Raul, et al.
Published: (2024)
by: Astudillo, Raul, et al.
Published: (2024)
Efficient Tail-Aware Generative Optimization via Flow Model Fine-Tuning
by: Wang, Zifan, et al.
Published: (2026)
by: Wang, Zifan, et al.
Published: (2026)
Control Theoretic Approach to Fine-Tuning and Transfer Learning
by: Bayram, Erkan, et al.
Published: (2024)
by: Bayram, Erkan, et al.
Published: (2024)
Secure LLM Fine-Tuning via Safety-Aware Probing
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
Efficient First-Order Optimization on the Pareto Set for Multi-Objective Learning under Preference Guidance
by: Chen, Lisha, et al.
Published: (2025)
by: Chen, Lisha, et al.
Published: (2025)
MOSS: Multi-Objective Optimization for Stable Rule Sets
by: Liu, Brian, et al.
Published: (2025)
by: Liu, Brian, et al.
Published: (2025)
A Unified Understanding of Offline Data Selection and Online Self-refining Generation for Post-training LLMs
by: Xiao, Quan, et al.
Published: (2025)
by: Xiao, Quan, et al.
Published: (2025)
A Unified Framework for Gradient Aggregation in Multi-Objective Optimization
by: Hu, Zeou, et al.
Published: (2026)
by: Hu, Zeou, et al.
Published: (2026)
Knowledge Gradient for Multi-Objective Bayesian Optimization with Decoupled Evaluations
by: Buckingham, Jack M., et al.
Published: (2023)
by: Buckingham, Jack M., et al.
Published: (2023)
Zero-Order Optimization for LLM Fine-Tuning via Learnable Direction Sampling
by: Parfenov, Valery, et al.
Published: (2026)
by: Parfenov, Valery, et al.
Published: (2026)
LoFT: Low-Rank Adaptation That Behaves Like Full Fine-Tuning
by: Tastan, Nurbek, et al.
Published: (2025)
by: Tastan, Nurbek, et al.
Published: (2025)
Collaborative Pareto Set Learning in Multiple Multi-Objective Optimization Problems
by: Shang, Chikai, et al.
Published: (2024)
by: Shang, Chikai, et al.
Published: (2024)
Multi-Objective Optimization-Based Anonymization of Structured Data for Machine Learning Application
by: Wei, Yusi, et al.
Published: (2025)
by: Wei, Yusi, et al.
Published: (2025)
Enhancing Multi-Objective Optimization through Machine Learning-Supported Multiphysics Simulation
by: Botache, Diego, et al.
Published: (2023)
by: Botache, Diego, et al.
Published: (2023)
Chebyshev Center-Based Direction Selection for Multi-Objective Optimization and Training PINNs
by: Yoon, Hoyeol, et al.
Published: (2026)
by: Yoon, Hoyeol, et al.
Published: (2026)
Pareto Front Shape-Agnostic Pareto Set Learning in Multi-Objective Optimization
by: Ye, Rongguang, et al.
Published: (2024)
by: Ye, Rongguang, et al.
Published: (2024)
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
by: Qiu, Shuang, et al.
Published: (2024)
by: Qiu, Shuang, et al.
Published: (2024)
New Hybrid Fine-Tuning Paradigm for LLMs: Algorithm Design and Convergence Analysis Framework
by: Ma, Shaocong, et al.
Published: (2026)
by: Ma, Shaocong, et al.
Published: (2026)
Multi-Objective Linear Ensembles for Robust and Sparse Training of Few-Bit Neural Networks
by: Bernardelli, Ambrogio Maria, et al.
Published: (2022)
by: Bernardelli, Ambrogio Maria, et al.
Published: (2022)
Gradient-Informed Monte Carlo Fine-Tuning of Diffusion Models for Low-Thrust Trajectory Design
by: Graebner, Jannik, et al.
Published: (2025)
by: Graebner, Jannik, et al.
Published: (2025)
Standardization of Multi-Objective QUBOs
by: Lee, Loong Kuan, et al.
Published: (2025)
by: Lee, Loong Kuan, et al.
Published: (2025)
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
by: Zeng, Sihan, et al.
Published: (2026)
by: Zeng, Sihan, et al.
Published: (2026)
On the Global Convergence of Risk-Averse Natural Policy Gradient Methods with Expected Conditional Risk Measures
by: Yu, Xian, et al.
Published: (2023)
by: Yu, Xian, et al.
Published: (2023)
Fine-grained Analysis of In-context Linear Estimation: Data, Architecture, and Beyond
by: Li, Yingcong, et al.
Published: (2024)
by: Li, Yingcong, et al.
Published: (2024)
Jacobian Descent for Multi-Objective Optimization
by: Quinton, Pierre, et al.
Published: (2024)
by: Quinton, Pierre, et al.
Published: (2024)
A Global Optimization Algorithm for K-Center Clustering of One Billion Samples
by: Ren, Jiayang, et al.
Published: (2022)
by: Ren, Jiayang, et al.
Published: (2022)
Multi-Objective Optimization for Sparse Deep Multi-Task Learning
by: Hotegni, S. S., et al.
Published: (2023)
by: Hotegni, S. S., et al.
Published: (2023)
Similar Items
-
Multi-Objective Optimization via Wasserstein-Fisher-Rao Gradient Flow
by: Ren, Yinuo, et al.
Published: (2023) -
A Sinkhorn-type Algorithm for Constrained Optimal Transport
by: Tang, Xun, et al.
Published: (2024) -
Accelerating Sinkhorn Algorithm with Sparse Newton Iterations
by: Tang, Xun, et al.
Published: (2024) -
An efficient algorithm for entropic optimal transport under martingale-type constraints
by: Tang, Xun, et al.
Published: (2025) -
A note on continuous-time online learning
by: Ying, Lexing
Published: (2024)