Controllable Preference Optimization: Toward Controllable Multi-Objective Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Guo, Yiju, Cui, Ganqu, Yuan, Lifan, Ding, Ning, Sun, Zexu, Sun, Bowen, Chen, Huimin, Xie, Ruobing, Zhou, Jie, Lin, Yankai, Liu, Zhiyuan, Sun, Maosong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Advancing LLM Reasoning Generalists with Preference Trees
di: Yuan, Lifan, et al.
Pubblicazione: (2024)
di: Yuan, Lifan, et al.
Pubblicazione: (2024)
Learning to Focus: Causal Attention Distillation via Gradient-Guided Token Pruning
di: Guo, Yiju, et al.
Pubblicazione: (2025)
di: Guo, Yiju, et al.
Pubblicazione: (2025)
UltraFeedback: Boosting Language Models with Scaled AI Feedback
di: Cui, Ganqu, et al.
Pubblicazione: (2023)
di: Cui, Ganqu, et al.
Pubblicazione: (2023)
Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification
di: Guo, Yiju, et al.
Pubblicazione: (2026)
di: Guo, Yiju, et al.
Pubblicazione: (2026)
Mastering Text, Code and Math Simultaneously via Fusing Highly Specialized Language Models
di: Ding, Ning, et al.
Pubblicazione: (2024)
di: Ding, Ning, et al.
Pubblicazione: (2024)
From $f(x)$ and $g(x)$ to $f(g(x))$: LLMs Learn New Skills in RL by Composing Old Ones
di: Yuan, Lifan, et al.
Pubblicazione: (2025)
di: Yuan, Lifan, et al.
Pubblicazione: (2025)
Representation Learning for Natural Language Processing
di: Liu, Zhiyuan, et al.
Pubblicazione: (2020)
di: Liu, Zhiyuan, et al.
Pubblicazione: (2020)
Meta-Aligner: Bidirectional Preference-Policy Optimization for Multi-Objective LLMs Alignment
di: Xu, Wenzhe, et al.
Pubblicazione: (2026)
di: Xu, Wenzhe, et al.
Pubblicazione: (2026)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
di: He, Bingxiang, et al.
Pubblicazione: (2024)
di: He, Bingxiang, et al.
Pubblicazione: (2024)
AIR: A Systematic Analysis of Annotations, Instructions, and Response Pairs in Preference Dataset
di: He, Bingxiang, et al.
Pubblicazione: (2025)
di: He, Bingxiang, et al.
Pubblicazione: (2025)
RLPR: Extrapolating RLVR to General Domains without Verifiers
di: Yu, Tianyu, et al.
Pubblicazione: (2025)
di: Yu, Tianyu, et al.
Pubblicazione: (2025)
Free Process Rewards without Process Labels
di: Yuan, Lifan, et al.
Pubblicazione: (2024)
di: Yuan, Lifan, et al.
Pubblicazione: (2024)
Exploring the Benefit of Activation Sparsity in Pre-training
di: Zhang, Zhengyan, et al.
Pubblicazione: (2024)
di: Zhang, Zhengyan, et al.
Pubblicazione: (2024)
The Overthinker's DIET: Cutting Token Calories with DIfficulty-AwarE Training
di: Chen, Weize, et al.
Pubblicazione: (2025)
di: Chen, Weize, et al.
Pubblicazione: (2025)
Noise Contrastive Alignment of Language Models with Explicit Rewards
di: Chen, Huayu, et al.
Pubblicazione: (2024)
di: Chen, Huayu, et al.
Pubblicazione: (2024)
RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback
di: Yu, Tianyu, et al.
Pubblicazione: (2023)
di: Yu, Tianyu, et al.
Pubblicazione: (2023)
LaSeR: Reinforcement Learning with Last-Token Self-Rewarding
di: Yang, Wenkai, et al.
Pubblicazione: (2025)
di: Yang, Wenkai, et al.
Pubblicazione: (2025)
Empowering Private Tutoring by Chaining Large Language Models
di: Chen, Yulin, et al.
Pubblicazione: (2023)
di: Chen, Yulin, et al.
Pubblicazione: (2023)
Variator: Accelerating Pre-trained Models with Plug-and-Play Compression Modules
di: Xiao, Chaojun, et al.
Pubblicazione: (2023)
di: Xiao, Chaojun, et al.
Pubblicazione: (2023)
Beyond One-Preference-Fits-All Alignment: Multi-Objective Direct Preference Optimization
di: Zhou, Zhanhui, et al.
Pubblicazione: (2023)
di: Zhou, Zhanhui, et al.
Pubblicazione: (2023)
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System
di: Chen, Weize, et al.
Pubblicazione: (2024)
di: Chen, Weize, et al.
Pubblicazione: (2024)
M$^3$TN: Multi-gate Mixture-of-Experts based Multi-valued Treatment Network for Uplift Modeling
di: Sun, Zexu, et al.
Pubblicazione: (2024)
di: Sun, Zexu, et al.
Pubblicazione: (2024)
Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
di: Wang, Haoxiang, et al.
Pubblicazione: (2024)
di: Wang, Haoxiang, et al.
Pubblicazione: (2024)
Mitigating Hallucination in Multimodal Large Language Model via Hallucination-targeted Direct Preference Optimization
di: Fu, Yuhan, et al.
Pubblicazione: (2024)
di: Fu, Yuhan, et al.
Pubblicazione: (2024)
Beyond Natural Language: LLMs Leveraging Alternative Formats for Enhanced Reasoning and Communication
di: Chen, Weize, et al.
Pubblicazione: (2024)
di: Chen, Weize, et al.
Pubblicazione: (2024)
Quality-Diversity Optimization as Multi-Objective Optimization
di: Lin, Xi, et al.
Pubblicazione: (2026)
di: Lin, Xi, et al.
Pubblicazione: (2026)
H-Neurons: On the Existence, Impact, and Origin of Hallucination-Associated Neurons in LLMs
di: Gao, Cheng, et al.
Pubblicazione: (2025)
di: Gao, Cheng, et al.
Pubblicazione: (2025)
Enhancing Legal Case Retrieval via Scaling High-quality Synthetic Query-Candidate Pairs
di: Gao, Cheng, et al.
Pubblicazione: (2024)
di: Gao, Cheng, et al.
Pubblicazione: (2024)
TTRL: Test-Time Reinforcement Learning
di: Zuo, Yuxin, et al.
Pubblicazione: (2025)
di: Zuo, Yuxin, et al.
Pubblicazione: (2025)
Preference Orchestrator: Prompt-Aware Multi-Objective Alignment for Large Language Models
di: Liu, Biao, et al.
Pubblicazione: (2025)
di: Liu, Biao, et al.
Pubblicazione: (2025)
Scalable Efficient Training of Large Language Models with Low-dimensional Projected Attention
di: Lv, Xingtai, et al.
Pubblicazione: (2024)
di: Lv, Xingtai, et al.
Pubblicazione: (2024)
Multi-Objective Preference Optimization: Improving Human Alignment of Generative Models
di: Agnihotri, Akhil, et al.
Pubblicazione: (2025)
di: Agnihotri, Akhil, et al.
Pubblicazione: (2025)
Self-Play Preference Optimization for Language Model Alignment
di: Wu, Yue, et al.
Pubblicazione: (2024)
di: Wu, Yue, et al.
Pubblicazione: (2024)
C-MORAL: Controllable Multi-Objective Molecular Optimization with Reinforcement Alignment for LLMs
di: Gao, Rui, et al.
Pubblicazione: (2026)
di: Gao, Rui, et al.
Pubblicazione: (2026)
Tell Me More! Towards Implicit User Intention Understanding of Language Model Driven Agents
di: Qian, Cheng, et al.
Pubblicazione: (2024)
di: Qian, Cheng, et al.
Pubblicazione: (2024)
Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search
di: Lu, Jialin, et al.
Pubblicazione: (2026)
di: Lu, Jialin, et al.
Pubblicazione: (2026)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
di: Sun, Yuhui, et al.
Pubblicazione: (2025)
di: Sun, Yuhui, et al.
Pubblicazione: (2025)
Self-Improvement Towards Pareto Optimality: Mitigating Preference Conflicts in Multi-Objective Alignment
di: Li, Moxin, et al.
Pubblicazione: (2025)
di: Li, Moxin, et al.
Pubblicazione: (2025)
Robust Multi-Objective Preference Alignment with Online DPO
di: Gupta, Raghav, et al.
Pubblicazione: (2025)
di: Gupta, Raghav, et al.
Pubblicazione: (2025)
INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair
di: Wang, Hanbin, et al.
Pubblicazione: (2023)
di: Wang, Hanbin, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Advancing LLM Reasoning Generalists with Preference Trees
di: Yuan, Lifan, et al.
Pubblicazione: (2024) -
Learning to Focus: Causal Attention Distillation via Gradient-Guided Token Pruning
di: Guo, Yiju, et al.
Pubblicazione: (2025) -
UltraFeedback: Boosting Language Models with Scaled AI Feedback
di: Cui, Ganqu, et al.
Pubblicazione: (2023) -
Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification
di: Guo, Yiju, et al.
Pubblicazione: (2026) -
Mastering Text, Code and Math Simultaneously via Fusing Highly Specialized Language Models
di: Ding, Ning, et al.
Pubblicazione: (2024)