Optimizing LLMs with Direct Preferences: A Data Efficiency Perspective
Fuente:
arXiv
Saved in:
| Main Authors: | Bernardelle, Pietro, Demartini, Gianluca |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Query-Document Dense Vectors for LLM Relevance Judgment Bias Analysis
by: Mohtadi, Samaneh, et al.
Published: (2026)
by: Mohtadi, Samaneh, et al.
Published: (2026)
The Impact of Persona-based Political Perspectives on Hateful Content Detection
by: Civelli, Stefano, et al.
Published: (2025)
by: Civelli, Stefano, et al.
Published: (2025)
On Softmax Direct Preference Optimization for Recommendation
by: Chen, Yuxin, et al.
Published: (2024)
by: Chen, Yuxin, et al.
Published: (2024)
The Effect of Document Summarization on LLM-Based Relevance Judgments
by: Mohtadi, Samaneh, et al.
Published: (2025)
by: Mohtadi, Samaneh, et al.
Published: (2025)
Causal Direct Preference Optimization for Distributionally Robust Generative Recommendation
by: Zhao, Chu, et al.
Published: (2026)
by: Zhao, Chu, et al.
Published: (2026)
On Negative-aware Preference Optimization for Recommendation
by: Ding, Chenlu, et al.
Published: (2025)
by: Ding, Chenlu, et al.
Published: (2025)
DynamicPO: Dynamic Preference Optimization for Recommendation
by: Hu, Xingyu, et al.
Published: (2026)
by: Hu, Xingyu, et al.
Published: (2026)
Generative Retrieval with Preference Optimization for E-commerce Search
by: Li, Mingming, et al.
Published: (2024)
by: Li, Mingming, et al.
Published: (2024)
Temporal User Profiling with LLMs: Balancing Short-Term and Long-Term Preferences for Recommendations
by: Sabouri, Milad, et al.
Published: (2025)
by: Sabouri, Milad, et al.
Published: (2025)
RLPO: Residual Listwise Preference Optimization for Long-Context Review Ranking
by: Jiang, Hao, et al.
Published: (2026)
by: Jiang, Hao, et al.
Published: (2026)
A Shared Geometry of Difficulty in Multilingual Language Models
by: Civelli, Stefano, et al.
Published: (2026)
by: Civelli, Stefano, et al.
Published: (2026)
ICPO: Intrinsic Confidence-Driven Group Relative Preference Optimization for Efficient Reinforcement Learning
by: Wang, Jinpeng, et al.
Published: (2025)
by: Wang, Jinpeng, et al.
Published: (2025)
Survey on Semantic Interpretation of Tabular Data: Challenges and Directions
by: Cremaschi, Marco, et al.
Published: (2024)
by: Cremaschi, Marco, et al.
Published: (2024)
Optimizing and Evaluating Enterprise Retrieval-Augmented Generation (RAG): A Content Design Perspective
by: Packowski, Sarah, et al.
Published: (2024)
by: Packowski, Sarah, et al.
Published: (2024)
CoPL: Collaborative Preference Learning for Personalizing LLMs
by: Choi, Youngbin, et al.
Published: (2025)
by: Choi, Youngbin, et al.
Published: (2025)
Flexible Generation of Preference Data for Recommendation Analysis
by: Mungari, Simone, et al.
Published: (2024)
by: Mungari, Simone, et al.
Published: (2024)
Preference Diffusion for Recommendation
by: Liu, Shuo, et al.
Published: (2024)
by: Liu, Shuo, et al.
Published: (2024)
Empirical and Experimental Perspectives on Big Data in Recommendation Systems: A Comprehensive Survey
by: Taha, Kamal, et al.
Published: (2024)
by: Taha, Kamal, et al.
Published: (2024)
Tree of Preferences for Diversified Recommendation
by: Yuan, Hanyang, et al.
Published: (2025)
by: Yuan, Hanyang, et al.
Published: (2025)
Integrating SPARQL and LLMs for Question Answering over Scholarly Data Sources
by: Fondi, Fomubad Borista, et al.
Published: (2024)
by: Fondi, Fomubad Borista, et al.
Published: (2024)
Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation
by: Zhang, Benyu, et al.
Published: (2026)
by: Zhang, Benyu, et al.
Published: (2026)
Greedy SLIM: A SLIM-Based Approach For Preference Elicitation
by: Proissl, Claudius, et al.
Published: (2024)
by: Proissl, Claudius, et al.
Published: (2024)
PRECTR-V2:Unified Relevance-CTR Framework with Cross-User Preference Mining, Exposure Bias Correction, and LLM-Distilled Encoder Optimization
by: Cao, Shuzhi, et al.
Published: (2026)
by: Cao, Shuzhi, et al.
Published: (2026)
GenRec: A Preference-Oriented Generative Framework for Large-Scale Recommendation
by: Zou, Yanyan, et al.
Published: (2026)
by: Zou, Yanyan, et al.
Published: (2026)
A Text-Based Recommender System that Leverages Explicit Affective State Preferences
by: Hasan, Tonmoy, et al.
Published: (2025)
by: Hasan, Tonmoy, et al.
Published: (2025)
Evaluating AI Recruitment Sourcing Tools by Human Preference
by: Slaykovskiy, Vladimir, et al.
Published: (2025)
by: Slaykovskiy, Vladimir, et al.
Published: (2025)
CROWN: A Novel Approach to Comprehending Users' Preferences for Accurate Personalized News Recommendation
by: Ko, Yunyong, et al.
Published: (2023)
by: Ko, Yunyong, et al.
Published: (2023)
Addressing Labelled Data Scarcity: Taxonomy-Agnostic Annotation of PII Values in HTTP Traffic using LLMs
by: Cory, Thomas, et al.
Published: (2026)
by: Cory, Thomas, et al.
Published: (2026)
Decoding Style: Efficient Fine-Tuning of LLMs for Image-Guided Outfit Recommendation with Preference
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
Reasoning over User Preferences: Knowledge Graph-Augmented LLMs for Explainable Conversational Recommendations
by: Qiu, Zhangchi, et al.
Published: (2024)
by: Qiu, Zhangchi, et al.
Published: (2024)
Dual Contrastive Transformer for Hierarchical Preference Modeling in Sequential Recommendation
by: Huang, Chengkai, et al.
Published: (2024)
by: Huang, Chengkai, et al.
Published: (2024)
Time-Aware Diffusion based on Preference Disentanglement for Generative Recommendation
by: Zhu, Bangguo, et al.
Published: (2026)
by: Zhu, Bangguo, et al.
Published: (2026)
Trustworthy Intelligent Education: A Systematic Perspective on Progress, Challenges, and Future Directions
by: Yu, Xiaoshan, et al.
Published: (2026)
by: Yu, Xiaoshan, et al.
Published: (2026)
Dynamic Evaluation Framework for Personalized and Trustworthy Agents: A Multi-Session Approach to Preference Adaptability
by: Shah, Chirag, et al.
Published: (2025)
by: Shah, Chirag, et al.
Published: (2025)
MTRec: Learning to Align with User Preferences via Mental Reward Models
by: Zhao, Mengchen, et al.
Published: (2025)
by: Zhao, Mengchen, et al.
Published: (2025)
The Challenge of Using LLMs to Simulate Human Behavior: A Causal Inference Perspective
by: Gui, George, et al.
Published: (2023)
by: Gui, George, et al.
Published: (2023)
A Preference-oriented Diversity Model Based on Mutual-information in Re-ranking for E-commerce Search
by: Wang, Huimu, et al.
Published: (2024)
by: Wang, Huimu, et al.
Published: (2024)
Cross-domain Transfer of Valence Preferences via a Meta-optimization Approach
by: Zhao, Chuang, et al.
Published: (2024)
by: Zhao, Chuang, et al.
Published: (2024)
AlignGroup: Learning and Aligning Group Consensus with Member Preferences for Group Recommendation
by: Xu, Jinfeng, et al.
Published: (2024)
by: Xu, Jinfeng, et al.
Published: (2024)
A Survey of Reasoning for Substitution Relationships: Definitions, Methods, and Directions
by: Yang, Anxin, et al.
Published: (2024)
by: Yang, Anxin, et al.
Published: (2024)
Similar Items
-
Query-Document Dense Vectors for LLM Relevance Judgment Bias Analysis
by: Mohtadi, Samaneh, et al.
Published: (2026) -
The Impact of Persona-based Political Perspectives on Hateful Content Detection
by: Civelli, Stefano, et al.
Published: (2025) -
On Softmax Direct Preference Optimization for Recommendation
by: Chen, Yuxin, et al.
Published: (2024) -
The Effect of Document Summarization on LLM-Based Relevance Judgments
by: Mohtadi, Samaneh, et al.
Published: (2025) -
Causal Direct Preference Optimization for Distributionally Robust Generative Recommendation
by: Zhao, Chu, et al.
Published: (2026)