Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Yuhui, Wang, Xiyao, Li, Zixi, Ding, YiTian, Ling, Tianyang, Chen, Jialuo, Yu, Tianyi, Yuan, Zhenlong, Zhao, Jinman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026)
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
von: Reddy, Sandeep, et al.
Veröffentlicht: (2025)
von: Reddy, Sandeep, et al.
Veröffentlicht: (2025)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025)
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025)
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
von: Mandal, Paul K.
Veröffentlicht: (2025)
von: Mandal, Paul K.
Veröffentlicht: (2025)
Less is More: Learning Graph Tasks with Just LLMs
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
von: Schneider, Felix, et al.
Veröffentlicht: (2026)
von: Schneider, Felix, et al.
Veröffentlicht: (2026)
Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
von: Resck, Lucas, et al.
Veröffentlicht: (2026)
von: Resck, Lucas, et al.
Veröffentlicht: (2026)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
von: Bachar, Or, et al.
Veröffentlicht: (2026)
von: Bachar, Or, et al.
Veröffentlicht: (2026)
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues
von: Kajare, Prajwal Vijay, et al.
Veröffentlicht: (2026)
von: Kajare, Prajwal Vijay, et al.
Veröffentlicht: (2026)
PersonalLLM: Tailoring LLMs to Individual Preferences
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences
von: Zhou, Yangyang, et al.
Veröffentlicht: (2026)
von: Zhou, Yangyang, et al.
Veröffentlicht: (2026)
A Comparative Study of Feature Selection in Tsetlin Machines
von: Halenka, Vojtech, et al.
Veröffentlicht: (2025)
von: Halenka, Vojtech, et al.
Veröffentlicht: (2025)
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
von: Han, Xudong, et al.
Veröffentlicht: (2025)
von: Han, Xudong, et al.
Veröffentlicht: (2025)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
von: Zolduoarrati, Elijah, et al.
Veröffentlicht: (2025)
von: Zolduoarrati, Elijah, et al.
Veröffentlicht: (2025)
Neural Attention: A Novel Mechanism for Enhanced Expressive Power in Transformer Models
von: DiGiugno, Andrew, et al.
Veröffentlicht: (2025)
von: DiGiugno, Andrew, et al.
Veröffentlicht: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Curveball Steering: The Right Direction To Steer Isn't Always Linear
von: Raval, Shivam, et al.
Veröffentlicht: (2026)
von: Raval, Shivam, et al.
Veröffentlicht: (2026)
Cytoarchitecture in Words: Weakly Supervised Vision-Language Modeling for Human Brain Microscopy
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
Let Your Graph Do the Talking: Encoding Structured Data for LLMs
von: Perozzi, Bryan, et al.
Veröffentlicht: (2024)
von: Perozzi, Bryan, et al.
Veröffentlicht: (2024)
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025)
von: Amini, Ali
Veröffentlicht: (2025)
Mixup Model Merge: Enhancing Model Merging Performance through Randomized Linear Interpolation
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
Scalable GPU-Accelerated Euler Characteristic Curves: Optimization and Differentiable Learning for PyTorch
von: Saxena, Udit
Veröffentlicht: (2025)
von: Saxena, Udit
Veröffentlicht: (2025)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Automated Bug Triaging using Instruction-Tuned Large Language Models
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
Merge-Bench: Resolve Merge Conflicts with Large Language Models
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
PRPO: Aligning Process Reward with Outcome Reward in Policy Optimization
von: Ding, Ruiyi, et al.
Veröffentlicht: (2026)
von: Ding, Ruiyi, et al.
Veröffentlicht: (2026)
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
von: Asanuma, Haruka, et al.
Veröffentlicht: (2025)
von: Asanuma, Haruka, et al.
Veröffentlicht: (2025)
ADALog: Adaptive Unsupervised Anomaly detection in Logs with Self-attention Masked Language Model
von: Pospieszny, Przemek, et al.
Veröffentlicht: (2025)
von: Pospieszny, Przemek, et al.
Veröffentlicht: (2025)
Optimized Gradient Clipping for Noisy Label Learning
von: Ye, Xichen, et al.
Veröffentlicht: (2024)
von: Ye, Xichen, et al.
Veröffentlicht: (2024)
MIMIC-SR-ICD11: A Dataset for Narrative-Based Diagnosis
von: Wu, Yuexin, et al.
Veröffentlicht: (2025)
von: Wu, Yuexin, et al.
Veröffentlicht: (2025)
DROID: Dual Representation for Out-of-Scope Intent Detection
von: Rashwan, Wael, et al.
Veröffentlicht: (2025)
von: Rashwan, Wael, et al.
Veröffentlicht: (2025)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026) -
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
von: Reddy, Sandeep, et al.
Veröffentlicht: (2025) -
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
von: Sharma, Anika, et al.
Veröffentlicht: (2025) -
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025) -
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
von: Mandal, Paul K.
Veröffentlicht: (2025)