PAL: Pluralistic Alignment Framework for Learning from Heterogeneous Preferences
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Daiwei, Chen, Yi, Rege, Aniket, Vinayak, Ramya Korlakai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Metric Learning from Limited Pairwise Preference Comparisons
by: Wang, Zhi, et al.
Published: (2024)
by: Wang, Zhi, et al.
Published: (2024)
Is Conformal Factuality for RAG-based LLMs Robust? Novel Metrics and Systematic Insights
by: Chen, Yi, et al.
Published: (2026)
by: Chen, Yi, et al.
Published: (2026)
Bridging Lifelong and Multi-Task Representation Learning via Algorithm and Complexity Measure
by: Wang, Zhi, et al.
Published: (2025)
by: Wang, Zhi, et al.
Published: (2025)
Metric Learning in an RKHS
by: Tatli, Gokcan, et al.
Published: (2025)
by: Tatli, Gokcan, et al.
Published: (2025)
Taming False Positives in Out-of-Distribution Detection with Human Feedback
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection
by: Yamada, Daisuke, et al.
Published: (2025)
by: Yamada, Daisuke, et al.
Published: (2025)
Agentic Very Long Video Understanding
by: Rege, Aniket, et al.
Published: (2026)
by: Rege, Aniket, et al.
Published: (2026)
Promises and Pitfalls of Threshold-based Auto-labeling
by: Vishwakarma, Harit, et al.
Published: (2022)
by: Vishwakarma, Harit, et al.
Published: (2022)
Pearls from Pebbles: Improved Confidence Functions for Auto-labeling
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
Almost Asymptotically Optimal Active Clustering Through Pairwise Observations
by: Teo, Rachel S. Y., et al.
Published: (2026)
by: Teo, Rachel S. Y., et al.
Published: (2026)
MixDPO: Modeling Preference Strength for Pluralistic Alignment
by: Imai, Saki, et al.
Published: (2026)
by: Imai, Saki, et al.
Published: (2026)
APPA: Adaptive Preference Pluralistic Alignment for Fair Federated RLHF of LLMs
by: Srewa, Mahmoud, et al.
Published: (2026)
by: Srewa, Mahmoud, et al.
Published: (2026)
Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
by: Harland, Hadassah, et al.
Published: (2024)
by: Harland, Hadassah, et al.
Published: (2024)
Pluralistic Alignment for Healthcare: A Role-Driven Framework
by: Zhong, Jiayou, et al.
Published: (2025)
by: Zhong, Jiayou, et al.
Published: (2025)
Data Attribution in Adaptive Learning
by: Rege, Amit Kiran
Published: (2026)
by: Rege, Amit Kiran
Published: (2026)
Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
by: Chen, Daiwei, et al.
Published: (2026)
by: Chen, Daiwei, et al.
Published: (2026)
Pairwise Calibrated Rewards for Pluralistic Alignment
by: Halpern, Daniel, et al.
Published: (2025)
by: Halpern, Daniel, et al.
Published: (2025)
Pluralistic Alignment Over Time
by: Klassen, Toryn Q., et al.
Published: (2024)
by: Klassen, Toryn Q., et al.
Published: (2024)
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
PluralLLM: Pluralistic Alignment in LLMs via Federated Learning
by: Srewa, Mahmoud, et al.
Published: (2025)
by: Srewa, Mahmoud, et al.
Published: (2025)
Direct Alignment with Heterogeneous Preferences
by: Shirali, Ali, et al.
Published: (2025)
by: Shirali, Ali, et al.
Published: (2025)
Counterfactual Reasoning for Steerable Pluralistic Value Alignment of Large Language Models
by: Guo, Hanze, et al.
Published: (2025)
by: Guo, Hanze, et al.
Published: (2025)
A Persona-Based Evaluation Framework for Pluralistic Alignment in Generative AI
by: Karagoz, Atahan
Published: (2026)
by: Karagoz, Atahan
Published: (2026)
CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems
by: Rege, Aniket, et al.
Published: (2025)
by: Rege, Aniket, et al.
Published: (2025)
Adaptive Pluralistic Alignment: A pipeline for dynamic artificial democracy
by: Freedman, Rachel
Published: (2026)
by: Freedman, Rachel
Published: (2026)
GRASP: Graph Agentic Search over Propositions for Multi-hop Question Answering
by: Jenkins, Stockton, et al.
Published: (2026)
by: Jenkins, Stockton, et al.
Published: (2026)
Learning Capacity: A Measure of the Effective Dimensionality of a Model
by: Chen, Daiwei, et al.
Published: (2023)
by: Chen, Daiwei, et al.
Published: (2023)
Exploring Chain-of-Thought Reasoning for Steerable Pluralistic Alignment
by: Zhang, Yunfan, et al.
Published: (2025)
by: Zhang, Yunfan, et al.
Published: (2025)
The Role of Generator Access in Autoregressive Post-Training
by: Rege, Amit Kiran
Published: (2026)
by: Rege, Amit Kiran
Published: (2026)
A Unified Framework for Locality in Scalable MARL
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
PAL: Prompting Analytic Learning with Missing Modality for Multi-Modal Class-Incremental Learning
by: Yue, Xianghu, et al.
Published: (2025)
by: Yue, Xianghu, et al.
Published: (2025)
Data Selection for LLM Alignment Using Fine-Grained Preferences
by: Zhang, Jia, et al.
Published: (2025)
by: Zhang, Jia, et al.
Published: (2025)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
by: Shetty, Anudeex, et al.
Published: (2025)
by: Shetty, Anudeex, et al.
Published: (2025)
VISPA: Pluralistic Alignment via Automatic Value Selection and Activation
by: Zheng, Shenyan, et al.
Published: (2026)
by: Zheng, Shenyan, et al.
Published: (2026)
Adversarial Preference Learning for Robust LLM Alignment
by: Wang, Yuanfu, et al.
Published: (2025)
by: Wang, Yuanfu, et al.
Published: (2025)
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
by: Alamdari, Parand A., et al.
Published: (2024)
by: Alamdari, Parand A., et al.
Published: (2024)
Response Time Enhances Alignment with Heterogeneous Preferences
by: Echenique, Federico, et al.
Published: (2026)
by: Echenique, Federico, et al.
Published: (2026)
Multi-Agent Lipschitz Bandits
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
Spectral Souping: A Unified Framework for Online Preference Alignment
by: Chow, Yinlam, et al.
Published: (2026)
by: Chow, Yinlam, et al.
Published: (2026)
Personalized Group Relative Policy Optimization for Heterogenous Preference Alignment
by: Wang, Jialu, et al.
Published: (2026)
by: Wang, Jialu, et al.
Published: (2026)
Similar Items
-
Metric Learning from Limited Pairwise Preference Comparisons
by: Wang, Zhi, et al.
Published: (2024) -
Is Conformal Factuality for RAG-based LLMs Robust? Novel Metrics and Systematic Insights
by: Chen, Yi, et al.
Published: (2026) -
Bridging Lifelong and Multi-Task Representation Learning via Algorithm and Complexity Measure
by: Wang, Zhi, et al.
Published: (2025) -
Metric Learning in an RKHS
by: Tatli, Gokcan, et al.
Published: (2025) -
Taming False Positives in Out-of-Distribution Detection with Human Feedback
by: Vishwakarma, Harit, et al.
Published: (2024)