Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Han, Wan, Xingchen, Liu, Yinhong, Collier, Nigel, Vulić, Ivan, Korhonen, Anna |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
por: Liu, Yinhong, et al.
Publicado: (2024)
por: Liu, Yinhong, et al.
Publicado: (2024)
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
por: Liu, Yinhong, et al.
Publicado: (2024)
por: Liu, Yinhong, et al.
Publicado: (2024)
TopViewRS: Vision-Language Models as Top-View Spatial Reasoners
por: Li, Chengzu, et al.
Publicado: (2024)
por: Li, Chengzu, et al.
Publicado: (2024)
Agentic Policy Optimization via Instruction-Policy Co-Evolution
por: Zhou, Han, et al.
Publicado: (2025)
por: Zhou, Han, et al.
Publicado: (2025)
AutoPEFT: Automatic Configuration Search for Parameter-Efficient Fine-Tuning
por: Zhou, Han, et al.
Publicado: (2023)
por: Zhou, Han, et al.
Publicado: (2023)
Beyond the Final Layer: Intermediate Representations for Better Multilingual Calibration in Large Language Models
por: Zhou, Ej, et al.
Publicado: (2025)
por: Zhou, Ej, et al.
Publicado: (2025)
Quantifying Language Disparities in Multilingual Large Language Models
por: Hu, Songbo, et al.
Publicado: (2025)
por: Hu, Songbo, et al.
Publicado: (2025)
Improving Word Translation via Two-Stage Contrastive Learning
por: Li, Yaoyiran, et al.
Publicado: (2022)
por: Li, Yaoyiran, et al.
Publicado: (2022)
Analyzing and Adapting Large Language Models for Few-Shot Multilingual NLU: Are We There Yet?
por: Razumovskaia, Evgeniia, et al.
Publicado: (2024)
por: Razumovskaia, Evgeniia, et al.
Publicado: (2024)
Prompt Compression for Large Language Models: A Survey
por: Li, Zongqian, et al.
Publicado: (2024)
por: Li, Zongqian, et al.
Publicado: (2024)
On Bilingual Lexicon Induction with Large Language Models
por: Li, Yaoyiran, et al.
Publicado: (2023)
por: Li, Yaoyiran, et al.
Publicado: (2023)
iNews: A Multimodal Dataset for Modeling Personalized Affective Responses to News
por: Hu, Tiancheng, et al.
Publicado: (2025)
por: Hu, Tiancheng, et al.
Publicado: (2025)
Large Language Models are Miscalibrated In-Context Learners
por: Li, Chengzu, et al.
Publicado: (2023)
por: Li, Chengzu, et al.
Publicado: (2023)
Improving Preference Extraction In LLMs By Identifying Latent Knowledge Through Classifying Probes
por: Maiya, Sharan, et al.
Publicado: (2025)
por: Maiya, Sharan, et al.
Publicado: (2025)
Quantifying the Persona Effect in LLM Simulations
por: Hu, Tiancheng, et al.
Publicado: (2024)
por: Hu, Tiancheng, et al.
Publicado: (2024)
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
por: Hu, Tiancheng, et al.
Publicado: (2025)
por: Hu, Tiancheng, et al.
Publicado: (2025)
Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation
por: Miranda, Lester James V., et al.
Publicado: (2026)
por: Miranda, Lester James V., et al.
Publicado: (2026)
DIALIGHT: Lightweight Multilingual Development and Evaluation of Task-Oriented Dialogue Systems with Large Language Models
por: Hu, Songbo, et al.
Publicado: (2024)
por: Hu, Songbo, et al.
Publicado: (2024)
SQATIN: Supervised Instruction Tuning Meets Question Answering for Improved Dialogue NLU
por: Razumovskaia, Evgeniia, et al.
Publicado: (2023)
por: Razumovskaia, Evgeniia, et al.
Publicado: (2023)
Visual Planning: Let's Think Only with Images
por: Xu, Yi, et al.
Publicado: (2025)
por: Xu, Yi, et al.
Publicado: (2025)
Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking
por: Hu, Songbo, et al.
Publicado: (2026)
por: Hu, Songbo, et al.
Publicado: (2026)
Artificial intelligence is creating a new global linguistic hierarchy
por: Occhini, Giulia, et al.
Publicado: (2026)
por: Occhini, Giulia, et al.
Publicado: (2026)
Generative Language Models Exhibit Social Identity Biases
por: Hu, Tiancheng, et al.
Publicado: (2023)
por: Hu, Tiancheng, et al.
Publicado: (2023)
Can LLM be a Personalized Judge?
por: Dong, Yijiang River, et al.
Publicado: (2024)
por: Dong, Yijiang River, et al.
Publicado: (2024)
Multilinguality at the Edge: Developing Language Models for the Global South
por: Miranda, Lester James V., et al.
Publicado: (2026)
por: Miranda, Lester James V., et al.
Publicado: (2026)
Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence
por: Liu, Yinhong, et al.
Publicado: (2024)
por: Liu, Yinhong, et al.
Publicado: (2024)
TOAD: Task-Oriented Automatic Dialogs with Diverse Response Styles
por: Liu, Yinhong, et al.
Publicado: (2024)
por: Liu, Yinhong, et al.
Publicado: (2024)
IROTE: Human-like Traits Elicitation of Large Language Model via In-Context Self-Reflective Optimization
por: Bai, Yuzhuo, et al.
Publicado: (2025)
por: Bai, Yuzhuo, et al.
Publicado: (2025)
Improving Bilingual Lexicon Induction with Cross-Encoder Reranking
por: Li, Yaoyiran, et al.
Publicado: (2022)
por: Li, Yaoyiran, et al.
Publicado: (2022)
When Personalization Meets Reality: A Multi-Faceted Analysis of Personalized Preference Learning
por: Dong, Yijiang River, et al.
Publicado: (2025)
por: Dong, Yijiang River, et al.
Publicado: (2025)
Scaling Sparse Fine-Tuning to Large Language Models
por: Ansell, Alan, et al.
Publicado: (2024)
por: Ansell, Alan, et al.
Publicado: (2024)
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models
por: Vijjini, Anvesh Rao, et al.
Publicado: (2024)
por: Vijjini, Anvesh Rao, et al.
Publicado: (2024)
Self-Augmented In-Context Learning for Unsupervised Word Translation
por: Li, Yaoyiran, et al.
Publicado: (2024)
por: Li, Yaoyiran, et al.
Publicado: (2024)
Can Large Language Models Simulate Human Responses? A Case Study of Stated Preference Experiments in the Context of Heating-related Choices
por: Wang, Han, et al.
Publicado: (2025)
por: Wang, Han, et al.
Publicado: (2025)
Debiasing Methods for Fairer Neural Models in Vision and Language Research: A Survey
por: Parraga, Otávio, et al.
Publicado: (2022)
por: Parraga, Otávio, et al.
Publicado: (2022)
Value of Information: A Framework for Human-Agent Communication
por: Dong, Yijiang River, et al.
Publicado: (2026)
por: Dong, Yijiang River, et al.
Publicado: (2026)
Reducing Biases towards Minoritized Populations in Medical Curricular Content via Artificial Intelligence for Fairer Health Outcomes
por: Salavati, Chiman, et al.
Publicado: (2024)
por: Salavati, Chiman, et al.
Publicado: (2024)
TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law
por: Hui, Zheng, et al.
Publicado: (2025)
por: Hui, Zheng, et al.
Publicado: (2025)
Multi-Agent Design: Optimizing Agents with Better Prompts and Topologies
por: Zhou, Han, et al.
Publicado: (2025)
por: Zhou, Han, et al.
Publicado: (2025)
Aligning Large Language Models with Healthcare Stakeholders: A Pathway to Trustworthy AI Integration
por: Ding, Kexin, et al.
Publicado: (2025)
por: Ding, Kexin, et al.
Publicado: (2025)
Ejemplares similares
-
Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
por: Liu, Yinhong, et al.
Publicado: (2024) -
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
por: Liu, Yinhong, et al.
Publicado: (2024) -
TopViewRS: Vision-Language Models as Top-View Spatial Reasoners
por: Li, Chengzu, et al.
Publicado: (2024) -
Agentic Policy Optimization via Instruction-Policy Co-Evolution
por: Zhou, Han, et al.
Publicado: (2025) -
AutoPEFT: Automatic Configuration Search for Parameter-Efficient Fine-Tuning
por: Zhou, Han, et al.
Publicado: (2023)