Beyond Performance: Quantifying and Mitigating Label Bias in LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Reif, Yuval, Schwartz, Roy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic
por: Reif, Yuval, et al.
Publicado: (2025)
por: Reif, Yuval, et al.
Publicado: (2025)
From Tokens to Words: On the Inner Lexicon of LLMs
por: Kaplan, Guy, et al.
Publicado: (2024)
por: Kaplan, Guy, et al.
Publicado: (2024)
Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models
por: Kaplan, Guy, et al.
Publicado: (2025)
por: Kaplan, Guy, et al.
Publicado: (2025)
The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
por: Shani, Chen, et al.
Publicado: (2026)
por: Shani, Chen, et al.
Publicado: (2026)
Disclosure and Mitigation of Gender Bias in LLMs
por: Dong, Xiangjue, et al.
Publicado: (2024)
por: Dong, Xiangjue, et al.
Publicado: (2024)
Mitigating Label Length Bias in Large Language Models
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
Do the Right Thing, Just Debias! Multi-Category Bias Mitigation Using LLMs
por: Roy, Amartya, et al.
Publicado: (2024)
por: Roy, Amartya, et al.
Publicado: (2024)
Why Fine-Tuning Encourages Hallucinations and How to Fix It
por: Kaplan, Guy, et al.
Publicado: (2026)
por: Kaplan, Guy, et al.
Publicado: (2026)
LLMCloudHunter: Harnessing LLMs for Automated Extraction of Detection Rules from Cloud-Based CTI
por: Schwartz, Yuval, et al.
Publicado: (2024)
por: Schwartz, Yuval, et al.
Publicado: (2024)
Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
por: Nahum, Omer, et al.
Publicado: (2024)
por: Nahum, Omer, et al.
Publicado: (2024)
Quantifying and Mitigating Premature Closure in Frontier LLMs
por: Handler, Rebecca, et al.
Publicado: (2026)
por: Handler, Rebecca, et al.
Publicado: (2026)
Relative Bias: A Comparative Framework for Quantifying Bias in LLMs
por: Arbabi, Alireza, et al.
Publicado: (2025)
por: Arbabi, Alireza, et al.
Publicado: (2025)
Quantifying and Mitigating Self-Preference Bias of LLM Judges
por: Yang, Jinming, et al.
Publicado: (2026)
por: Yang, Jinming, et al.
Publicado: (2026)
On Pruning State-Space LLMs
por: Ghattas, Tamer, et al.
Publicado: (2025)
por: Ghattas, Tamer, et al.
Publicado: (2025)
The power of Prompts: Evaluating and Mitigating Gender Bias in MT with LLMs
por: Sant, Aleix, et al.
Publicado: (2024)
por: Sant, Aleix, et al.
Publicado: (2024)
Quantifying Label-Induced Bias in Large Language Model Self- and Cross-Evaluations
por: Saraf, Muskan, et al.
Publicado: (2025)
por: Saraf, Muskan, et al.
Publicado: (2025)
How Quantization Shapes Bias in Large Language Models
por: Marcuzzi, Federico, et al.
Publicado: (2025)
por: Marcuzzi, Federico, et al.
Publicado: (2025)
Can We Trust LLMs? Mitigate Overconfidence Bias in LLMs through Knowledge Transfer
por: Yang, Haoyan, et al.
Publicado: (2024)
por: Yang, Haoyan, et al.
Publicado: (2024)
Steering Towards Fairness: Mitigating Political Bias in LLMs
por: Nadeem, Afrozah, et al.
Publicado: (2025)
por: Nadeem, Afrozah, et al.
Publicado: (2025)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
por: Yamamoto, Taisei, et al.
Publicado: (2025)
por: Yamamoto, Taisei, et al.
Publicado: (2025)
Understanding and Mitigating Gender Bias in LLMs via Interpretable Neuron Editing
por: Yu, Zeping, et al.
Publicado: (2025)
por: Yu, Zeping, et al.
Publicado: (2025)
Quantifying and Mitigating Selection Bias in LLMs: A Transferable LoRA Fine-Tuning and Efficient Majority Voting Approach
por: Guda, Blessed, et al.
Publicado: (2025)
por: Guda, Blessed, et al.
Publicado: (2025)
Format as a Prior: Quantifying and Analyzing Bias in LLMs for Heterogeneous Data
por: Liu, Jiacheng, et al.
Publicado: (2025)
por: Liu, Jiacheng, et al.
Publicado: (2025)
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
por: Long, Do Xuan, et al.
Publicado: (2024)
por: Long, Do Xuan, et al.
Publicado: (2024)
Augmenting In-Context-Learning in LLMs via Automatic Data Labeling and Refinement
por: Shtok, Joseph, et al.
Publicado: (2024)
por: Shtok, Joseph, et al.
Publicado: (2024)
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
por: Wei, Kangda, et al.
Publicado: (2025)
por: Wei, Kangda, et al.
Publicado: (2025)
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2026)
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2026)
CalibraEval: Calibrating Prediction Distribution to Mitigate Selection Bias in LLMs-as-Judges
por: Li, Haitao, et al.
Publicado: (2024)
por: Li, Haitao, et al.
Publicado: (2024)
Discrimination by LLMs: Cross-lingual Bias Assessment and Mitigation in Decision-Making and Summarisation
por: Huijzer, Willem, et al.
Publicado: (2025)
por: Huijzer, Willem, et al.
Publicado: (2025)
Position Bias Mitigates Position Bias:Mitigate Position Bias Through Inter-Position Knowledge Distillation
por: Wang, Yifei, et al.
Publicado: (2025)
por: Wang, Yifei, et al.
Publicado: (2025)
Interpreting and Mitigating Unwanted Uncertainty in LLMs
por: Roy, Tiasa Singha, et al.
Publicado: (2025)
por: Roy, Tiasa Singha, et al.
Publicado: (2025)
Detecting and Mitigating Bias in LLMs through Knowledge Graph-Augmented Training
por: Kumar, Rajeev, et al.
Publicado: (2025)
por: Kumar, Rajeev, et al.
Publicado: (2025)
Bias Testing and Mitigation in Black Box LLMs using Metamorphic Relations
por: Salimian, Sina, et al.
Publicado: (2025)
por: Salimian, Sina, et al.
Publicado: (2025)
Veracity Bias and Beyond: Uncovering LLMs' Hidden Beliefs in Problem-Solving Reasoning
por: Zhou, Yue, et al.
Publicado: (2025)
por: Zhou, Yue, et al.
Publicado: (2025)
Breaking Bias, Building Bridges: Evaluation and Mitigation of Social Biases in LLMs via Contact Hypothesis
por: Raj, Chahat, et al.
Publicado: (2024)
por: Raj, Chahat, et al.
Publicado: (2024)
Sometimes the Model doth Preach: Quantifying Religious Bias in Open LLMs through Demographic Analysis in Asian Nations
por: Shankar, Hari, et al.
Publicado: (2025)
por: Shankar, Hari, et al.
Publicado: (2025)
AGR: Age Group fairness Reward for Bias Mitigation in LLMs
por: Cao, Shuirong, et al.
Publicado: (2024)
por: Cao, Shuirong, et al.
Publicado: (2024)
Shifting Perspectives: Steering Vectors for Robust Bias Mitigation in LLMs
por: Siddique, Zara, et al.
Publicado: (2025)
por: Siddique, Zara, et al.
Publicado: (2025)
From Oracle to Noisy Context: Mitigating Contextual Exposure Bias in Speech-LLMs
por: Guo, Xiaoyong, et al.
Publicado: (2026)
por: Guo, Xiaoyong, et al.
Publicado: (2026)
Beyond the Vision Encoder: Identifying and Mitigating Spatial Bias in Large Vision-Language Models
por: Zhu, Yingjie, et al.
Publicado: (2025)
por: Zhu, Yingjie, et al.
Publicado: (2025)
Ejemplares similares
-
Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic
por: Reif, Yuval, et al.
Publicado: (2025) -
From Tokens to Words: On the Inner Lexicon of LLMs
por: Kaplan, Guy, et al.
Publicado: (2024) -
Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models
por: Kaplan, Guy, et al.
Publicado: (2025) -
The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
por: Shani, Chen, et al.
Publicado: (2026) -
Disclosure and Mitigation of Gender Bias in LLMs
por: Dong, Xiangjue, et al.
Publicado: (2024)