Self-Preference Bias in LLM-as-a-Judge
Fuente:
arXiv
Salvato in:
| Autori principali: | Wataoka, Koki, Takahashi, Tsubasa, Ri, Ryokan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Self-Translate-Train: Enhancing Cross-Lingual Transfer of Large Language Models via Inherent Capability
di: Ri, Ryokan, et al.
Pubblicazione: (2024)
di: Ri, Ryokan, et al.
Pubblicazione: (2024)
LEIA: Facilitating Cross-lingual Knowledge Transfer in Language Models with Entity-based Data Augmentation
di: Yamada, Ikuya, et al.
Pubblicazione: (2024)
di: Yamada, Ikuya, et al.
Pubblicazione: (2024)
Natural Fingerprints of Large Language Models
di: Suzuki, Teppei, et al.
Pubblicazione: (2025)
di: Suzuki, Teppei, et al.
Pubblicazione: (2025)
Quantifying and Mitigating Self-Preference Bias of LLM Judges
di: Yang, Jinming, et al.
Pubblicazione: (2026)
di: Yang, Jinming, et al.
Pubblicazione: (2026)
Large Vocabulary Size Improves Large Language Models
di: Takase, Sho, et al.
Pubblicazione: (2024)
di: Takase, Sho, et al.
Pubblicazione: (2024)
Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming
di: Kavumba, Pride, et al.
Pubblicazione: (2026)
di: Kavumba, Pride, et al.
Pubblicazione: (2026)
Assistant-Guided Mitigation of Teacher Preference Bias in LLM-as-a-Judge
di: Liu, Zhuo, et al.
Pubblicazione: (2025)
di: Liu, Zhuo, et al.
Pubblicazione: (2025)
Dynamic Injection of Entity Knowledge into Dense Retrievers
di: Yamada, Ikuya, et al.
Pubblicazione: (2025)
di: Yamada, Ikuya, et al.
Pubblicazione: (2025)
MergePrint: Merge-Resistant Fingerprints for Robust Black-box Ownership Verification of Large Language Models
di: Yamabe, Shojiro, et al.
Pubblicazione: (2024)
di: Yamabe, Shojiro, et al.
Pubblicazione: (2024)
The Silent Judge: Unacknowledged Shortcut Bias in LLM-as-a-Judge
di: Marioriyad, Arash, et al.
Pubblicazione: (2025)
di: Marioriyad, Arash, et al.
Pubblicazione: (2025)
Evaluating Scoring Bias in LLM-as-a-Judge
di: Li, Qingquan, et al.
Pubblicazione: (2025)
di: Li, Qingquan, et al.
Pubblicazione: (2025)
Judging the Judges: A Systematic Study of Position Bias in LLM-as-a-Judge
di: Shi, Lin, et al.
Pubblicazione: (2024)
di: Shi, Lin, et al.
Pubblicazione: (2024)
CyclicJudge: Mitigating Judge Bias Efficiently in LLM-based Evaluation
di: Zhu, Ziyi, et al.
Pubblicazione: (2026)
di: Zhu, Ziyi, et al.
Pubblicazione: (2026)
Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge
di: Spiliopoulou, Evangelia, et al.
Pubblicazione: (2025)
di: Spiliopoulou, Evangelia, et al.
Pubblicazione: (2025)
Automated Concept Discovery for LLM-as-a-Judge Preference Analysis
di: Wedgwood, James, et al.
Pubblicazione: (2026)
di: Wedgwood, James, et al.
Pubblicazione: (2026)
Faithful or Fabricated? A Causal Framework for Rationalization Bias in LLM Judges
di: Tapwal, Riya, et al.
Pubblicazione: (2026)
di: Tapwal, Riya, et al.
Pubblicazione: (2026)
BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation
di: Lai, Peng, et al.
Pubblicazione: (2026)
di: Lai, Peng, et al.
Pubblicazione: (2026)
Fairness or Fluency? An Investigation into Language Bias of Pairwise LLM-as-a-Judge
di: Zhou, Xiaolin, et al.
Pubblicazione: (2026)
di: Zhou, Xiaolin, et al.
Pubblicazione: (2026)
Contrastive Decoding Mitigates Score Range Bias in LLM-as-a-Judge
di: Fujinuma, Yoshinari
Pubblicazione: (2025)
di: Fujinuma, Yoshinari
Pubblicazione: (2025)
Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization
di: Zhou, Hongli, et al.
Pubblicazione: (2026)
di: Zhou, Hongli, et al.
Pubblicazione: (2026)
Explaining Length Bias in LLM-Based Preference Evaluations
di: Hu, Zhengyu, et al.
Pubblicazione: (2024)
di: Hu, Zhengyu, et al.
Pubblicazione: (2024)
Am I More Pointwise or Pairwise? Revealing Position Bias in Rubric-Based LLM-as-a-Judge
di: Xu, Yuzheng, et al.
Pubblicazione: (2026)
di: Xu, Yuzheng, et al.
Pubblicazione: (2026)
Rating Roulette: Self-Inconsistency in LLM-As-A-Judge Frameworks
di: Haldar, Rajarshi, et al.
Pubblicazione: (2025)
di: Haldar, Rajarshi, et al.
Pubblicazione: (2025)
Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges
di: Ding, Ruomeng, et al.
Pubblicazione: (2026)
di: Ding, Ruomeng, et al.
Pubblicazione: (2026)
Mitigating Translationese Bias in Multilingual LLM-as-a-Judge via Disentangled Information Bottleneck
di: Zhang, Hongbin, et al.
Pubblicazione: (2026)
di: Zhang, Hongbin, et al.
Pubblicazione: (2026)
Judging with Confidence: Calibrating Autoraters to Preference Distributions
di: Li, Zhuohang, et al.
Pubblicazione: (2025)
di: Li, Zhuohang, et al.
Pubblicazione: (2025)
FairJudge: An Adaptive, Debiased, and Consistent LLM-as-a-Judge
di: Yang, Bo, et al.
Pubblicazione: (2026)
di: Yang, Bo, et al.
Pubblicazione: (2026)
Who Judges the Judge? Evaluating LLM-as-a-Judge for French Medical open-ended QA
di: Belmadani, Ikram, et al.
Pubblicazione: (2026)
di: Belmadani, Ikram, et al.
Pubblicazione: (2026)
Decoding Biases: Automated Methods and LLM Judges for Gender Bias Detection in Language Models
di: Kumar, Shachi H, et al.
Pubblicazione: (2024)
di: Kumar, Shachi H, et al.
Pubblicazione: (2024)
Beyond the Surface: Measuring Self-Preference in LLM Judgments
di: Chen, Zhi-Yuan, et al.
Pubblicazione: (2025)
di: Chen, Zhi-Yuan, et al.
Pubblicazione: (2025)
Can LLM be a Personalized Judge?
di: Dong, Yijiang River, et al.
Pubblicazione: (2024)
di: Dong, Yijiang River, et al.
Pubblicazione: (2024)
Self-Preference Bias in Rubric-Based Evaluation of Large Language Models
di: Pombal, José, et al.
Pubblicazione: (2026)
di: Pombal, José, et al.
Pubblicazione: (2026)
JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems
di: Bellibatlu, Rohith Reddy, et al.
Pubblicazione: (2026)
di: Bellibatlu, Rohith Reddy, et al.
Pubblicazione: (2026)
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
di: Xu, Wenda, et al.
Pubblicazione: (2024)
di: Xu, Wenda, et al.
Pubblicazione: (2024)
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
di: Wang, Qian, et al.
Pubblicazione: (2025)
di: Wang, Qian, et al.
Pubblicazione: (2025)
Benchmarking Adversarial Robustness to Bias Elicitation in Large Language Models: Scalable Automated Assessment with LLM-as-a-Judge
di: Cantini, Riccardo, et al.
Pubblicazione: (2025)
di: Cantini, Riccardo, et al.
Pubblicazione: (2025)
RankJudge: A Multi-Turn LLM-as-a-Judge Synthetic Benchmark Generator
di: Tang, Zhenwei, et al.
Pubblicazione: (2026)
di: Tang, Zhenwei, et al.
Pubblicazione: (2026)
MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge
di: Lee, Sua, et al.
Pubblicazione: (2026)
di: Lee, Sua, et al.
Pubblicazione: (2026)
The Necessity of Setting Temperature in LLM-as-a-Judge
di: Li, Lujun, et al.
Pubblicazione: (2026)
di: Li, Lujun, et al.
Pubblicazione: (2026)
How Reliable is Multilingual LLM-as-a-Judge?
di: Fu, Xiyan, et al.
Pubblicazione: (2025)
di: Fu, Xiyan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Self-Translate-Train: Enhancing Cross-Lingual Transfer of Large Language Models via Inherent Capability
di: Ri, Ryokan, et al.
Pubblicazione: (2024) -
LEIA: Facilitating Cross-lingual Knowledge Transfer in Language Models with Entity-based Data Augmentation
di: Yamada, Ikuya, et al.
Pubblicazione: (2024) -
Natural Fingerprints of Large Language Models
di: Suzuki, Teppei, et al.
Pubblicazione: (2025) -
Quantifying and Mitigating Self-Preference Bias of LLM Judges
di: Yang, Jinming, et al.
Pubblicazione: (2026) -
Large Vocabulary Size Improves Large Language Models
di: Takase, Sho, et al.
Pubblicazione: (2024)