Faithful or Fabricated? A Causal Framework for Rationalization Bias in LLM Judges
Fuente:
arXiv
Guardado en:
| Autores principales: | Tapwal, Riya, Kumar, Abhishek, Maple, Carsten |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DriveSafe: A Hierarchical Risk Taxonomy for Safety-Critical LLM-Based Driving Assistants
por: Kumar, Abhishek, et al.
Publicado: (2026)
por: Kumar, Abhishek, et al.
Publicado: (2026)
PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines
por: Tapwal, Riya, et al.
Publicado: (2026)
por: Tapwal, Riya, et al.
Publicado: (2026)
Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success
por: Maple, Carsten, et al.
Publicado: (2026)
por: Maple, Carsten, et al.
Publicado: (2026)
Field-Localized Forgery Detection for Digital Identity Documents
por: Kumar, Abhishek, et al.
Publicado: (2026)
por: Kumar, Abhishek, et al.
Publicado: (2026)
C2-Faith: Benchmarking LLM Judges for Causal and Coverage Faithfulness in Chain-of-Thought Reasoning
por: Mittal, Avni, et al.
Publicado: (2026)
por: Mittal, Avni, et al.
Publicado: (2026)
The Silent Judge: Unacknowledged Shortcut Bias in LLM-as-a-Judge
por: Marioriyad, Arash, et al.
Publicado: (2025)
por: Marioriyad, Arash, et al.
Publicado: (2025)
Judging the Judges: A Systematic Study of Position Bias in LLM-as-a-Judge
por: Shi, Lin, et al.
Publicado: (2024)
por: Shi, Lin, et al.
Publicado: (2024)
CyclicJudge: Mitigating Judge Bias Efficiently in LLM-based Evaluation
por: Zhu, Ziyi, et al.
Publicado: (2026)
por: Zhu, Ziyi, et al.
Publicado: (2026)
Evaluating Scoring Bias in LLM-as-a-Judge
por: Li, Qingquan, et al.
Publicado: (2025)
por: Li, Qingquan, et al.
Publicado: (2025)
Self-Preference Bias in LLM-as-a-Judge
por: Wataoka, Koki, et al.
Publicado: (2024)
por: Wataoka, Koki, et al.
Publicado: (2024)
Exploring Causal Effect of Social Bias on Faithfulness Hallucinations in Large Language Models
por: Zhang, Zhenliang, et al.
Publicado: (2025)
por: Zhang, Zhenliang, et al.
Publicado: (2025)
Are More Tokens Rational? Inference-Time Scaling in Language Models as Adaptive Resource Rationality
por: Hu, Zhimin, et al.
Publicado: (2026)
por: Hu, Zhimin, et al.
Publicado: (2026)
Assistant-Guided Mitigation of Teacher Preference Bias in LLM-as-a-Judge
por: Liu, Zhuo, et al.
Publicado: (2025)
por: Liu, Zhuo, et al.
Publicado: (2025)
On Positional Bias of Faithfulness for Long-form Summarization
por: Wan, David, et al.
Publicado: (2024)
por: Wan, David, et al.
Publicado: (2024)
Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization
por: Zhou, Hongli, et al.
Publicado: (2026)
por: Zhou, Hongli, et al.
Publicado: (2026)
Decoding Biases: Automated Methods and LLM Judges for Gender Bias Detection in Language Models
por: Kumar, Shachi H, et al.
Publicado: (2024)
por: Kumar, Shachi H, et al.
Publicado: (2024)
Quantifying and Mitigating Self-Preference Bias of LLM Judges
por: Yang, Jinming, et al.
Publicado: (2026)
por: Yang, Jinming, et al.
Publicado: (2026)
BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation
por: Lai, Peng, et al.
Publicado: (2026)
por: Lai, Peng, et al.
Publicado: (2026)
Fairness or Fluency? An Investigation into Language Bias of Pairwise LLM-as-a-Judge
por: Zhou, Xiaolin, et al.
Publicado: (2026)
por: Zhou, Xiaolin, et al.
Publicado: (2026)
Rating Roulette: Self-Inconsistency in LLM-As-A-Judge Frameworks
por: Haldar, Rajarshi, et al.
Publicado: (2025)
por: Haldar, Rajarshi, et al.
Publicado: (2025)
Contrastive Decoding Mitigates Score Range Bias in LLM-as-a-Judge
por: Fujinuma, Yoshinari
Publicado: (2025)
por: Fujinuma, Yoshinari
Publicado: (2025)
Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge
por: Spiliopoulou, Evangelia, et al.
Publicado: (2025)
por: Spiliopoulou, Evangelia, et al.
Publicado: (2025)
ClickGuard: A Trustworthy Adaptive Fusion Framework for Clickbait Detection
por: Dhiman, Chhavi, et al.
Publicado: (2026)
por: Dhiman, Chhavi, et al.
Publicado: (2026)
eDIF: A European Deep Inference Fabric for Remote Interpretability of LLM
por: Guggenberger, Irma Heithoff. Marc, et al.
Publicado: (2025)
por: Guggenberger, Irma Heithoff. Marc, et al.
Publicado: (2025)
Am I More Pointwise or Pairwise? Revealing Position Bias in Rubric-Based LLM-as-a-Judge
por: Xu, Yuzheng, et al.
Publicado: (2026)
por: Xu, Yuzheng, et al.
Publicado: (2026)
Mitigating Translationese Bias in Multilingual LLM-as-a-Judge via Disentangled Information Bottleneck
por: Zhang, Hongbin, et al.
Publicado: (2026)
por: Zhang, Hongbin, et al.
Publicado: (2026)
NeuroFaith: Evaluating LLM Self-Explanation Faithfulness via Internal Representation Alignment
por: Bhan, Milan, et al.
Publicado: (2025)
por: Bhan, Milan, et al.
Publicado: (2025)
LLM-as-a-Judge for Time Series Explanations
por: Sivalingam, Preetham, et al.
Publicado: (2026)
por: Sivalingam, Preetham, et al.
Publicado: (2026)
FairJudge: An Adaptive, Debiased, and Consistent LLM-as-a-Judge
por: Yang, Bo, et al.
Publicado: (2026)
por: Yang, Bo, et al.
Publicado: (2026)
PRECISE: Reducing the Bias of LLM Evaluations Using Prediction-Powered Ranking Estimation
por: Divekar, Abhishek, et al.
Publicado: (2026)
por: Divekar, Abhishek, et al.
Publicado: (2026)
JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems
por: Bellibatlu, Rohith Reddy, et al.
Publicado: (2026)
por: Bellibatlu, Rohith Reddy, et al.
Publicado: (2026)
A Causal Lens for Evaluating Faithfulness Metrics
por: Zaman, Kerem, et al.
Publicado: (2025)
por: Zaman, Kerem, et al.
Publicado: (2025)
Who Judges the Judge? Evaluating LLM-as-a-Judge for French Medical open-ended QA
por: Belmadani, Ikram, et al.
Publicado: (2026)
por: Belmadani, Ikram, et al.
Publicado: (2026)
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
por: Wang, Qian, et al.
Publicado: (2025)
por: Wang, Qian, et al.
Publicado: (2025)
RankJudge: A Multi-Turn LLM-as-a-Judge Synthetic Benchmark Generator
por: Tang, Zhenwei, et al.
Publicado: (2026)
por: Tang, Zhenwei, et al.
Publicado: (2026)
MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge
por: Lee, Sua, et al.
Publicado: (2026)
por: Lee, Sua, et al.
Publicado: (2026)
Are LLM Decisions Faithful to Verbal Confidence?
por: Wang, Jiawei, et al.
Publicado: (2026)
por: Wang, Jiawei, et al.
Publicado: (2026)
A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework
por: Li, Chenyu, et al.
Publicado: (2026)
por: Li, Chenyu, et al.
Publicado: (2026)
Bias Association Discovery Framework for Open-Ended LLM Generations
por: Pan, Jinhao, et al.
Publicado: (2025)
por: Pan, Jinhao, et al.
Publicado: (2025)
Quantitative LLM Judges
por: Sahoo, Aishwarya, et al.
Publicado: (2025)
por: Sahoo, Aishwarya, et al.
Publicado: (2025)
Ejemplares similares
-
DriveSafe: A Hierarchical Risk Taxonomy for Safety-Critical LLM-Based Driving Assistants
por: Kumar, Abhishek, et al.
Publicado: (2026) -
PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines
por: Tapwal, Riya, et al.
Publicado: (2026) -
Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success
por: Maple, Carsten, et al.
Publicado: (2026) -
Field-Localized Forgery Detection for Digital Identity Documents
por: Kumar, Abhishek, et al.
Publicado: (2026) -
C2-Faith: Benchmarking LLM Judges for Causal and Coverage Faithfulness in Chain-of-Thought Reasoning
por: Mittal, Avni, et al.
Publicado: (2026)