Justice in Judgment: Unveiling (Hidden) Bias in LLM-assisted Peer Reviews
Fuente:
arXiv
Saved in:
| Main Authors: | Vasu, Sai Suresh Macharla, Sheth, Ivaxi, Wang, Hui-Po, Binkyte, Ruta, Fritz, Mario |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safety Must Precede the Deployment of Open-Ended AI
by: Sheth, Ivaxi, et al.
Published: (2025)
by: Sheth, Ivaxi, et al.
Published: (2025)
Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution
by: Binkyte, Ruta, et al.
Published: (2026)
by: Binkyte, Ruta, et al.
Published: (2026)
LLM4GRN: Discovering Causal Gene Regulatory Networks with LLMs -- Evaluation through Synthetic Data Generation
by: Afonja, Tejumade, et al.
Published: (2024)
by: Afonja, Tejumade, et al.
Published: (2024)
Causality Is Key to Understand and Balance Multiple Goals in Trustworthy ML and Foundation Models
by: Binkyte, Ruta, et al.
Published: (2025)
by: Binkyte, Ruta, et al.
Published: (2025)
Safe for Whom? Rethinking How We Evaluate the Safety of LLMs for Real Users
by: Kempermann, Manon, et al.
Published: (2025)
by: Kempermann, Manon, et al.
Published: (2025)
Interactional Fairness in LLM Multi-Agent Systems: An Evaluation Framework
by: Binkyte, Ruta
Published: (2025)
by: Binkyte, Ruta
Published: (2025)
ProtocolLLM: RTL Benchmark for SystemVerilog Generation of Communication Protocols
by: Sheth, Arnav, et al.
Published: (2025)
by: Sheth, Arnav, et al.
Published: (2025)
BaBE: Enhancing Fairness via Estimation of Latent Explaining Variables
by: Binkyte, Ruta, et al.
Published: (2023)
by: Binkyte, Ruta, et al.
Published: (2023)
On the Need and Applicability of Causality for Fairness: A Unified Framework for AI Auditing and Legal Analysis
by: Binkyte, Ruta, et al.
Published: (2022)
by: Binkyte, Ruta, et al.
Published: (2022)
AuditCopilot: Leveraging LLMs for Fraud Detection in Double-Entry Bookkeeping
by: Kadir, Md Abdul, et al.
Published: (2025)
by: Kadir, Md Abdul, et al.
Published: (2025)
IV Co-Scientist: Multi-Agent LLM Framework for Causal Instrumental Variable Discovery
by: Sheth, Ivaxi, et al.
Published: (2026)
by: Sheth, Ivaxi, et al.
Published: (2026)
Survey on AI Ethics: A Socio-technical Perspective
by: Mbiazi, Dave, et al.
Published: (2023)
by: Mbiazi, Dave, et al.
Published: (2023)
Hidden in Memory: Sleeper Memory Poisoning in LLM Agents
by: Pulipaka, Sidharth, et al.
Published: (2026)
by: Pulipaka, Sidharth, et al.
Published: (2026)
Inspectable AI for Science: A Research Object Approach to Generative AI Governance
by: Binkyte, Ruta, et al.
Published: (2026)
by: Binkyte, Ruta, et al.
Published: (2026)
Funny or Persuasive, but Not Both: Evaluating Fine-Grained Multi-Concept Control in LLMs
by: Labroo, Arya, et al.
Published: (2026)
by: Labroo, Arya, et al.
Published: (2026)
Gender Bias of LLM in Economics: An Existentialism Perspective
by: Zhong, Hui, et al.
Published: (2024)
by: Zhong, Hui, et al.
Published: (2024)
Hidden Prompts in Manuscripts Exploit AI-Assisted Peer Review
by: Lin, Zhicheng
Published: (2025)
by: Lin, Zhicheng
Published: (2025)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
by: Fernandes, Gustavo Lúcius, et al.
Published: (2026)
by: Fernandes, Gustavo Lúcius, et al.
Published: (2026)
Sparks of Rationality: Do Reasoning LLMs Align with Human Judgment and Choice?
by: Tak, Ala N., et al.
Published: (2026)
by: Tak, Ala N., et al.
Published: (2026)
Implicit Humanization in Everyday LLM Moral Judgments
by: Ayad, Hoda, et al.
Published: (2026)
by: Ayad, Hoda, et al.
Published: (2026)
Insights from the ICLR Peer Review and Rebuttal Process
by: Kargaran, Amir Hossein, et al.
Published: (2025)
by: Kargaran, Amir Hossein, et al.
Published: (2025)
Position on LLM-Assisted Peer Review: Addressing Reviewer Gap through Mentoring and Feedback
by: Yun, JungMin, et al.
Published: (2026)
by: Yun, JungMin, et al.
Published: (2026)
Peer Identity Bias in Multi-Agent LLM Evaluation: An Empirical Study Using the TRUST Democratic Discourse Analysis Pipeline
by: Dietrich, Juergen
Published: (2026)
by: Dietrich, Juergen
Published: (2026)
Position: The AI Conference Peer Review Crisis Demands Author Feedback and Reviewer Rewards
by: Kim, Jaeho, et al.
Published: (2025)
by: Kim, Jaeho, et al.
Published: (2025)
Do LLMs Favor LLMs? Quantifying Interaction Effects in Peer Review
by: Sharma, Vibhhu, et al.
Published: (2026)
by: Sharma, Vibhhu, et al.
Published: (2026)
Policies Permitting LLM Use for Polishing Peer Reviews Are Currently Not Enforceable
by: Saha, Rounak, et al.
Published: (2026)
by: Saha, Rounak, et al.
Published: (2026)
Patentformer: A demonstration of AI-assisted automated patent drafting
by: Mudhiganti, Sai Krishna Reddy, et al.
Published: (2025)
by: Mudhiganti, Sai Krishna Reddy, et al.
Published: (2025)
Revolutionizing Pharma: Unveiling the AI and LLM Trends in the Pharmaceutical Industry
by: Han, Yu, et al.
Published: (2024)
by: Han, Yu, et al.
Published: (2024)
When Agents See Humans as the Outgroup: Belief-Dependent Bias in LLM-Powered Agents
by: Wang, Zongwei, et al.
Published: (2026)
by: Wang, Zongwei, et al.
Published: (2026)
Legal Fact Prediction: The Missing Piece in Legal Judgment Prediction
by: Liu, Junkai, et al.
Published: (2024)
by: Liu, Junkai, et al.
Published: (2024)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review
by: Weng, Yixuan, et al.
Published: (2026)
by: Weng, Yixuan, et al.
Published: (2026)
The Good, the Bad and the Constructive: Automatically Measuring Peer Review's Utility for Authors
by: Sadallah, Abdelrahman, et al.
Published: (2025)
by: Sadallah, Abdelrahman, et al.
Published: (2025)
The Switch, the Ladder, and the Matrix: Models for Classifying AI Systems
by: Mokander, Jakob, et al.
Published: (2024)
by: Mokander, Jakob, et al.
Published: (2024)
Between Rules and Reality: On the Context Sensitivity of LLM Moral Judgment
by: Sauter, Adrian, et al.
Published: (2026)
by: Sauter, Adrian, et al.
Published: (2026)
Street-Level AI: Are Large Language Models Ready for Real-World Judgments?
by: Pokharel, Gaurab, et al.
Published: (2025)
by: Pokharel, Gaurab, et al.
Published: (2025)
Diagnosing Korean-Language LLM Political Bias via Census-Grounded Agent Simulation
by: Kang, Sungwoo
Published: (2026)
by: Kang, Sungwoo
Published: (2026)
Echoes of AI Harms: A Human-LLM Synergistic Framework for Bias-Driven Harm Anticipation
by: Tantalaki, Nicoleta, et al.
Published: (2025)
by: Tantalaki, Nicoleta, et al.
Published: (2025)
From Delegates to Trustees: How Optimizing for Long-Term Interests Shapes Bias and Alignment in LLM
by: Fulay, Suyash, et al.
Published: (2025)
by: Fulay, Suyash, et al.
Published: (2025)
GLARE: Agentic Reasoning for Legal Judgment Prediction
by: Yang, Xinyu, et al.
Published: (2025)
by: Yang, Xinyu, et al.
Published: (2025)
Similar Items
-
Safety Must Precede the Deployment of Open-Ended AI
by: Sheth, Ivaxi, et al.
Published: (2025) -
Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution
by: Binkyte, Ruta, et al.
Published: (2026) -
LLM4GRN: Discovering Causal Gene Regulatory Networks with LLMs -- Evaluation through Synthetic Data Generation
by: Afonja, Tejumade, et al.
Published: (2024) -
Causality Is Key to Understand and Balance Multiple Goals in Trustworthy ML and Foundation Models
by: Binkyte, Ruta, et al.
Published: (2025) -
Safe for Whom? Rethinking How We Evaluate the Safety of LLMs for Real Users
by: Kempermann, Manon, et al.
Published: (2025)