Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mammen, Priyanka Mary, Joswin, Emil, Venkitachalam, Shankar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aligning Language Models with Clinical Expertise: DPO for Heart Failure Nursing Documentation in Critical Care
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
LIBRA: Measuring Bias of Large Language Model from a Local Context
von: Pang, Bo, et al.
Veröffentlicht: (2025)
von: Pang, Bo, et al.
Veröffentlicht: (2025)
Bias Similarity Measurement: A Black-Box Audit of Fairness Across LLMs
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
A Representation-Level Assessment of Bias Mitigation in Foundation Models
von: Nizhnichenkov, Svetoslav, et al.
Veröffentlicht: (2026)
von: Nizhnichenkov, Svetoslav, et al.
Veröffentlicht: (2026)
Having Beer after Prayer? Measuring Cultural Bias in Large Language Models
von: Naous, Tarek, et al.
Veröffentlicht: (2023)
von: Naous, Tarek, et al.
Veröffentlicht: (2023)
Dynamic Subset Tuning: Expanding the Operational Range of Parameter-Efficient Training for Large Language Models
von: Stahlberg, Felix, et al.
Veröffentlicht: (2024)
von: Stahlberg, Felix, et al.
Veröffentlicht: (2024)
What if I ask in \textit{alia lingua}? Measuring Functional Similarity Across Languages
von: Mishra, Debangan, et al.
Veröffentlicht: (2025)
von: Mishra, Debangan, et al.
Veröffentlicht: (2025)
Probing Cultural Signals in Large Language Models through Author Profiling
von: Lafargue, Valentin, et al.
Veröffentlicht: (2026)
von: Lafargue, Valentin, et al.
Veröffentlicht: (2026)
Race, Ethnicity and Their Implication on Bias in Large Language Models
von: Hu, Shiyue, et al.
Veröffentlicht: (2026)
von: Hu, Shiyue, et al.
Veröffentlicht: (2026)
Bias in Large Language Models: Origin, Evaluation, and Mitigation
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
How Quantization Shapes Bias in Large Language Models
von: Marcuzzi, Federico, et al.
Veröffentlicht: (2025)
von: Marcuzzi, Federico, et al.
Veröffentlicht: (2025)
An Analysis for Reasoning Bias of Language Models with Small Initialization
von: Yao, Junjie, et al.
Veröffentlicht: (2025)
von: Yao, Junjie, et al.
Veröffentlicht: (2025)
Transferring Linear Features Across Language Models With Model Stitching
von: Chen, Alan, et al.
Veröffentlicht: (2025)
von: Chen, Alan, et al.
Veröffentlicht: (2025)
Exploring Gender Bias in Large Language Models: An In-depth Dive into the German Language
von: Gnadt, Kristin, et al.
Veröffentlicht: (2025)
von: Gnadt, Kristin, et al.
Veröffentlicht: (2025)
Failing to Falsify: Evaluating and Mitigating Confirmation Bias in Language Models
von: Jhaveri, Ayush Rajesh, et al.
Veröffentlicht: (2026)
von: Jhaveri, Ayush Rajesh, et al.
Veröffentlicht: (2026)
Bias after Prompting: Persistent Discrimination in Large Language Models
von: Sivakumar, Nivedha, et al.
Veröffentlicht: (2025)
von: Sivakumar, Nivedha, et al.
Veröffentlicht: (2025)
Red-Teaming for Inducing Societal Bias in Large Language Models
von: Luo, Chu Fei, et al.
Veröffentlicht: (2024)
von: Luo, Chu Fei, et al.
Veröffentlicht: (2024)
Eliminating Position Bias of Language Models: A Mechanistic Approach
von: Wang, Ziqi, et al.
Veröffentlicht: (2024)
von: Wang, Ziqi, et al.
Veröffentlicht: (2024)
Can Large Language Models Generalize Procedures Across Representations?
von: Lin, Fangru, et al.
Veröffentlicht: (2026)
von: Lin, Fangru, et al.
Veröffentlicht: (2026)
Generalizing Large Language Model Usability Across Resource-Constrained
von: Tsai, Yun-Da
Veröffentlicht: (2025)
von: Tsai, Yun-Da
Veröffentlicht: (2025)
Circuit Component Reuse Across Tasks in Transformer Language Models
von: Merullo, Jack, et al.
Veröffentlicht: (2023)
von: Merullo, Jack, et al.
Veröffentlicht: (2023)
P^3SUM: Preserving Author's Perspective in News Summarization with Diffusion Language Models
von: Liu, Yuhan, et al.
Veröffentlicht: (2023)
von: Liu, Yuhan, et al.
Veröffentlicht: (2023)
Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training
von: Rahmati, Elnaz, et al.
Veröffentlicht: (2026)
von: Rahmati, Elnaz, et al.
Veröffentlicht: (2026)
Speculative Decoding Across Languages
von: Paudel, Nirajan, et al.
Veröffentlicht: (2026)
von: Paudel, Nirajan, et al.
Veröffentlicht: (2026)
Twin-Merging: Dynamic Integration of Modular Expertise in Model Merging
von: Lu, Zhenyi, et al.
Veröffentlicht: (2024)
von: Lu, Zhenyi, et al.
Veröffentlicht: (2024)
REFLEX: Reference-Free Evaluation of Log Summarization via Large Language Model Judgment
von: Mudgal, Priyanka
Veröffentlicht: (2025)
von: Mudgal, Priyanka
Veröffentlicht: (2025)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
von: Das, Amit, et al.
Veröffentlicht: (2024)
von: Das, Amit, et al.
Veröffentlicht: (2024)
Faithfulness Measurable Masked Language Models
von: Madsen, Andreas, et al.
Veröffentlicht: (2023)
von: Madsen, Andreas, et al.
Veröffentlicht: (2023)
On the Inductive Bias of Stacking Towards Improving Reasoning
von: Saunshi, Nikunj, et al.
Veröffentlicht: (2024)
von: Saunshi, Nikunj, et al.
Veröffentlicht: (2024)
Predicting Compact Phrasal Rewrites with Large Language Models for ASR Post Editing
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
Neuron-Level Knowledge Attribution in Large Language Models
von: Yu, Zeping, et al.
Veröffentlicht: (2023)
von: Yu, Zeping, et al.
Veröffentlicht: (2023)
Simultaneous Reward Distillation and Preference Learning: Get You a Language Model Who Can Do Both
von: Nath, Abhijnan, et al.
Veröffentlicht: (2024)
von: Nath, Abhijnan, et al.
Veröffentlicht: (2024)
Measuring Moral Inconsistencies in Large Language Models
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2024)
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2024)
Omitted Variable Bias in Language Models Under Distribution Shift
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
Bias in LLMs as Annotators: The Effect of Party Cues on Labelling Decision by Large Language Models
von: Vera, Sebastian Vallejo, et al.
Veröffentlicht: (2024)
von: Vera, Sebastian Vallejo, et al.
Veröffentlicht: (2024)
Prompt-Based Bias Calibration for Better Zero/Few-Shot Learning of Language Models
von: He, Kang, et al.
Veröffentlicht: (2024)
von: He, Kang, et al.
Veröffentlicht: (2024)
Mitigate Position Bias in Large Language Models via Scaling a Single Dimension
von: Yu, Yijiong, et al.
Veröffentlicht: (2024)
von: Yu, Yijiong, et al.
Veröffentlicht: (2024)
Understanding and Mitigating Tokenization Bias in Language Models
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
On the "Induction Bias" in Sequence Models
von: Ebrahimi, M. Reza, et al.
Veröffentlicht: (2026)
von: Ebrahimi, M. Reza, et al.
Veröffentlicht: (2026)
Fair Representation in Parliamentary Summaries: Measuring and Mitigating Inclusion Bias
von: Cunningham, Eoghan, et al.
Veröffentlicht: (2025)
von: Cunningham, Eoghan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Aligning Language Models with Clinical Expertise: DPO for Heart Failure Nursing Documentation in Critical Care
von: Fan, Junyi, et al.
Veröffentlicht: (2025) -
LIBRA: Measuring Bias of Large Language Model from a Local Context
von: Pang, Bo, et al.
Veröffentlicht: (2025) -
Bias Similarity Measurement: A Black-Box Audit of Fairness Across LLMs
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024) -
A Representation-Level Assessment of Bias Mitigation in Foundation Models
von: Nizhnichenkov, Svetoslav, et al.
Veröffentlicht: (2026) -
Having Beer after Prayer? Measuring Cultural Bias in Large Language Models
von: Naous, Tarek, et al.
Veröffentlicht: (2023)