Gespeichert in:
| Hauptverfasser: | Zhou, Ej, Lu, Weiming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2504.11183 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
OffsetBias: Leveraging Debiased Data for Tuning Evaluators
von: Park, Junsoo, et al.
Veröffentlicht: (2024)
von: Park, Junsoo, et al.
Veröffentlicht: (2024)
Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization
von: Zhou, Hongli, et al.
Veröffentlicht: (2026)
von: Zhou, Hongli, et al.
Veröffentlicht: (2026)
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
von: Chen, Yupeng, et al.
Veröffentlicht: (2025)
von: Chen, Yupeng, et al.
Veröffentlicht: (2025)
On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection
von: Elsafoury, Fatma, et al.
Veröffentlicht: (2023)
von: Elsafoury, Fatma, et al.
Veröffentlicht: (2023)
Evaluating and Mitigating Social Bias for Large Language Models in Open-ended Settings
von: Liu, Zhao, et al.
Veröffentlicht: (2024)
von: Liu, Zhao, et al.
Veröffentlicht: (2024)
Debiasing CLIP: Interpreting and Correcting Bias in Attention Heads
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
Are Bias Evaluation Methods Biased ?
von: Berrayana, Lina, et al.
Veröffentlicht: (2025)
von: Berrayana, Lina, et al.
Veröffentlicht: (2025)
Towards Multimodal Sentiment Analysis Debiasing via Bias Purification
von: Yang, Dingkang, et al.
Veröffentlicht: (2024)
von: Yang, Dingkang, et al.
Veröffentlicht: (2024)
BiasFilter: An Inference-Time Debiasing Framework for Large Language Models
von: Cheng, Xiaoqing, et al.
Veröffentlicht: (2025)
von: Cheng, Xiaoqing, et al.
Veröffentlicht: (2025)
Unboxing Occupational Bias: Grounded Debiasing of LLMs with U.S. Labor Data
von: Gorti, Atmika, et al.
Veröffentlicht: (2024)
von: Gorti, Atmika, et al.
Veröffentlicht: (2024)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
von: Du, Yongkang, et al.
Veröffentlicht: (2025)
von: Du, Yongkang, et al.
Veröffentlicht: (2025)
Towards Understanding Task-agnostic Debiasing Through the Lenses of Intrinsic Bias and Forgetfulness
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
Open-DeBias: Toward Mitigating Open-Set Bias in Language Models
von: Rani, Arti, et al.
Veröffentlicht: (2025)
von: Rani, Arti, et al.
Veröffentlicht: (2025)
Gender Bias in English-to-Greek Machine Translation
von: Gkovedarou, Eleni, et al.
Veröffentlicht: (2025)
von: Gkovedarou, Eleni, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Beyond the Final Layer: Intermediate Representations for Better Multilingual Calibration in Large Language Models
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
Religious Bias Landscape in Language and Text-to-Image Models: Analysis, Detection, and Debiasing Strategies
von: Abrar, Ajwad, et al.
Veröffentlicht: (2025)
von: Abrar, Ajwad, et al.
Veröffentlicht: (2025)
Mind the Language Gap: Automated and Augmented Evaluation of Bias in LLMs for High- and Low-Resource Languages
von: Buscemi, Alessio, et al.
Veröffentlicht: (2025)
von: Buscemi, Alessio, et al.
Veröffentlicht: (2025)
Does Reasoning Introduce Bias? A Study of Social Bias Evaluation and Mitigation in LLM Reasoning
von: Wu, Xuyang, et al.
Veröffentlicht: (2025)
von: Wu, Xuyang, et al.
Veröffentlicht: (2025)
Evaluating Social Bias in RAG Systems: When External Context Helps and Reasoning Hurts
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
Trustworthy Social Bias Measurement
von: Bommasani, Rishi, et al.
Veröffentlicht: (2022)
von: Bommasani, Rishi, et al.
Veröffentlicht: (2022)
VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models
von: Raj, Chahat, et al.
Veröffentlicht: (2025)
von: Raj, Chahat, et al.
Veröffentlicht: (2025)
FIBER: A Multilingual Evaluation Resource for Factual Inference Bias
von: Munis, Evren Ayberk, et al.
Veröffentlicht: (2025)
von: Munis, Evren Ayberk, et al.
Veröffentlicht: (2025)
From Measurement to Mitigation: Exploring the Transferability of Debiasing Approaches to Gender Bias in Maltese Language Models
von: Galea, Melanie, et al.
Veröffentlicht: (2025)
von: Galea, Melanie, et al.
Veröffentlicht: (2025)
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs
von: Fan, Zhiting, et al.
Veröffentlicht: (2024)
von: Fan, Zhiting, et al.
Veröffentlicht: (2024)
Any Large Language Model Can Be a Reliable Judge: Debiasing with a Reasoning-based Bias Detector
von: Yang, Haoyan, et al.
Veröffentlicht: (2025)
von: Yang, Haoyan, et al.
Veröffentlicht: (2025)
Bias in Language Models: Beyond Trick Tests and Toward RUTEd Evaluation
von: Lum, Kristian, et al.
Veröffentlicht: (2024)
von: Lum, Kristian, et al.
Veröffentlicht: (2024)
Bi-directional Bias Attribution: Debiasing Large Language Models without Modifying Prompts
von: Lin, Yujie, et al.
Veröffentlicht: (2026)
von: Lin, Yujie, et al.
Veröffentlicht: (2026)
Towards Resource Efficient and Interpretable Bias Mitigation in Large Language Models
von: Tong, Schrasing, et al.
Veröffentlicht: (2024)
von: Tong, Schrasing, et al.
Veröffentlicht: (2024)
Social Bias Evaluation for Large Language Models Requires Prompt Variations
von: Hida, Rem, et al.
Veröffentlicht: (2024)
von: Hida, Rem, et al.
Veröffentlicht: (2024)
Exploring Gender Bias Beyond Occupational Titles
von: Sabir, Ahmed, et al.
Veröffentlicht: (2025)
von: Sabir, Ahmed, et al.
Veröffentlicht: (2025)
Social Bias in Large Language Models For Bangla: An Empirical Study on Gender and Religious Bias
von: Sadhu, Jayanta, et al.
Veröffentlicht: (2024)
von: Sadhu, Jayanta, et al.
Veröffentlicht: (2024)
Mitigating the Bias of Large Language Model Evaluation
von: Zhou, Hongli, et al.
Veröffentlicht: (2024)
von: Zhou, Hongli, et al.
Veröffentlicht: (2024)
Bias Beyond Borders: Political Ideology Evaluation and Steering in Multilingual LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
Veracity Bias and Beyond: Uncovering LLMs' Hidden Beliefs in Problem-Solving Reasoning
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
Bias Dynamics in BabyLMs: Towards a Compute-Efficient Sandbox for Democratising Pre-Training Debiasing
von: Trhlik, Filip, et al.
Veröffentlicht: (2026)
von: Trhlik, Filip, et al.
Veröffentlicht: (2026)
Mitigating Social Bias in English and Urdu Language Models Using PRM-Guided Candidate Selection and Sequential Refinement
von: Khan, Muneeb Ur Raheem
Veröffentlicht: (2025)
von: Khan, Muneeb Ur Raheem
Veröffentlicht: (2025)
Evaluating Scoring Bias in LLM-as-a-Judge
von: Li, Qingquan, et al.
Veröffentlicht: (2025)
von: Li, Qingquan, et al.
Veröffentlicht: (2025)
BiasCause: Evaluate Socially Biased Causal Reasoning of Large Language Models
von: Xie, Tian, et al.
Veröffentlicht: (2025)
von: Xie, Tian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024) -
OffsetBias: Leveraging Debiased Data for Tuning Evaluators
von: Park, Junsoo, et al.
Veröffentlicht: (2024) -
Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization
von: Zhou, Hongli, et al.
Veröffentlicht: (2026) -
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
von: Chen, Yupeng, et al.
Veröffentlicht: (2025) -
On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection
von: Elsafoury, Fatma, et al.
Veröffentlicht: (2023)