On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Elsafoury, Fatma, Katsigiannis, Stamos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Systematic Offensive Stereotyping (SOS) Bias in Language Models
von: Elsafoury, Fatma
Veröffentlicht: (2023)
von: Elsafoury, Fatma
Veröffentlicht: (2023)
Out of Sight Out of Mind, Out of Sight Out of Mind: Measuring Bias in Language Models Against Overlooked Marginalized Groups in Regional Contexts
von: Elsafoury, Fatma, et al.
Veröffentlicht: (2025)
von: Elsafoury, Fatma, et al.
Veröffentlicht: (2025)
SKDU at De-Factify 4.0: Natural Language Features for AI-Generated Text-Detection
von: Malviya, Shrikant, et al.
Veröffentlicht: (2025)
von: Malviya, Shrikant, et al.
Veröffentlicht: (2025)
Surface Fairness, Deep Bias: A Comparative Study of Bias in Language Models
von: Sorokovikova, Aleksandra, et al.
Veröffentlicht: (2025)
von: Sorokovikova, Aleksandra, et al.
Veröffentlicht: (2025)
Are Non-English Papers Reviewed Fairly? Language-of-Study Bias in NLP Peer Reviews
von: Barkhordar, Ehsan, et al.
Veröffentlicht: (2026)
von: Barkhordar, Ehsan, et al.
Veröffentlicht: (2026)
Social Bias Probing: Fairness Benchmarking for Language Models
von: Manerba, Marta Marchiori, et al.
Veröffentlicht: (2023)
von: Manerba, Marta Marchiori, et al.
Veröffentlicht: (2023)
Fairness or Fluency? An Investigation into Language Bias of Pairwise LLM-as-a-Judge
von: Zhou, Xiaolin, et al.
Veröffentlicht: (2026)
von: Zhou, Xiaolin, et al.
Veröffentlicht: (2026)
Fair-GPTQ: Bias-Aware Quantization for Large Language Models
von: Proskurina, Irina, et al.
Veröffentlicht: (2025)
von: Proskurina, Irina, et al.
Veröffentlicht: (2025)
Bias and Fairness in Large Language Models: A Survey
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2023)
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2023)
The Impact of Disability Disclosure on Fairness and Bias in LLM-Driven Candidate Selection
von: Kamruzzaman, Mahammed, et al.
Veröffentlicht: (2025)
von: Kamruzzaman, Mahammed, et al.
Veröffentlicht: (2025)
Intrinsic Meets Extrinsic Fairness: Assessing the Downstream Impact of Bias Mitigation in Large Language Models
von: Arzaghi', 'Mina, et al.
Veröffentlicht: (2025)
von: Arzaghi', 'Mina, et al.
Veröffentlicht: (2025)
A Trip Towards Fairness: Bias and De-Biasing in Large Language Models
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2023)
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2023)
Fairness and Bias in Multimodal AI: A Survey
von: Adewumi, Tosin, et al.
Veröffentlicht: (2024)
von: Adewumi, Tosin, et al.
Veröffentlicht: (2024)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
Religious Bias Landscape in Language and Text-to-Image Models: Analysis, Detection, and Debiasing Strategies
von: Abrar, Ajwad, et al.
Veröffentlicht: (2025)
von: Abrar, Ajwad, et al.
Veröffentlicht: (2025)
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
FairI Tales: Evaluation of Fairness in Indian Contexts with a Focus on Bias and Stereotypes
von: Nawale, Janki Atul, et al.
Veröffentlicht: (2025)
von: Nawale, Janki Atul, et al.
Veröffentlicht: (2025)
SKDU at De-Factify 4.0: Vision Transformer with Data Augmentation for AI-Generated Image Detection
von: Malviya, Shrikant, et al.
Veröffentlicht: (2025)
von: Malviya, Shrikant, et al.
Veröffentlicht: (2025)
Thinking Fair and Slow: On the Efficacy of Structured Prompts for Debiasing Language Models
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
BiasFilter: An Inference-Time Debiasing Framework for Large Language Models
von: Cheng, Xiaoqing, et al.
Veröffentlicht: (2025)
von: Cheng, Xiaoqing, et al.
Veröffentlicht: (2025)
Towards Fair Rankings: Leveraging LLMs for Gender Bias Detection and Measurement
von: Mousavian, Maryam, et al.
Veröffentlicht: (2025)
von: Mousavian, Maryam, et al.
Veröffentlicht: (2025)
Bias Neutralization Framework: Measuring Fairness in Large Language Models with Bias Intelligence Quotient (BiQ)
von: Narayan, Malur, et al.
Veröffentlicht: (2024)
von: Narayan, Malur, et al.
Veröffentlicht: (2024)
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models
von: Wang, Ze, et al.
Veröffentlicht: (2024)
von: Wang, Ze, et al.
Veröffentlicht: (2024)
Steering Towards Fairness: Mitigating Political Bias in LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
Persuasiveness and Bias in LLM: Investigating the Impact of Persuasiveness and Reinforcement of Bias in Language Models
von: Roy, Saumya
Veröffentlicht: (2025)
von: Roy, Saumya
Veröffentlicht: (2025)
Does Differential Privacy Impact Bias in Pretrained NLP Models?
von: Islam, Md. Khairul, et al.
Veröffentlicht: (2024)
von: Islam, Md. Khairul, et al.
Veröffentlicht: (2024)
TriCon-Fair: Triplet Contrastive Learning for Mitigating Social Bias in Pre-trained Language Models
von: Lyu, Chong, et al.
Veröffentlicht: (2025)
von: Lyu, Chong, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Unveiling the "Fairness Seesaw": Discovering and Mitigating Gender and Race Bias in Vision-Language Models
von: Lan, Jian, et al.
Veröffentlicht: (2025)
von: Lan, Jian, et al.
Veröffentlicht: (2025)
Fair Representation in Parliamentary Summaries: Measuring and Mitigating Inclusion Bias
von: Cunningham, Eoghan, et al.
Veröffentlicht: (2025)
von: Cunningham, Eoghan, et al.
Veröffentlicht: (2025)
Structured Reasoning for Fairness: A Multi-Agent Approach to Bias Detection in Textual Data
von: Huang, Tianyi, et al.
Veröffentlicht: (2025)
von: Huang, Tianyi, et al.
Veröffentlicht: (2025)
FairImagen: Post-Processing for Bias Mitigation in Text-to-Image Models
von: Fu, Zihao, et al.
Veröffentlicht: (2025)
von: Fu, Zihao, et al.
Veröffentlicht: (2025)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
von: Du, Yongkang, et al.
Veröffentlicht: (2025)
von: Du, Yongkang, et al.
Veröffentlicht: (2025)
Investigating Gender Bias in Turkish Language Models
von: Caglidil, Orhun Mersin, et al.
Veröffentlicht: (2024)
von: Caglidil, Orhun Mersin, et al.
Veröffentlicht: (2024)
Information Representation Fairness in Long-Document Embeddings: The Peculiar Interaction of Positional and Language Bias
von: Schuhmacher, Elias, et al.
Veröffentlicht: (2026)
von: Schuhmacher, Elias, et al.
Veröffentlicht: (2026)
Social Debiasing for Fair Multi-modal LLMs
von: Cheng, Harry, et al.
Veröffentlicht: (2024)
von: Cheng, Harry, et al.
Veröffentlicht: (2024)
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT
von: Reddy, Harishwar, et al.
Veröffentlicht: (2025)
von: Reddy, Harishwar, et al.
Veröffentlicht: (2025)
AXOLOTL: Fairness through Assisted Self-Debiasing of Large Language Model Outputs
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
OffsetBias: Leveraging Debiased Data for Tuning Evaluators
von: Park, Junsoo, et al.
Veröffentlicht: (2024)
von: Park, Junsoo, et al.
Veröffentlicht: (2024)
Bias Beyond English: Evaluating Social Bias and Debiasing Methods in a Low-Resource Setting
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Systematic Offensive Stereotyping (SOS) Bias in Language Models
von: Elsafoury, Fatma
Veröffentlicht: (2023) -
Out of Sight Out of Mind, Out of Sight Out of Mind: Measuring Bias in Language Models Against Overlooked Marginalized Groups in Regional Contexts
von: Elsafoury, Fatma, et al.
Veröffentlicht: (2025) -
SKDU at De-Factify 4.0: Natural Language Features for AI-Generated Text-Detection
von: Malviya, Shrikant, et al.
Veröffentlicht: (2025) -
Surface Fairness, Deep Bias: A Comparative Study of Bias in Language Models
von: Sorokovikova, Aleksandra, et al.
Veröffentlicht: (2025) -
Are Non-English Papers Reviewed Fairly? Language-of-Study Bias in NLP Peer Reviews
von: Barkhordar, Ehsan, et al.
Veröffentlicht: (2026)