Disclosure and Mitigation of Gender Bias in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Dong, Xiangjue, Wang, Yibo, Yu, Philip S., Caverlee, James |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
$DA^3$: A Distribution-Aware Adversarial Attack against Language Models
di: Wang, Yibo, et al.
Pubblicazione: (2023)
di: Wang, Yibo, et al.
Pubblicazione: (2023)
A Survey on LLM Inference-Time Self-Improvement
di: Dong, Xiangjue, et al.
Pubblicazione: (2024)
di: Dong, Xiangjue, et al.
Pubblicazione: (2024)
Masculine Defaults via Gendered Discourse in Podcasts and Large Language Models
di: Teleki, Maria, et al.
Pubblicazione: (2025)
di: Teleki, Maria, et al.
Pubblicazione: (2025)
Language Models as Semantic Augmenters for Sequential Recommenders
di: Valizadeh, Mahsa, et al.
Pubblicazione: (2025)
di: Valizadeh, Mahsa, et al.
Pubblicazione: (2025)
CHOIR: Collaborative Harmonization fOr Inference Robustness
di: Dong, Xiangjue, et al.
Pubblicazione: (2025)
di: Dong, Xiangjue, et al.
Pubblicazione: (2025)
Understanding and Mitigating Gender Bias in LLMs via Interpretable Neuron Editing
di: Yu, Zeping, et al.
Pubblicazione: (2025)
di: Yu, Zeping, et al.
Pubblicazione: (2025)
The power of Prompts: Evaluating and Mitigating Gender Bias in MT with LLMs
di: Sant, Aleix, et al.
Pubblicazione: (2024)
di: Sant, Aleix, et al.
Pubblicazione: (2024)
The Neglected Tails in Vision-Language Models
di: Parashar, Shubham, et al.
Pubblicazione: (2024)
di: Parashar, Shubham, et al.
Pubblicazione: (2024)
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
di: Wei, Kangda, et al.
Pubblicazione: (2025)
di: Wei, Kangda, et al.
Pubblicazione: (2025)
DisastQA: A Comprehensive Benchmark for Evaluating Question Answering in Disaster Management
di: Chen, Zhitong, et al.
Pubblicazione: (2026)
di: Chen, Zhitong, et al.
Pubblicazione: (2026)
ReasoningRec: Bridging Personalized Recommendations and Human-Interpretable Explanations through LLM Reasoning
di: Bismay, Millennium, et al.
Pubblicazione: (2024)
di: Bismay, Millennium, et al.
Pubblicazione: (2024)
GenderAlign: An Alignment Dataset for Mitigating Gender Bias in Large Language Models
di: Zhang, Tao, et al.
Pubblicazione: (2024)
di: Zhang, Tao, et al.
Pubblicazione: (2024)
Overview of the NLPCC 2025 Shared Task: Gender Bias Mitigation Challenge
di: Li, Yizhi, et al.
Pubblicazione: (2025)
di: Li, Yizhi, et al.
Pubblicazione: (2025)
Mitigating Gender Bias in Contextual Word Embeddings
di: Yarrabelly, Navya, et al.
Pubblicazione: (2024)
di: Yarrabelly, Navya, et al.
Pubblicazione: (2024)
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
Mitigating Extrinsic Gender Bias for Bangla Classification Tasks
di: Joy, Sajib Kumar Saha, et al.
Pubblicazione: (2024)
di: Joy, Sajib Kumar Saha, et al.
Pubblicazione: (2024)
Detection, Classification, and Mitigation of Gender Bias in Large Language Models
di: Cheng, Xiaoqing, et al.
Pubblicazione: (2025)
di: Cheng, Xiaoqing, et al.
Pubblicazione: (2025)
Exploring the Linear Subspace Hypothesis in Gender Bias Mitigation
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
Evaluating Gender Bias of LLMs in Making Morality Judgements
di: Bajaj, Divij, et al.
Pubblicazione: (2024)
di: Bajaj, Divij, et al.
Pubblicazione: (2024)
Locating and Mitigating Gender Bias in Large Language Models
di: Cai, Yuchen, et al.
Pubblicazione: (2024)
di: Cai, Yuchen, et al.
Pubblicazione: (2024)
Projective Methods for Mitigating Gender Bias in Pre-trained Language Models
di: Dawkins, Hillary, et al.
Pubblicazione: (2024)
di: Dawkins, Hillary, et al.
Pubblicazione: (2024)
Does Context Help Mitigate Gender Bias in Neural Machine Translation?
di: Gete, Harritxu, et al.
Pubblicazione: (2024)
di: Gete, Harritxu, et al.
Pubblicazione: (2024)
CalibraEval: Calibrating Prediction Distribution to Mitigate Selection Bias in LLMs-as-Judges
di: Li, Haitao, et al.
Pubblicazione: (2024)
di: Li, Haitao, et al.
Pubblicazione: (2024)
Can We Trust LLMs? Mitigate Overconfidence Bias in LLMs through Knowledge Transfer
di: Yang, Haoyan, et al.
Pubblicazione: (2024)
di: Yang, Haoyan, et al.
Pubblicazione: (2024)
Beyond Performance: Quantifying and Mitigating Label Bias in LLMs
di: Reif, Yuval, et al.
Pubblicazione: (2024)
di: Reif, Yuval, et al.
Pubblicazione: (2024)
Equilibrium Dynamics and Mitigation of Gender Bias in Synthetically Generated Data
di: Kattamuri, Ashish, et al.
Pubblicazione: (2025)
di: Kattamuri, Ashish, et al.
Pubblicazione: (2025)
MoESD: Mixture of Experts Stable Diffusion to Mitigate Gender Bias
di: Wang, Guorun, et al.
Pubblicazione: (2024)
di: Wang, Guorun, et al.
Pubblicazione: (2024)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
di: Chen, Yuen, et al.
Pubblicazione: (2022)
di: Chen, Yuen, et al.
Pubblicazione: (2022)
The Confidence Trap: Gender Bias and Predictive Certainty in LLMs
di: Sabir, Ahmed, et al.
Pubblicazione: (2026)
di: Sabir, Ahmed, et al.
Pubblicazione: (2026)
Position Bias Mitigates Position Bias:Mitigate Position Bias Through Inter-Position Knowledge Distillation
di: Wang, Yifei, et al.
Pubblicazione: (2025)
di: Wang, Yifei, et al.
Pubblicazione: (2025)
Evaluating Bias in LLMs for Job-Resume Matching: Gender, Race, and Education
di: Iso, Hayate, et al.
Pubblicazione: (2025)
di: Iso, Hayate, et al.
Pubblicazione: (2025)
SPAGBias: Uncovering and Tracing Structured Spatial Gender Bias in Large Language Models
di: Su, Binxian, et al.
Pubblicazione: (2026)
di: Su, Binxian, et al.
Pubblicazione: (2026)
Auto-Search and Refinement: An Automated Framework for Gender Bias Mitigation in Large Language Models
di: Xu, Yue, et al.
Pubblicazione: (2025)
di: Xu, Yue, et al.
Pubblicazione: (2025)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
di: Fernandes, Gustavo Lúcius, et al.
Pubblicazione: (2026)
di: Fernandes, Gustavo Lúcius, et al.
Pubblicazione: (2026)
Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs
di: Pearman, Edie, et al.
Pubblicazione: (2026)
di: Pearman, Edie, et al.
Pubblicazione: (2026)
Steering Towards Fairness: Mitigating Political Bias in LLMs
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups
di: Liu, Geng, et al.
Pubblicazione: (2025)
di: Liu, Geng, et al.
Pubblicazione: (2025)
Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian
di: Kalhor, Ghazal, et al.
Pubblicazione: (2025)
di: Kalhor, Ghazal, et al.
Pubblicazione: (2025)
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
di: Long, Do Xuan, et al.
Pubblicazione: (2024)
di: Long, Do Xuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
$DA^3$: A Distribution-Aware Adversarial Attack against Language Models
di: Wang, Yibo, et al.
Pubblicazione: (2023) -
A Survey on LLM Inference-Time Self-Improvement
di: Dong, Xiangjue, et al.
Pubblicazione: (2024) -
Masculine Defaults via Gendered Discourse in Podcasts and Large Language Models
di: Teleki, Maria, et al.
Pubblicazione: (2025) -
Language Models as Semantic Augmenters for Sequential Recommenders
di: Valizadeh, Mahsa, et al.
Pubblicazione: (2025) -
CHOIR: Collaborative Harmonization fOr Inference Robustness
di: Dong, Xiangjue, et al.
Pubblicazione: (2025)