DR.GAP: Mitigating Bias in Large Language Models using Gender-Aware Prompting with Demonstration and Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiu, Hongye, Xu, Yue, Qiu, Meikang, Wang, Wenjie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Auto-Search and Refinement: An Automated Framework for Gender Bias Mitigation in Large Language Models
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
From Individuals to Interactions: Benchmarking Gender Bias in Multimodal Large Language Models from the Lens of Social Relationship
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
Locating and Mitigating Gender Bias in Large Language Models
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
$\textit{LinkPrompt}$: Natural and Universal Adversarial Attacks on Prompt-based Language Models
von: Xu, Yue, et al.
Veröffentlicht: (2024)
von: Xu, Yue, et al.
Veröffentlicht: (2024)
GenderAlign: An Alignment Dataset for Mitigating Gender Bias in Large Language Models
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
LFTF: Locating First and Then Fine-Tuning for Mitigating Gender Bias in Large Language Models
von: Qin, Zhanyue, et al.
Veröffentlicht: (2025)
von: Qin, Zhanyue, et al.
Veröffentlicht: (2025)
Context-Aware Counterfactual Data Augmentation for Gender Bias Mitigation in Language Models
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
Mitigating Gender Bias in Code Large Language Models via Model Editing
von: Qin, Zhanyue, et al.
Veröffentlicht: (2024)
von: Qin, Zhanyue, et al.
Veröffentlicht: (2024)
Promoting Equality in Large Language Models: Identifying and Mitigating the Implicit Bias based on Bayesian Theory
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
Continual Learning Using Only Large Language Model Prompting
von: Qiu, Jiabao, et al.
Veröffentlicht: (2024)
von: Qiu, Jiabao, et al.
Veröffentlicht: (2024)
Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs
von: Pearman, Edie, et al.
Veröffentlicht: (2026)
von: Pearman, Edie, et al.
Veröffentlicht: (2026)
Stop Before You Fail: Operational Capability Boundaries for Mitigating Unproductive Reasoning in Large Reasoning Models
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
A Survey on Multilingual Large Language Models: Corpora, Alignment, and Bias
von: Xu, Yuemei, et al.
Veröffentlicht: (2024)
von: Xu, Yuemei, et al.
Veröffentlicht: (2024)
When Audio and Text Disagree: Revealing Text Bias in Large Audio-Language Models
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
Large Language Models are Clinical Reasoners: Reasoning-Aware Diagnosis Framework with Prompt-Generated Rationales
von: Kwon, Taeyoon, et al.
Veröffentlicht: (2023)
von: Kwon, Taeyoon, et al.
Veröffentlicht: (2023)
Struct-X: Enhancing Large Language Models Reasoning with Structured Data
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
Towards Explainable Temporal Reasoning in Large Language Models: A Structure-Aware Generative Framework
von: Jiang, Zihao, et al.
Veröffentlicht: (2025)
von: Jiang, Zihao, et al.
Veröffentlicht: (2025)
Mitigating Selection Bias in Large Language Models via Permutation-Aware GRPO
von: Zheng, Jinquan, et al.
Veröffentlicht: (2026)
von: Zheng, Jinquan, et al.
Veröffentlicht: (2026)
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
von: Oi, Masanari, et al.
Veröffentlicht: (2024)
von: Oi, Masanari, et al.
Veröffentlicht: (2024)
Multi-Persona Thinking for Bias Mitigation in Large Language Models
von: Chen, Yuxing, et al.
Veröffentlicht: (2026)
von: Chen, Yuxing, et al.
Veröffentlicht: (2026)
GenderCARE: A Comprehensive Framework for Assessing and Reducing Gender Bias in Large Language Models
von: Tang, Kunsheng, et al.
Veröffentlicht: (2024)
von: Tang, Kunsheng, et al.
Veröffentlicht: (2024)
Prompt Programming for Cultural Bias and Alignment of Large Language Models
von: Eren, Maksim, et al.
Veröffentlicht: (2026)
von: Eren, Maksim, et al.
Veröffentlicht: (2026)
Gender Bias in Machine Translation and The Era of Large Language Models
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
Investigating Thinking Behaviours of Reasoning-Based Language Models for Social Bias Mitigation
von: Luo, Guoqing, et al.
Veröffentlicht: (2025)
von: Luo, Guoqing, et al.
Veröffentlicht: (2025)
Promptomatix: An Automatic Prompt Optimization Framework for Large Language Models
von: Murthy, Rithesh, et al.
Veröffentlicht: (2025)
von: Murthy, Rithesh, et al.
Veröffentlicht: (2025)
TInR: Exploring Tool-Internalized Reasoning in Large Language Models
von: Xu, Qiancheng, et al.
Veröffentlicht: (2026)
von: Xu, Qiancheng, et al.
Veröffentlicht: (2026)
Large Language Model Bias Mitigation from the Perspective of Knowledge Editing
von: Chen, Ruizhe, et al.
Veröffentlicht: (2024)
von: Chen, Ruizhe, et al.
Veröffentlicht: (2024)
Mitigating Hallucinations in Large Language Models via Causal Reasoning
von: Li, Yuangang, et al.
Veröffentlicht: (2025)
von: Li, Yuangang, et al.
Veröffentlicht: (2025)
Locale-Conditioned Few-Shot Prompting Mitigates Demonstration Regurgitation in On-Device PII Substitution with Small Language Models
von: Sadani, Anuj, et al.
Veröffentlicht: (2026)
von: Sadani, Anuj, et al.
Veröffentlicht: (2026)
Walking in Others' Shoes: How Perspective-Taking Guides Large Language Models in Reducing Toxicity and Bias
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
Inclusivity in Large Language Models: Personality Traits and Gender Bias in Scientific Abstracts
von: Pervez, Naseela, et al.
Veröffentlicht: (2024)
von: Pervez, Naseela, et al.
Veröffentlicht: (2024)
Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach
von: Cao, Hongyu, et al.
Veröffentlicht: (2026)
von: Cao, Hongyu, et al.
Veröffentlicht: (2026)
Mitigating Extrinsic Gender Bias for Bangla Classification Tasks
von: Joy, Sajib Kumar Saha, et al.
Veröffentlicht: (2024)
von: Joy, Sajib Kumar Saha, et al.
Veröffentlicht: (2024)
Mitigation of Gender and Ethnicity Bias in AI-Generated Stories through Model Explanations
von: Dimgba, Martha O., et al.
Veröffentlicht: (2025)
von: Dimgba, Martha O., et al.
Veröffentlicht: (2025)
Robust Prompt Optimization for Large Language Models Against Distribution Shifts
von: Li, Moxin, et al.
Veröffentlicht: (2023)
von: Li, Moxin, et al.
Veröffentlicht: (2023)
Bayesian Teaching Enables Probabilistic Reasoning in Large Language Models
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Harnessing the Reasoning Economy: A Survey of Efficient Reasoning for Large Language Models
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Take Care of Your Prompt Bias! Investigating and Mitigating Prompt Bias in Factual Knowledge Extraction
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
Are Large Language Models Really Bias-Free? Jailbreak Prompts for Assessing Adversarial Robustness to Bias Elicitation
von: Cantini, Riccardo, et al.
Veröffentlicht: (2024)
von: Cantini, Riccardo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Auto-Search and Refinement: An Automated Framework for Gender Bias Mitigation in Large Language Models
von: Xu, Yue, et al.
Veröffentlicht: (2025) -
From Individuals to Interactions: Benchmarking Gender Bias in Multimodal Large Language Models from the Lens of Social Relationship
von: Xu, Yue, et al.
Veröffentlicht: (2025) -
Locating and Mitigating Gender Bias in Large Language Models
von: Cai, Yuchen, et al.
Veröffentlicht: (2024) -
$\textit{LinkPrompt}$: Natural and Universal Adversarial Attacks on Prompt-based Language Models
von: Xu, Yue, et al.
Veröffentlicht: (2024) -
GenderAlign: An Alignment Dataset for Mitigating Gender Bias in Large Language Models
von: Zhang, Tao, et al.
Veröffentlicht: (2024)