Mitigation of Gender and Ethnicity Bias in AI-Generated Stories through Model Explanations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dimgba, Martha O., Oba, Sharon, Agrawal, Ameeta, Giabbanelli, Philippe J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization
von: Giabbanelli, Philippe J., et al.
Veröffentlicht: (2025)
von: Giabbanelli, Philippe J., et al.
Veröffentlicht: (2025)
Broadening Access to Simulations for End-Users via Large Language Models: Challenges and Opportunities
von: Giabbanelli, Philippe J., et al.
Veröffentlicht: (2024)
von: Giabbanelli, Philippe J., et al.
Veröffentlicht: (2024)
Understanding Position Bias Effects on Fairness in Social Multi-Document Summarization
von: Olabisi, Olubusayo, et al.
Veröffentlicht: (2024)
von: Olabisi, Olubusayo, et al.
Veröffentlicht: (2024)
No-Worse Context-Aware Decoding: Preventing Neutral Regression in Context-Conditioned Generation
von: Tao, Yufei, et al.
Veröffentlicht: (2026)
von: Tao, Yufei, et al.
Veröffentlicht: (2026)
MTQ-Eval: Multilingual Text Quality Evaluation for Language Models
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2025)
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2025)
Cross-Lingual Activation Steering for Multilingual Language Models
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2026)
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2026)
Beyond Data Quantity: Key Factors Driving Performance in Multilingual Language Models
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
CAPO: Confidence Aware Preference Optimization Learning for Multilingual Preferences
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2025)
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2025)
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
von: Masoudian, Shahed, et al.
Veröffentlicht: (2025)
von: Masoudian, Shahed, et al.
Veröffentlicht: (2025)
"Lost-in-the-Later": Framework for Quantifying Contextual Grounding in Large Language Models
von: Tao, Yufei, et al.
Veröffentlicht: (2025)
von: Tao, Yufei, et al.
Veröffentlicht: (2025)
Narrating Causal Graphs with Large Language Models
von: Phatak, Atharva, et al.
Veröffentlicht: (2024)
von: Phatak, Atharva, et al.
Veröffentlicht: (2024)
From Policy to Logic for Efficient and Interpretable Coverage Assessment
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2026)
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2026)
Fair Summarization: Bridging Quality and Diversity in Extractive Summaries
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
Locating and Mitigating Gender Bias in Large Language Models
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
The Impact of Model Scaling on Seen and Unseen Language Performance
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2025)
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2025)
Equilibrium Dynamics and Mitigation of Gender Bias in Synthetically Generated Data
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
GenderAlign: An Alignment Dataset for Mitigating Gender Bias in Large Language Models
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
When Context Leads but Parametric Memory Follows in Large Language Models
von: Tao, Yufei, et al.
Veröffentlicht: (2024)
von: Tao, Yufei, et al.
Veröffentlicht: (2024)
Explain-then-Process: Using Grammar Prompting to Enhance Grammatical Acceptability Judgments
von: Scheinberg, Russell, et al.
Veröffentlicht: (2025)
von: Scheinberg, Russell, et al.
Veröffentlicht: (2025)
Correct-Detect: Balancing Performance and Ambiguity Through the Lens of Coreference Resolution in LLMs
von: Shore, Amber, et al.
Veröffentlicht: (2025)
von: Shore, Amber, et al.
Veröffentlicht: (2025)
Context-Aware Counterfactual Data Augmentation for Gender Bias Mitigation in Language Models
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
A Guide to Large Language Models in Modeling and Simulation: From Core Techniques to Critical Challenges
von: Giabbanelli, Philippe J.
Veröffentlicht: (2026)
von: Giabbanelli, Philippe J.
Veröffentlicht: (2026)
LFTF: Locating First and Then Fine-Tuning for Mitigating Gender Bias in Large Language Models
von: Qin, Zhanyue, et al.
Veröffentlicht: (2025)
von: Qin, Zhanyue, et al.
Veröffentlicht: (2025)
Auto-Search and Refinement: An Automated Framework for Gender Bias Mitigation in Large Language Models
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
FAIRE: Assessing Racial and Gender Bias in AI-Driven Resume Evaluations
von: Wen, Athena, et al.
Veröffentlicht: (2025)
von: Wen, Athena, et al.
Veröffentlicht: (2025)
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
Mitigating Gender Bias in Code Large Language Models via Model Editing
von: Qin, Zhanyue, et al.
Veröffentlicht: (2024)
von: Qin, Zhanyue, et al.
Veröffentlicht: (2024)
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation
von: Prahallad, Lavanya, et al.
Veröffentlicht: (2024)
von: Prahallad, Lavanya, et al.
Veröffentlicht: (2024)
Mitigating Extrinsic Gender Bias for Bangla Classification Tasks
von: Joy, Sajib Kumar Saha, et al.
Veröffentlicht: (2024)
von: Joy, Sajib Kumar Saha, et al.
Veröffentlicht: (2024)
DR.GAP: Mitigating Bias in Large Language Models using Gender-Aware Prompting with Demonstration and Reasoning
von: Qiu, Hongye, et al.
Veröffentlicht: (2025)
von: Qiu, Hongye, et al.
Veröffentlicht: (2025)
Making a Long Story Short in Conversation Modeling
von: Tao, Yufei, et al.
Veröffentlicht: (2024)
von: Tao, Yufei, et al.
Veröffentlicht: (2024)
UniBias: Unveiling and Mitigating LLM Bias through Internal Attention and FFN Manipulation
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2024)
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2024)
A Close Reading Approach to Gender Narrative Biases in AI-Generated Stories
von: Raffini, Daniel, et al.
Veröffentlicht: (2025)
von: Raffini, Daniel, et al.
Veröffentlicht: (2025)
Mitigating Length Bias in RLHF through a Causal Lens
von: Kim, Hyeonji, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonji, et al.
Veröffentlicht: (2025)
MoESD: Mixture of Experts Stable Diffusion to Mitigate Gender Bias
von: Wang, Guorun, et al.
Veröffentlicht: (2024)
von: Wang, Guorun, et al.
Veröffentlicht: (2024)
Unveiling the "Fairness Seesaw": Discovering and Mitigating Gender and Race Bias in Vision-Language Models
von: Lan, Jian, et al.
Veröffentlicht: (2025)
von: Lan, Jian, et al.
Veröffentlicht: (2025)
Detecting and Mitigating Bias in LLMs through Knowledge Graph-Augmented Training
von: Kumar, Rajeev, et al.
Veröffentlicht: (2025)
von: Kumar, Rajeev, et al.
Veröffentlicht: (2025)
StoryAlign: Evaluating and Training Reward Models for Story Generation
von: Xia, Haotian, et al.
Veröffentlicht: (2026)
von: Xia, Haotian, et al.
Veröffentlicht: (2026)
LLM-Guided Synthetic Augmentation (LGSA) for Mitigating Bias in AI Systems
von: Karri, Sai Suhruth Reddy, et al.
Veröffentlicht: (2025)
von: Karri, Sai Suhruth Reddy, et al.
Veröffentlicht: (2025)
Word2World: Generating Stories and Worlds through Large Language Models
von: Nasir, Muhammad U., et al.
Veröffentlicht: (2024)
von: Nasir, Muhammad U., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization
von: Giabbanelli, Philippe J., et al.
Veröffentlicht: (2025) -
Broadening Access to Simulations for End-Users via Large Language Models: Challenges and Opportunities
von: Giabbanelli, Philippe J., et al.
Veröffentlicht: (2024) -
Understanding Position Bias Effects on Fairness in Social Multi-Document Summarization
von: Olabisi, Olubusayo, et al.
Veröffentlicht: (2024) -
No-Worse Context-Aware Decoding: Preventing Neutral Regression in Context-Conditioned Generation
von: Tao, Yufei, et al.
Veröffentlicht: (2026) -
MTQ-Eval: Multilingual Text Quality Evaluation for Language Models
von: Pokharel, Rhitabrat, et al.
Veröffentlicht: (2025)