White Men Lead, Black Women Help? Benchmarking and Mitigating Language Agency Social Biases in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wan, Yixin, Chang, Kai-Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
The Male CEO and the Female Assistant: Evaluation and Mitigation of Gender Biases in Text-To-Image Generation of Dual Subjects
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
The Factuality Tax of Diversity-Intervened Text-to-Image Generation: Benchmark and Fact-Augmented Intervention
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
"Not in My Backyard": LLMs Uncover Online and Offline Social Biases Against Homelessness
von: Karr Jr., Jonathan A., et al.
Veröffentlicht: (2025)
von: Karr Jr., Jonathan A., et al.
Veröffentlicht: (2025)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
von: Christian, Brian, et al.
Veröffentlicht: (2026)
von: Christian, Brian, et al.
Veröffentlicht: (2026)
MEDEQUALQA: Evaluating Biases in LLMs with Counterfactual Reasoning
von: Ghosh, Rajarshi, et al.
Veröffentlicht: (2025)
von: Ghosh, Rajarshi, et al.
Veröffentlicht: (2025)
Large Language Models Develop Novel Social Biases Through Adaptive Exploration
von: Wu, Addison J., et al.
Veröffentlicht: (2025)
von: Wu, Addison J., et al.
Veröffentlicht: (2025)
No Free Lunch in Language Model Bias Mitigation? Targeted Bias Reduction Can Exacerbate Unmitigated LLM Biases
von: Chand, Shireen, et al.
Veröffentlicht: (2025)
von: Chand, Shireen, et al.
Veröffentlicht: (2025)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
von: Ren, Ruiping, et al.
Veröffentlicht: (2024)
von: Ren, Ruiping, et al.
Veröffentlicht: (2024)
Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs
von: de Landa, Joseba Fernandez, et al.
Veröffentlicht: (2026)
von: de Landa, Joseba Fernandez, et al.
Veröffentlicht: (2026)
Generative AI Carries Non-Democratic Biases and Stereotypes: Representation of Women, Black Individuals, Age Groups, and People with Disability in AI-Generated Images across Occupations
von: Sadeghiani, Ayoob
Veröffentlicht: (2024)
von: Sadeghiani, Ayoob
Veröffentlicht: (2024)
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
CompAlign: Improving Compositional Text-to-Image Generation with a Complex Benchmark and Fine-Grained Feedback
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
A Systematic Analysis of Biases in Large Language Models
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
Open Source Language Models Can Provide Feedback: Evaluating LLMs' Ability to Help Students Using GPT-4-As-A-Judge
von: Koutcheme, Charles, et al.
Veröffentlicht: (2024)
von: Koutcheme, Charles, et al.
Veröffentlicht: (2024)
Identifying Implicit Social Biases in Vision-Language Models
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
PRISM: A Methodology for Auditing Biases in Large Language Models
von: Azzopardi, Leif, et al.
Veröffentlicht: (2024)
von: Azzopardi, Leif, et al.
Veröffentlicht: (2024)
From "Help" to Helpful: A Hierarchical Assessment of LLMs in Mental e-Health Applications
von: Steigerwald, Philipp, et al.
Veröffentlicht: (2026)
von: Steigerwald, Philipp, et al.
Veröffentlicht: (2026)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
von: Salecha, Aadesh, et al.
Veröffentlicht: (2024)
von: Salecha, Aadesh, et al.
Veröffentlicht: (2024)
WHBench: Evaluating Frontier LLMs with Expert-in-the-Loop Validation on Women's Health Topics
von: Maurya, Sneha, et al.
Veröffentlicht: (2026)
von: Maurya, Sneha, et al.
Veröffentlicht: (2026)
Large Language Models are Geographically Biased
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
Happy Young Women, Grumpy Old Men? Emotion-Driven Demographic Biases in Synthetic Face Generation
von: Wei, Mengting, et al.
Veröffentlicht: (2026)
von: Wei, Mengting, et al.
Veröffentlicht: (2026)
Ask LLMs Directly, "What shapes your bias?": Measuring Social Bias in Large Language Models
von: Shin, Jisu, et al.
Veröffentlicht: (2024)
von: Shin, Jisu, et al.
Veröffentlicht: (2024)
HugAgent: Benchmarking LLMs for Simulation of Individualized Human Reasoning
von: Li, Chance Jiajie, et al.
Veröffentlicht: (2025)
von: Li, Chance Jiajie, et al.
Veröffentlicht: (2025)
Misaligned by Reward: Socially Undesirable Preferences in LLMs
von: Ghazaryan, Gayane, et al.
Veröffentlicht: (2026)
von: Ghazaryan, Gayane, et al.
Veröffentlicht: (2026)
Mitigating Social Biases in Language Models through Unlearning
von: Dige, Omkar, et al.
Veröffentlicht: (2024)
von: Dige, Omkar, et al.
Veröffentlicht: (2024)
Social Bias in Popular Question-Answering Benchmarks
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
PLawBench: A Rubric-Based Benchmark for Evaluating LLMs in Real-World Legal Practice
von: Shi, Yuzhen, et al.
Veröffentlicht: (2026)
von: Shi, Yuzhen, et al.
Veröffentlicht: (2026)
From Individuals to Interactions: Benchmarking Gender Bias in Multimodal Large Language Models from the Lens of Social Relationship
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
Transforming Agency. On the mode of existence of Large Language Models
von: Barandiaran, Xabier E., et al.
Veröffentlicht: (2024)
von: Barandiaran, Xabier E., et al.
Veröffentlicht: (2024)
GECOBench: A Gender-Controlled Text Dataset and Benchmark for Quantifying Biases in Explanations
von: Wilming, Rick, et al.
Veröffentlicht: (2024)
von: Wilming, Rick, et al.
Veröffentlicht: (2024)
LocalBench: Benchmarking LLMs on County-Level Local Knowledge and Reasoning
von: Gao, Zihan, et al.
Veröffentlicht: (2025)
von: Gao, Zihan, et al.
Veröffentlicht: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
The Balancing Act: Unmasking and Alleviating ASR Biases in Portuguese
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2024)
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2024)
Writing in Symbiosis: Mapping Human Creative Agency in the AI Era
von: Doshi, Vivan, et al.
Veröffentlicht: (2025)
von: Doshi, Vivan, et al.
Veröffentlicht: (2025)
XCR-Bench: A Multi-Task Benchmark for Evaluating Cultural Reasoning in LLMs
von: Kabir, Mohsinul, et al.
Veröffentlicht: (2026)
von: Kabir, Mohsinul, et al.
Veröffentlicht: (2026)
Defining bias in AI-systems: Biased models are fair models
von: Lindloff, Chiara, et al.
Veröffentlicht: (2025)
von: Lindloff, Chiara, et al.
Veröffentlicht: (2025)
Laissez-Faire Harms: Algorithmic Biases in Generative Language Models
von: Shieh, Evan, et al.
Veröffentlicht: (2024)
von: Shieh, Evan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation
von: Wan, Yixin, et al.
Veröffentlicht: (2025) -
The Male CEO and the Female Assistant: Evaluation and Mitigation of Gender Biases in Text-To-Image Generation of Dual Subjects
von: Wan, Yixin, et al.
Veröffentlicht: (2024) -
The Factuality Tax of Diversity-Intervened Text-to-Image Generation: Benchmark and Fact-Augmented Intervention
von: Wan, Yixin, et al.
Veröffentlicht: (2024) -
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
von: Ding, Junchen, et al.
Veröffentlicht: (2025) -
"Not in My Backyard": LLMs Uncover Online and Offline Social Biases Against Homelessness
von: Karr Jr., Jonathan A., et al.
Veröffentlicht: (2025)