What's in a Name? Auditing Large Language Models for Race and Gender Bias
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Salinas, Alejandro, Haim, Amit, Nyarko, Julian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies
von: Ma, Sibo, et al.
Veröffentlicht: (2025)
von: Ma, Sibo, et al.
Veröffentlicht: (2025)
Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval
von: Wilson, Kyra, et al.
Veröffentlicht: (2024)
von: Wilson, Kyra, et al.
Veröffentlicht: (2024)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Bias and Fairness in Large Language Models: A Survey
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2023)
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2023)
Who Gets Which Message? Auditing Demographic Bias in LLM-Generated Targeted Text
von: Islam, Tunazzina
Veröffentlicht: (2026)
von: Islam, Tunazzina
Veröffentlicht: (2026)
Contextual StereoSet: Stress-Testing Bias Alignment Robustness in Large Language Models
von: Basu, Abhinaba, et al.
Veröffentlicht: (2026)
von: Basu, Abhinaba, et al.
Veröffentlicht: (2026)
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
von: Urchs, Stefanie, et al.
Veröffentlicht: (2023)
von: Urchs, Stefanie, et al.
Veröffentlicht: (2023)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
Empowering Many, Biasing a Few: Generalist Credit Scoring through Large Language Models
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
Gender Bias in Machine Translation and The Era of Large Language Models
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
Hypothesis Generation with Large Language Models
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
Large Language Models are Geographically Biased
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
Correlated Errors in Large Language Models
von: Kim, Elliot, et al.
Veröffentlicht: (2025)
von: Kim, Elliot, et al.
Veröffentlicht: (2025)
AccessEval: Benchmarking Disability Bias in Large Language Models
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
Say My Name: a Model's Bias Discovery Framework
von: Ciranni, Massimiliano, et al.
Veröffentlicht: (2024)
von: Ciranni, Massimiliano, et al.
Veröffentlicht: (2024)
Psychological Counseling Ability of Large Language Models
von: Peng, Fangyu, et al.
Veröffentlicht: (2025)
von: Peng, Fangyu, et al.
Veröffentlicht: (2025)
Assessing Large Language Models on Climate Information
von: Bulian, Jannis, et al.
Veröffentlicht: (2023)
von: Bulian, Jannis, et al.
Veröffentlicht: (2023)
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Masculine Defaults via Gendered Discourse in Podcasts and Large Language Models
von: Teleki, Maria, et al.
Veröffentlicht: (2025)
von: Teleki, Maria, et al.
Veröffentlicht: (2025)
Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
von: Kıcıman, Emre, et al.
Veröffentlicht: (2023)
von: Kıcıman, Emre, et al.
Veröffentlicht: (2023)
A Taxonomy of Stereotype Content in Large Language Models
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
Transforming Agency. On the mode of existence of Large Language Models
von: Barandiaran, Xabier E., et al.
Veröffentlicht: (2024)
von: Barandiaran, Xabier E., et al.
Veröffentlicht: (2024)
Benchmarking Gender and Political Bias in Large Language Models
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models
von: Faiz, Ahmad, et al.
Veröffentlicht: (2023)
von: Faiz, Ahmad, et al.
Veröffentlicht: (2023)
ResumeAtlas: Revisiting Resume Classification with Large-Scale Datasets and Large Language Models
von: Heakl, Ahmed, et al.
Veröffentlicht: (2024)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2024)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
What is in a name? Mitigating Name Bias in Text Embeddings via Anonymization
von: Manchanda, Sahil, et al.
Veröffentlicht: (2025)
von: Manchanda, Sahil, et al.
Veröffentlicht: (2025)
Foundational Challenges in Assuring Alignment and Safety of Large Language Models
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
Exploring Accuracy-Fairness Trade-off in Large Language Models
von: Zhang, Qingquan, et al.
Veröffentlicht: (2024)
von: Zhang, Qingquan, et al.
Veröffentlicht: (2024)
Are Large Language Models Chameleons? An Attempt to Simulate Social Surveys
von: Geng, Mingmeng, et al.
Veröffentlicht: (2024)
von: Geng, Mingmeng, et al.
Veröffentlicht: (2024)
Towards Large Language Models that Benefit for All: Benchmarking Group Fairness in Reward Models
von: Song, Kefan, et al.
Veröffentlicht: (2025)
von: Song, Kefan, et al.
Veröffentlicht: (2025)
REQUAL-LM: Reliability and Equity through Aggregation in Large Language Models
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
A Moral Imperative: The Need for Continual Superalignment of Large Language Models
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2024)
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2024)
Emissions and Performance Trade-off Between Small and Large Language Models
von: Garg, Anandita, et al.
Veröffentlicht: (2025)
von: Garg, Anandita, et al.
Veröffentlicht: (2025)
AuditWen:An Open-Source Large Language Model for Audit
von: Huang, Jiajia, et al.
Veröffentlicht: (2024)
von: Huang, Jiajia, et al.
Veröffentlicht: (2024)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
von: Zhou, Han, et al.
Veröffentlicht: (2024)
von: Zhou, Han, et al.
Veröffentlicht: (2024)
Large Language Models Assume People are More Rational than We Really are
von: Liu, Ryan, et al.
Veröffentlicht: (2024)
von: Liu, Ryan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies
von: Ma, Sibo, et al.
Veröffentlicht: (2025) -
Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval
von: Wilson, Kyra, et al.
Veröffentlicht: (2024) -
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
von: Xu, Xin, et al.
Veröffentlicht: (2025) -
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025) -
Bias and Fairness in Large Language Models: A Survey
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2023)