Assessing GPT's Bias Towards Stigmatized Social Groups: An Intersectional Case Study on Nationality Prejudice and Psychophobia
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kashif, Afifah, Patel, Heer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Quite Good, but Not Enough: Nationality Bias in Large Language Models -- A Case Study of ChatGPT
von: Zhu, Shucheng, et al.
Veröffentlicht: (2024)
von: Zhu, Shucheng, et al.
Veröffentlicht: (2024)
From Preferences to Prejudice: The Role of Alignment Tuning in Shaping Social Bias in Video Diffusion Models
von: Cai, Zefan, et al.
Veröffentlicht: (2025)
von: Cai, Zefan, et al.
Veröffentlicht: (2025)
Identifying Features Associated with Bias Against 93 Stigmatized Groups in Language Models and Guardrail Model Safety Mitigation
von: Gueorguieva, Anna-Maria, et al.
Veröffentlicht: (2025)
von: Gueorguieva, Anna-Maria, et al.
Veröffentlicht: (2025)
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
Prompt and Prejudice
von: Berlincioni, Lorenzo, et al.
Veröffentlicht: (2024)
von: Berlincioni, Lorenzo, et al.
Veröffentlicht: (2024)
Assessing Agentic Large Language Models in Multilingual National Bias
von: Liu, Qianying, et al.
Veröffentlicht: (2025)
von: Liu, Qianying, et al.
Veröffentlicht: (2025)
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT
von: Reddy, Harishwar, et al.
Veröffentlicht: (2025)
von: Reddy, Harishwar, et al.
Veröffentlicht: (2025)
Personalisation or Prejudice? Addressing Geographic Bias in Hate Speech Detection using Debias Tuning in Large Language Models
von: Piot, Paloma, et al.
Veröffentlicht: (2025)
von: Piot, Paloma, et al.
Veröffentlicht: (2025)
GPT and Prejudice: A Sparse Approach to Understanding Learned Representations in Large Language Models
von: Mahran, Mariam, et al.
Veröffentlicht: (2025)
von: Mahran, Mariam, et al.
Veröffentlicht: (2025)
Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups
von: Liu, Geng, et al.
Veröffentlicht: (2025)
von: Liu, Geng, et al.
Veröffentlicht: (2025)
Understanding Stigmatizing Language Lexicons: A Comparative Analysis in Clinical Contexts
von: Zhou, Yiliang, et al.
Veröffentlicht: (2025)
von: Zhou, Yiliang, et al.
Veröffentlicht: (2025)
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
von: Wang, Qian, et al.
Veröffentlicht: (2025)
von: Wang, Qian, et al.
Veröffentlicht: (2025)
Reducing Large Language Model Bias with Emphasis on 'Restricted Industries': Automated Dataset Augmentation and Prejudice Quantification
von: Mondal, Devam, et al.
Veröffentlicht: (2024)
von: Mondal, Devam, et al.
Veröffentlicht: (2024)
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4
von: Yadav, Sachin, et al.
Veröffentlicht: (2024)
von: Yadav, Sachin, et al.
Veröffentlicht: (2024)
Silver-Tongued and Sundry: Exploring Intersectional Pronouns with ChatGPT
von: Fujii, Takao, et al.
Veröffentlicht: (2024)
von: Fujii, Takao, et al.
Veröffentlicht: (2024)
Social Bias in Large Language Models For Bangla: An Empirical Study on Gender and Religious Bias
von: Sadhu, Jayanta, et al.
Veröffentlicht: (2024)
von: Sadhu, Jayanta, et al.
Veröffentlicht: (2024)
Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making
von: Huang, Jen-tse, et al.
Veröffentlicht: (2026)
von: Huang, Jen-tse, et al.
Veröffentlicht: (2026)
Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
von: Ye, Jiayi, et al.
Veröffentlicht: (2024)
von: Ye, Jiayi, et al.
Veröffentlicht: (2024)
ChatGPT v.s. Media Bias: A Comparative Study of GPT-3.5 and Fine-tuned Language Models
von: Wen, Zehao, et al.
Veröffentlicht: (2024)
von: Wen, Zehao, et al.
Veröffentlicht: (2024)
Bias in Opinion Summarisation from Pre-training to Adaptation: A Case Study in Political Bias
von: Huang, Nannan, et al.
Veröffentlicht: (2024)
von: Huang, Nannan, et al.
Veröffentlicht: (2024)
Probability of Differentiation Reveals Brittleness of Homogeneity Bias in GPT-4
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
Assessing and Refining ChatGPT's Performance in Identifying Targeting and Inappropriate Language: A Comparative Study
von: Baran, Barbarestani, et al.
Veröffentlicht: (2025)
von: Baran, Barbarestani, et al.
Veröffentlicht: (2025)
The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
von: Sauter, Adrian, et al.
Veröffentlicht: (2025)
von: Sauter, Adrian, et al.
Veröffentlicht: (2025)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
von: Chen, Yuen, et al.
Veröffentlicht: (2022)
von: Chen, Yuen, et al.
Veröffentlicht: (2022)
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
ChatGPT as a Translation Engine: A Case Study on Japanese-English
von: Sutanto, Vincent Michael, et al.
Veröffentlicht: (2025)
von: Sutanto, Vincent Michael, et al.
Veröffentlicht: (2025)
Quantitative Assessment of Intersectional Empathetic Bias and Understanding
von: Formanek, Vojtech, et al.
Veröffentlicht: (2024)
von: Formanek, Vojtech, et al.
Veröffentlicht: (2024)
Gender Bias Detection in Court Decisions: A Brazilian Case Study
von: Benatti, Raysa, et al.
Veröffentlicht: (2024)
von: Benatti, Raysa, et al.
Veröffentlicht: (2024)
Obscured but Not Erased: Evaluating Nationality Bias in LLMs via Name-Based Bias Benchmarks
von: Pelosio, Giulio, et al.
Veröffentlicht: (2025)
von: Pelosio, Giulio, et al.
Veröffentlicht: (2025)
Intersectional Bias in Japanese Large Language Models from a Contextualized Perspective
von: Yanaka, Hitomi, et al.
Veröffentlicht: (2025)
von: Yanaka, Hitomi, et al.
Veröffentlicht: (2025)
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2026)
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2026)
ABLEIST: Intersectional Disability Bias in LLM-Generated Hiring Scenarios
von: Phutane, Mahika, et al.
Veröffentlicht: (2025)
von: Phutane, Mahika, et al.
Veröffentlicht: (2025)
Compounding Disadvantage: Auditing Intersectional Bias in LLM-Generated Explanations Across Indian and American STEM Education
von: Gupta, Amogh, et al.
Veröffentlicht: (2026)
von: Gupta, Amogh, et al.
Veröffentlicht: (2026)
Large Language Models Portray Socially Subordinate Groups as More Homogeneous, Consistent with a Bias Observed in Humans
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2026)
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2026)
From Prejudice to Parity: A New Approach to Debiasing Large Language Model Word Embeddings
von: Rakshit, Aishik, et al.
Veröffentlicht: (2024)
von: Rakshit, Aishik, et al.
Veröffentlicht: (2024)
Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian
von: Kalhor, Ghazal, et al.
Veröffentlicht: (2025)
von: Kalhor, Ghazal, et al.
Veröffentlicht: (2025)
Assessing the Reliability and Validity of GPT-4 in Annotating Emotion Appraisal Ratings
von: Ruder, Deniss, et al.
Veröffentlicht: (2025)
von: Ruder, Deniss, et al.
Veröffentlicht: (2025)
Assessing the Level of Toxicity Against Distinct Groups in Bangla Social Media Comments: A Comprehensive Investigation
von: Moin, Mukaffi Bin, et al.
Veröffentlicht: (2024)
von: Moin, Mukaffi Bin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Quite Good, but Not Enough: Nationality Bias in Large Language Models -- A Case Study of ChatGPT
von: Zhu, Shucheng, et al.
Veröffentlicht: (2024) -
From Preferences to Prejudice: The Role of Alignment Tuning in Shaping Social Bias in Video Diffusion Models
von: Cai, Zefan, et al.
Veröffentlicht: (2025) -
Identifying Features Associated with Bias Against 93 Stigmatized Groups in Language Models and Guardrail Model Safety Mitigation
von: Gueorguieva, Anna-Maria, et al.
Veröffentlicht: (2025) -
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
von: Xu, Wenda, et al.
Veröffentlicht: (2024) -
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)