Assessing GPT's Bias Towards Stigmatized Social Groups: An Intersectional Case Study on Nationality Prejudice and Psychophobia
Fuente:
arXiv
Salvato in:
| Autori principali: | Kashif, Afifah, Patel, Heer |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Quite Good, but Not Enough: Nationality Bias in Large Language Models -- A Case Study of ChatGPT
di: Zhu, Shucheng, et al.
Pubblicazione: (2024)
di: Zhu, Shucheng, et al.
Pubblicazione: (2024)
From Preferences to Prejudice: The Role of Alignment Tuning in Shaping Social Bias in Video Diffusion Models
di: Cai, Zefan, et al.
Pubblicazione: (2025)
di: Cai, Zefan, et al.
Pubblicazione: (2025)
Identifying Features Associated with Bias Against 93 Stigmatized Groups in Language Models and Guardrail Model Safety Mitigation
di: Gueorguieva, Anna-Maria, et al.
Pubblicazione: (2025)
di: Gueorguieva, Anna-Maria, et al.
Pubblicazione: (2025)
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
di: Xu, Wenda, et al.
Pubblicazione: (2024)
di: Xu, Wenda, et al.
Pubblicazione: (2024)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
Prompt and Prejudice
di: Berlincioni, Lorenzo, et al.
Pubblicazione: (2024)
di: Berlincioni, Lorenzo, et al.
Pubblicazione: (2024)
Assessing Agentic Large Language Models in Multilingual National Bias
di: Liu, Qianying, et al.
Pubblicazione: (2025)
di: Liu, Qianying, et al.
Pubblicazione: (2025)
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT
di: Reddy, Harishwar, et al.
Pubblicazione: (2025)
di: Reddy, Harishwar, et al.
Pubblicazione: (2025)
Personalisation or Prejudice? Addressing Geographic Bias in Hate Speech Detection using Debias Tuning in Large Language Models
di: Piot, Paloma, et al.
Pubblicazione: (2025)
di: Piot, Paloma, et al.
Pubblicazione: (2025)
GPT and Prejudice: A Sparse Approach to Understanding Learned Representations in Large Language Models
di: Mahran, Mariam, et al.
Pubblicazione: (2025)
di: Mahran, Mariam, et al.
Pubblicazione: (2025)
Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups
di: Liu, Geng, et al.
Pubblicazione: (2025)
di: Liu, Geng, et al.
Pubblicazione: (2025)
Understanding Stigmatizing Language Lexicons: A Comparative Analysis in Clinical Contexts
di: Zhou, Yiliang, et al.
Pubblicazione: (2025)
di: Zhou, Yiliang, et al.
Pubblicazione: (2025)
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
di: Wang, Qian, et al.
Pubblicazione: (2025)
di: Wang, Qian, et al.
Pubblicazione: (2025)
Reducing Large Language Model Bias with Emphasis on 'Restricted Industries': Automated Dataset Augmentation and Prejudice Quantification
di: Mondal, Devam, et al.
Pubblicazione: (2024)
di: Mondal, Devam, et al.
Pubblicazione: (2024)
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4
di: Yadav, Sachin, et al.
Pubblicazione: (2024)
di: Yadav, Sachin, et al.
Pubblicazione: (2024)
Silver-Tongued and Sundry: Exploring Intersectional Pronouns with ChatGPT
di: Fujii, Takao, et al.
Pubblicazione: (2024)
di: Fujii, Takao, et al.
Pubblicazione: (2024)
Social Bias in Large Language Models For Bangla: An Empirical Study on Gender and Religious Bias
di: Sadhu, Jayanta, et al.
Pubblicazione: (2024)
di: Sadhu, Jayanta, et al.
Pubblicazione: (2024)
Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making
di: Huang, Jen-tse, et al.
Pubblicazione: (2026)
di: Huang, Jen-tse, et al.
Pubblicazione: (2026)
Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
di: Ye, Jiayi, et al.
Pubblicazione: (2024)
di: Ye, Jiayi, et al.
Pubblicazione: (2024)
ChatGPT v.s. Media Bias: A Comparative Study of GPT-3.5 and Fine-tuned Language Models
di: Wen, Zehao, et al.
Pubblicazione: (2024)
di: Wen, Zehao, et al.
Pubblicazione: (2024)
Bias in Opinion Summarisation from Pre-training to Adaptation: A Case Study in Political Bias
di: Huang, Nannan, et al.
Pubblicazione: (2024)
di: Huang, Nannan, et al.
Pubblicazione: (2024)
Probability of Differentiation Reveals Brittleness of Homogeneity Bias in GPT-4
di: Lee, Messi H. J., et al.
Pubblicazione: (2024)
di: Lee, Messi H. J., et al.
Pubblicazione: (2024)
Assessing and Refining ChatGPT's Performance in Identifying Targeting and Inappropriate Language: A Comparative Study
di: Baran, Barbarestani, et al.
Pubblicazione: (2025)
di: Baran, Barbarestani, et al.
Pubblicazione: (2025)
The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
di: Sauter, Adrian, et al.
Pubblicazione: (2025)
di: Sauter, Adrian, et al.
Pubblicazione: (2025)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
di: Chen, Yuen, et al.
Pubblicazione: (2022)
di: Chen, Yuen, et al.
Pubblicazione: (2022)
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination
di: Fleisig, Eve, et al.
Pubblicazione: (2024)
di: Fleisig, Eve, et al.
Pubblicazione: (2024)
ChatGPT as a Translation Engine: A Case Study on Japanese-English
di: Sutanto, Vincent Michael, et al.
Pubblicazione: (2025)
di: Sutanto, Vincent Michael, et al.
Pubblicazione: (2025)
Quantitative Assessment of Intersectional Empathetic Bias and Understanding
di: Formanek, Vojtech, et al.
Pubblicazione: (2024)
di: Formanek, Vojtech, et al.
Pubblicazione: (2024)
Gender Bias Detection in Court Decisions: A Brazilian Case Study
di: Benatti, Raysa, et al.
Pubblicazione: (2024)
di: Benatti, Raysa, et al.
Pubblicazione: (2024)
Obscured but Not Erased: Evaluating Nationality Bias in LLMs via Name-Based Bias Benchmarks
di: Pelosio, Giulio, et al.
Pubblicazione: (2025)
di: Pelosio, Giulio, et al.
Pubblicazione: (2025)
Intersectional Bias in Japanese Large Language Models from a Contextualized Perspective
di: Yanaka, Hitomi, et al.
Pubblicazione: (2025)
di: Yanaka, Hitomi, et al.
Pubblicazione: (2025)
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
ABLEIST: Intersectional Disability Bias in LLM-Generated Hiring Scenarios
di: Phutane, Mahika, et al.
Pubblicazione: (2025)
di: Phutane, Mahika, et al.
Pubblicazione: (2025)
Compounding Disadvantage: Auditing Intersectional Bias in LLM-Generated Explanations Across Indian and American STEM Education
di: Gupta, Amogh, et al.
Pubblicazione: (2026)
di: Gupta, Amogh, et al.
Pubblicazione: (2026)
Large Language Models Portray Socially Subordinate Groups as More Homogeneous, Consistent with a Bias Observed in Humans
di: Lee, Messi H. J., et al.
Pubblicazione: (2024)
di: Lee, Messi H. J., et al.
Pubblicazione: (2024)
From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
From Prejudice to Parity: A New Approach to Debiasing Large Language Model Word Embeddings
di: Rakshit, Aishik, et al.
Pubblicazione: (2024)
di: Rakshit, Aishik, et al.
Pubblicazione: (2024)
Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian
di: Kalhor, Ghazal, et al.
Pubblicazione: (2025)
di: Kalhor, Ghazal, et al.
Pubblicazione: (2025)
Assessing the Reliability and Validity of GPT-4 in Annotating Emotion Appraisal Ratings
di: Ruder, Deniss, et al.
Pubblicazione: (2025)
di: Ruder, Deniss, et al.
Pubblicazione: (2025)
Assessing the Level of Toxicity Against Distinct Groups in Bangla Social Media Comments: A Comprehensive Investigation
di: Moin, Mukaffi Bin, et al.
Pubblicazione: (2024)
di: Moin, Mukaffi Bin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Quite Good, but Not Enough: Nationality Bias in Large Language Models -- A Case Study of ChatGPT
di: Zhu, Shucheng, et al.
Pubblicazione: (2024) -
From Preferences to Prejudice: The Role of Alignment Tuning in Shaping Social Bias in Video Diffusion Models
di: Cai, Zefan, et al.
Pubblicazione: (2025) -
Identifying Features Associated with Bias Against 93 Stigmatized Groups in Language Models and Guardrail Model Safety Mitigation
di: Gueorguieva, Anna-Maria, et al.
Pubblicazione: (2025) -
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
di: Xu, Wenda, et al.
Pubblicazione: (2024) -
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)