Who is better at math, Jenny or Jingzhen? Uncovering Stereotypes in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Siddique, Zara, Turner, Liam D., Espinosa-Anke, Luis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Shifting Perspectives: Steering Vectors for Robust Bias Mitigation in LLMs
von: Siddique, Zara, et al.
Veröffentlicht: (2025)
von: Siddique, Zara, et al.
Veröffentlicht: (2025)
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
von: Nakanishi, Akito, et al.
Veröffentlicht: (2025)
von: Nakanishi, Akito, et al.
Veröffentlicht: (2025)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
DECASTE: Unveiling Caste Stereotypes in Large Language Models through Multi-Dimensional Bias Analysis
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2025)
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2025)
Addressing Stereotypes in Large Language Models: A Critical Examination and Mitigation
von: Kazi, Fatima
Veröffentlicht: (2025)
von: Kazi, Fatima
Veröffentlicht: (2025)
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
von: Liu, Yiran, et al.
Veröffentlicht: (2024)
von: Liu, Yiran, et al.
Veröffentlicht: (2024)
A Taxonomy of Stereotype Content in Large Language Models
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
A Survey on Stereotype Detection in Natural Language Processing
von: Cignarella, Alessandra Teresa, et al.
Veröffentlicht: (2025)
von: Cignarella, Alessandra Teresa, et al.
Veröffentlicht: (2025)
Large Language Models as Students Who Think Aloud: Overly Coherent, Verbose, and Confident
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
von: Nejadgholi, Isar, et al.
Veröffentlicht: (2024)
von: Nejadgholi, Isar, et al.
Veröffentlicht: (2024)
An Empirical Investigation of Gender Stereotype Representation in Large Language Models: The Italian Case
von: Giachino, Gioele, et al.
Veröffentlicht: (2025)
von: Giachino, Gioele, et al.
Veröffentlicht: (2025)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2024)
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2024)
Local Contrastive Editing of Gender Stereotypes
von: Lutz, Marlene, et al.
Veröffentlicht: (2024)
von: Lutz, Marlene, et al.
Veröffentlicht: (2024)
From General Reasoning to Domain Expertise: Uncovering the Limits of Generalization in Large Language Models
von: Alsagheer, Dana, et al.
Veröffentlicht: (2025)
von: Alsagheer, Dana, et al.
Veröffentlicht: (2025)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
von: Robles, Melissa, et al.
Veröffentlicht: (2025)
von: Robles, Melissa, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Automatic Extraction of Metaphoric Analogies from Literary Texts: Task Formulation, Dataset Construction, and Evaluation
von: Boisson, Joanne, et al.
Veröffentlicht: (2024)
von: Boisson, Joanne, et al.
Veröffentlicht: (2024)
Who Shares Fake News? Uncovering Insights from Social Media Users' Post Histories
von: Schoenmueller, Verena, et al.
Veröffentlicht: (2022)
von: Schoenmueller, Verena, et al.
Veröffentlicht: (2022)
Surfacing Subtle Stereotypes: A Multilingual, Debate-Oriented Evaluation of Modern LLMs
von: Saeed, Muhammed, et al.
Veröffentlicht: (2025)
von: Saeed, Muhammed, et al.
Veröffentlicht: (2025)
Probing the Subtle Ideological Manipulation of Large Language Models
von: Paschalides, Demetris, et al.
Veröffentlicht: (2025)
von: Paschalides, Demetris, et al.
Veröffentlicht: (2025)
On The Role of Reasoning in the Identification of Subtle Stereotypes in Natural Language
von: Tian, Jacob-Junqi, et al.
Veröffentlicht: (2023)
von: Tian, Jacob-Junqi, et al.
Veröffentlicht: (2023)
Motivation in Large Language Models
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
Dialz: A Python Toolkit for Steering Vectors
von: Siddique, Zara, et al.
Veröffentlicht: (2025)
von: Siddique, Zara, et al.
Veröffentlicht: (2025)
Language of Thought Shapes Output Diversity in Large Language Models
von: Xu, Shaoyang, et al.
Veröffentlicht: (2026)
von: Xu, Shaoyang, et al.
Veröffentlicht: (2026)
Investigating Cultural Alignment of Large Language Models
von: AlKhamissi, Badr, et al.
Veröffentlicht: (2024)
von: AlKhamissi, Badr, et al.
Veröffentlicht: (2024)
Multilingual Large Language Models and Curse of Multilinguality
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
On Classification with Large Language Models in Cultural Analytics
von: Bamman, David, et al.
Veröffentlicht: (2024)
von: Bamman, David, et al.
Veröffentlicht: (2024)
Will Large Language Models Transform Clinical Prediction?
von: Yildiz, Yusuf, et al.
Veröffentlicht: (2025)
von: Yildiz, Yusuf, et al.
Veröffentlicht: (2025)
Large Language Models in the Abuse Detection Pipeline
von: Kath, Suraj, et al.
Veröffentlicht: (2026)
von: Kath, Suraj, et al.
Veröffentlicht: (2026)
Climate Change from Large Language Models
von: Zhu, Hongyin, et al.
Veröffentlicht: (2023)
von: Zhu, Hongyin, et al.
Veröffentlicht: (2023)
Urban Computing in the Era of Large Language Models
von: Li, Zhonghang, et al.
Veröffentlicht: (2025)
von: Li, Zhonghang, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Actionable Course Evaluation Student Feedback to Lecturers
von: Zhang, Mike, et al.
Veröffentlicht: (2024)
von: Zhang, Mike, et al.
Veröffentlicht: (2024)
Development of Application-Specific Large Language Models to Facilitate Research Ethics Review
von: Mann, Sebastian Porsdam, et al.
Veröffentlicht: (2025)
von: Mann, Sebastian Porsdam, et al.
Veröffentlicht: (2025)
Reinforcing Stereotypes of Anger: Emotion AI on African American Vernacular English
von: Dorn, Rebecca, et al.
Veröffentlicht: (2025)
von: Dorn, Rebecca, et al.
Veröffentlicht: (2025)
Uncovering Regulatory Affairs Complexity in Medical Products: A Qualitative Assessment Utilizing Open Coding and Natural Language Processing (NLP)
von: Han, Yu, et al.
Veröffentlicht: (2023)
von: Han, Yu, et al.
Veröffentlicht: (2023)
The World of Generative AI: Deepfakes and Large Language Models
von: Mitra, Alakananda, et al.
Veröffentlicht: (2024)
von: Mitra, Alakananda, et al.
Veröffentlicht: (2024)
Caveat Lector: Large Language Models in Legal Practice
von: Mik, Eliza
Veröffentlicht: (2024)
von: Mik, Eliza
Veröffentlicht: (2024)
Extrinsic Evaluation of Cultural Competence in Large Language Models
von: Bhatt, Shaily, et al.
Veröffentlicht: (2024)
von: Bhatt, Shaily, et al.
Veröffentlicht: (2024)
Large Language Models as Search Engines: Societal Challenges
von: Sadeddine, Zacchary, et al.
Veröffentlicht: (2025)
von: Sadeddine, Zacchary, et al.
Veröffentlicht: (2025)
Evaluating Proactive Risk Awareness of Large Language Models
von: Luo, Xuan, et al.
Veröffentlicht: (2026)
von: Luo, Xuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Shifting Perspectives: Steering Vectors for Robust Bias Mitigation in LLMs
von: Siddique, Zara, et al.
Veröffentlicht: (2025) -
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
von: Nakanishi, Akito, et al.
Veröffentlicht: (2025) -
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024) -
DECASTE: Unveiling Caste Stereotypes in Large Language Models through Multi-Dimensional Bias Analysis
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2025) -
Addressing Stereotypes in Large Language Models: A Critical Examination and Mitigation
von: Kazi, Fatima
Veröffentlicht: (2025)