Quantifying Stereotypes in Language
Fuente:
arXiv
Salvato in:
| Autore principale: | Liu, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
di: Liu, Yiran, et al.
Pubblicazione: (2024)
di: Liu, Yiran, et al.
Pubblicazione: (2024)
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
di: Nakanishi, Akito, et al.
Pubblicazione: (2025)
di: Nakanishi, Akito, et al.
Pubblicazione: (2025)
Systematic Offensive Stereotyping (SOS) Bias in Language Models
di: Elsafoury, Fatma
Pubblicazione: (2023)
di: Elsafoury, Fatma
Pubblicazione: (2023)
Measuring Stereotype and Deviation Biases in Large Language Models
di: Wang, Daniel, et al.
Pubblicazione: (2025)
di: Wang, Daniel, et al.
Pubblicazione: (2025)
A Survey on Stereotype Detection in Natural Language Processing
di: Cignarella, Alessandra Teresa, et al.
Pubblicazione: (2025)
di: Cignarella, Alessandra Teresa, et al.
Pubblicazione: (2025)
Do Multilingual Large Language Models Mitigate Stereotype Bias?
di: Nie, Shangrui, et al.
Pubblicazione: (2024)
di: Nie, Shangrui, et al.
Pubblicazione: (2024)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
di: Nejadgholi, Isar, et al.
Pubblicazione: (2024)
di: Nejadgholi, Isar, et al.
Pubblicazione: (2024)
Engagement Undermines Safety: How Stereotypes and Toxicity Shape Humor in Language Models
di: Dogra, Atharvan, et al.
Pubblicazione: (2025)
di: Dogra, Atharvan, et al.
Pubblicazione: (2025)
Women Are Beautiful, Men Are Leaders: Gender Stereotypes in Machine Translation and Language Modeling
di: Pikuliak, Matúš, et al.
Pubblicazione: (2023)
di: Pikuliak, Matúš, et al.
Pubblicazione: (2023)
Who is better at math, Jenny or Jingzhen? Uncovering Stereotypes in Large Language Models
di: Siddique, Zara, et al.
Pubblicazione: (2024)
di: Siddique, Zara, et al.
Pubblicazione: (2024)
Evaluation of Large Language Models: STEM education and Gender Stereotypes
di: Due, Smilla, et al.
Pubblicazione: (2024)
di: Due, Smilla, et al.
Pubblicazione: (2024)
Detecting Linguistic Indicators for Stereotype Assessment with Large Language Models
di: Görge, Rebekka, et al.
Pubblicazione: (2025)
di: Görge, Rebekka, et al.
Pubblicazione: (2025)
Local Contrastive Editing of Gender Stereotypes
di: Lutz, Marlene, et al.
Pubblicazione: (2024)
di: Lutz, Marlene, et al.
Pubblicazione: (2024)
A Taxonomy of Stereotype Content in Large Language Models
di: Nicolas, Gandalf, et al.
Pubblicazione: (2024)
di: Nicolas, Gandalf, et al.
Pubblicazione: (2024)
FairMonitor: A Dual-framework for Detecting Stereotypes and Biases in Large Language Models
di: Bai, Yanhong, et al.
Pubblicazione: (2024)
di: Bai, Yanhong, et al.
Pubblicazione: (2024)
An Empirical Study of Gendered Stereotypes in Emotional Attributes for Bangla in Multilingual Large Language Models
di: Sadhu, Jayanta, et al.
Pubblicazione: (2024)
di: Sadhu, Jayanta, et al.
Pubblicazione: (2024)
Addressing Stereotypes in Large Language Models: A Critical Examination and Mitigation
di: Kazi, Fatima
Pubblicazione: (2025)
di: Kazi, Fatima
Pubblicazione: (2025)
REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning
di: Qureshi, Rameez, et al.
Pubblicazione: (2024)
di: Qureshi, Rameez, et al.
Pubblicazione: (2024)
Unmasking and Quantifying Racial Bias of Large Language Models in Medical Report Generation
di: Yang, Yifan, et al.
Pubblicazione: (2024)
di: Yang, Yifan, et al.
Pubblicazione: (2024)
LLMs Reproduce Stereotypes of Sexual and Gender Minorities
di: Ostrow, Ruby, et al.
Pubblicazione: (2025)
di: Ostrow, Ruby, et al.
Pubblicazione: (2025)
Angry Men, Sad Women: Large Language Models Reflect Gendered Stereotypes in Emotion Attribution
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
di: Robles, Melissa, et al.
Pubblicazione: (2025)
di: Robles, Melissa, et al.
Pubblicazione: (2025)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
DECASTE: Unveiling Caste Stereotypes in Large Language Models through Multi-Dimensional Bias Analysis
di: Vijayaraghavan, Prashanth, et al.
Pubblicazione: (2025)
di: Vijayaraghavan, Prashanth, et al.
Pubblicazione: (2025)
Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping
di: Aich, Ankit, et al.
Pubblicazione: (2024)
di: Aich, Ankit, et al.
Pubblicazione: (2024)
Quantifying Semantic Emergence in Language Models
di: Chen, Hang, et al.
Pubblicazione: (2024)
di: Chen, Hang, et al.
Pubblicazione: (2024)
On The Role of Reasoning in the Identification of Subtle Stereotypes in Natural Language
di: Tian, Jacob-Junqi, et al.
Pubblicazione: (2023)
di: Tian, Jacob-Junqi, et al.
Pubblicazione: (2023)
Large Language Models Reproduce Racial Stereotypes When Used for Text Annotation
di: Törnberg, Petter
Pubblicazione: (2026)
di: Törnberg, Petter
Pubblicazione: (2026)
Investigating Gender Stereotypes in Large Language Models via Social Determinants of Health
di: Ngo, Trung Hieu, et al.
Pubblicazione: (2026)
di: Ngo, Trung Hieu, et al.
Pubblicazione: (2026)
Stereotype or Personalization? User Identity Biases Chatbot Recommendations
di: Kantharuban, Anjali, et al.
Pubblicazione: (2024)
di: Kantharuban, Anjali, et al.
Pubblicazione: (2024)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
Quantifying Language Disparities in Multilingual Large Language Models
di: Hu, Songbo, et al.
Pubblicazione: (2025)
di: Hu, Songbo, et al.
Pubblicazione: (2025)
Diagnosing the Performance Trade-off in Moral Alignment: A Case Study on Gender Stereotypes
di: Liu, Guangliang, et al.
Pubblicazione: (2025)
di: Liu, Guangliang, et al.
Pubblicazione: (2025)
Biased or Flawed? Mitigating Stereotypes in Generative Language Models by Addressing Task-Specific Flaws
di: Jha, Akshita, et al.
Pubblicazione: (2024)
di: Jha, Akshita, et al.
Pubblicazione: (2024)
Analysing Differences in Persuasive Language in LLM-Generated Text: Uncovering Stereotypical Gender Patterns
di: Pauli, Amalie Brogaard, et al.
Pubblicazione: (2026)
di: Pauli, Amalie Brogaard, et al.
Pubblicazione: (2026)
More Women, Same Stereotypes: Unpacking the Gender Bias Paradox in Large Language Models
di: Chen, Evan, et al.
Pubblicazione: (2025)
di: Chen, Evan, et al.
Pubblicazione: (2025)
Dual Debiasing: Remove Stereotypes and Keep Factual Gender for Fair Language Modeling and Translation
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2025)
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2025)
Examining Alignment of Large Language Models through Representative Heuristics: The Case of Political Stereotypes
di: Jeoung, Sullam, et al.
Pubblicazione: (2025)
di: Jeoung, Sullam, et al.
Pubblicazione: (2025)
An Empirical Investigation of Gender Stereotype Representation in Large Language Models: The Italian Case
di: Giachino, Gioele, et al.
Pubblicazione: (2025)
di: Giachino, Gioele, et al.
Pubblicazione: (2025)
Profiling Bias in LLMs: Stereotype Dimensions in Contextual Word Embeddings
di: Schuster, Carolin M., et al.
Pubblicazione: (2024)
di: Schuster, Carolin M., et al.
Pubblicazione: (2024)
Documenti analoghi
-
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
di: Liu, Yiran, et al.
Pubblicazione: (2024) -
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
di: Nakanishi, Akito, et al.
Pubblicazione: (2025) -
Systematic Offensive Stereotyping (SOS) Bias in Language Models
di: Elsafoury, Fatma
Pubblicazione: (2023) -
Measuring Stereotype and Deviation Biases in Large Language Models
di: Wang, Daniel, et al.
Pubblicazione: (2025) -
A Survey on Stereotype Detection in Natural Language Processing
di: Cignarella, Alessandra Teresa, et al.
Pubblicazione: (2025)