Characterizing Stereotypical Bias from Privacy-preserving Pre-Training
Fuente:
arXiv
Saved in:
| Main Authors: | Arnold, Stefan, Gröbner, Rene, Schreiner, Annika |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Steering Prepositional Phrases in Language Models: A Case of with-headed Adjectival and Adverbial Complements in Gemma-2
by: Arnold, Stefan, et al.
Published: (2025)
by: Arnold, Stefan, et al.
Published: (2025)
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
by: Masoudian, Shahed, et al.
Published: (2025)
by: Masoudian, Shahed, et al.
Published: (2025)
Experiments in News Bias Detection with Pre-Trained Neural Transformers
by: Menzner, Tim, et al.
Published: (2024)
by: Menzner, Tim, et al.
Published: (2024)
More Women, Same Stereotypes: Unpacking the Gender Bias Paradox in Large Language Models
by: Chen, Evan, et al.
Published: (2025)
by: Chen, Evan, et al.
Published: (2025)
A Stereotype Content Analysis on Color-related Social Bias in Large Vision Language Models
by: Choi, Junhyuk, et al.
Published: (2025)
by: Choi, Junhyuk, et al.
Published: (2025)
The LLM Wears Prada: Analysing Gender Bias and Stereotypes through Online Shopping Data
by: Luca, Massimiliano, et al.
Published: (2025)
by: Luca, Massimiliano, et al.
Published: (2025)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
by: Powers, Maximus, et al.
Published: (2024)
by: Powers, Maximus, et al.
Published: (2024)
ASCenD-BDS: Adaptable, Stochastic and Context-aware framework for Detection of Bias, Discrimination and Stereotyping
by: Bahl, Rajiv, et al.
Published: (2025)
by: Bahl, Rajiv, et al.
Published: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Bias Dynamics in BabyLMs: Towards a Compute-Efficient Sandbox for Democratising Pre-Training Debiasing
by: Trhlik, Filip, et al.
Published: (2026)
by: Trhlik, Filip, et al.
Published: (2026)
Enhancing Small Medical Learners with Privacy-preserving Contextual Prompting
by: Zhang, Xinlu, et al.
Published: (2023)
by: Zhang, Xinlu, et al.
Published: (2023)
AfriStereo: A Culturally Grounded Dataset for Evaluating Stereotypical Bias in Large Language Models
by: Beux, Yann Le, et al.
Published: (2025)
by: Beux, Yann Le, et al.
Published: (2025)
PLACID: Privacy-preserving Large language models for Acronym Clinical Inference and Disambiguation
by: Aithal, Manjushree B., et al.
Published: (2026)
by: Aithal, Manjushree B., et al.
Published: (2026)
Can We Locate and Prevent Stereotypes in LLMs?
by: D'Souza, Alex
Published: (2026)
by: D'Souza, Alex
Published: (2026)
Rethinking Reflection in Pre-Training
by: AI, Essential, et al.
Published: (2025)
by: AI, Essential, et al.
Published: (2025)
Locating and Editing Figure-Ground Organization in Vision Transformers
by: Arnold, Stefan, et al.
Published: (2026)
by: Arnold, Stefan, et al.
Published: (2026)
Are Stereotypes Leading LLMs' Zero-Shot Stance Detection ?
by: Dubreuil, Anthony, et al.
Published: (2025)
by: Dubreuil, Anthony, et al.
Published: (2025)
Cultural Bias and Cultural Alignment of Large Language Models
by: Tao, Yan, et al.
Published: (2023)
by: Tao, Yan, et al.
Published: (2023)
Unraveling Emotions with Pre-Trained Models
by: Pajón-Sanmartín, Alejandro, et al.
Published: (2025)
by: Pajón-Sanmartín, Alejandro, et al.
Published: (2025)
Evaluation of Large Language Models: STEM education and Gender Stereotypes
by: Due, Smilla, et al.
Published: (2024)
by: Due, Smilla, et al.
Published: (2024)
Detecting Linguistic Indicators for Stereotype Assessment with Large Language Models
by: Görge, Rebekka, et al.
Published: (2025)
by: Görge, Rebekka, et al.
Published: (2025)
Evolution of Concepts in Language Model Pre-Training
by: Ge, Xuyang, et al.
Published: (2025)
by: Ge, Xuyang, et al.
Published: (2025)
REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning
by: Qureshi, Rameez, et al.
Published: (2024)
by: Qureshi, Rameez, et al.
Published: (2024)
Stereotype Detection in LLMs: A Multiclass, Explainable, and Benchmark-Driven Approach
by: Wu, Zekun, et al.
Published: (2024)
by: Wu, Zekun, et al.
Published: (2024)
Blind Men and the Elephant: Diverse Perspectives on Gender Stereotypes in Benchmark Datasets
by: Zakizadeh, Mahdi, et al.
Published: (2025)
by: Zakizadeh, Mahdi, et al.
Published: (2025)
Med-UniC: Unifying Cross-Lingual Medical Vision-Language Pre-Training by Diminishing Bias
by: Wan, Zhongwei, et al.
Published: (2023)
by: Wan, Zhongwei, et al.
Published: (2023)
Revealing the Inherent Instructability of Pre-Trained Language Models
by: An, Seokhyun, et al.
Published: (2024)
by: An, Seokhyun, et al.
Published: (2024)
Beyond Public Access in LLM Pre-Training Data
by: Rosenblat, Sruly, et al.
Published: (2025)
by: Rosenblat, Sruly, et al.
Published: (2025)
Identifying Gender Stereotypes and Biases in Automated Translation from English to Italian using Similarity Networks
by: Mohammadi, Fatemeh, et al.
Published: (2025)
by: Mohammadi, Fatemeh, et al.
Published: (2025)
Large Language Models Reproduce Racial Stereotypes When Used for Text Annotation
by: Törnberg, Petter
Published: (2026)
by: Törnberg, Petter
Published: (2026)
Investigating Gender Stereotypes in Large Language Models via Social Determinants of Health
by: Ngo, Trung Hieu, et al.
Published: (2026)
by: Ngo, Trung Hieu, et al.
Published: (2026)
"According to ...": Prompting Language Models Improves Quoting from Pre-Training Data
by: Weller, Orion, et al.
Published: (2023)
by: Weller, Orion, et al.
Published: (2023)
Bias-Augmented Consistency Training Reduces Biased Reasoning in Chain-of-Thought
by: Chua, James, et al.
Published: (2024)
by: Chua, James, et al.
Published: (2024)
Detecting and Mitigating Bias in LLMs through Knowledge Graph-Augmented Training
by: Kumar, Rajeev, et al.
Published: (2025)
by: Kumar, Rajeev, et al.
Published: (2025)
Document-Level In-Context Few-Shot Relation Extraction via Pre-Trained Language Models
by: Ozyurt, Yilmazcan, et al.
Published: (2023)
by: Ozyurt, Yilmazcan, et al.
Published: (2023)
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
by: Wehnert, Sabine, et al.
Published: (2025)
by: Wehnert, Sabine, et al.
Published: (2025)
Unifying Structured Data as Graph for Data-to-Text Pre-Training
by: Li, Shujie, et al.
Published: (2024)
by: Li, Shujie, et al.
Published: (2024)
Pre-Trained Language Models for Keyphrase Prediction: A Review
by: Umair, Muhammad, et al.
Published: (2024)
by: Umair, Muhammad, et al.
Published: (2024)
Pre-Training Curriculum for Multi-Token Prediction in Language Models
by: Aynetdinov, Ansar, et al.
Published: (2025)
by: Aynetdinov, Ansar, et al.
Published: (2025)
Does Differential Privacy Impact Bias in Pretrained NLP Models?
by: Islam, Md. Khairul, et al.
Published: (2024)
by: Islam, Md. Khairul, et al.
Published: (2024)
Similar Items
-
Steering Prepositional Phrases in Language Models: A Case of with-headed Adjectival and Adverbial Complements in Gemma-2
by: Arnold, Stefan, et al.
Published: (2025) -
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
by: Masoudian, Shahed, et al.
Published: (2025) -
Experiments in News Bias Detection with Pre-Trained Neural Transformers
by: Menzner, Tim, et al.
Published: (2024) -
More Women, Same Stereotypes: Unpacking the Gender Bias Paradox in Large Language Models
by: Chen, Evan, et al.
Published: (2025) -
A Stereotype Content Analysis on Color-related Social Bias in Large Vision Language Models
by: Choi, Junhyuk, et al.
Published: (2025)