Simulating Identity, Propagating Bias: Abstraction and Stereotypes in LLM-Generated Text
Fuente:
arXiv
Salvato in:
| Autori principali: | Sommerauer, Pia, Rambelli, Giulia, Caselli, Tommaso |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving Causal Interventions in Amnesic Probing with Mean Projection or LEACE
di: Dobrzeniecka, Alicja, et al.
Pubblicazione: (2025)
di: Dobrzeniecka, Alicja, et al.
Pubblicazione: (2025)
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
di: Masoudian, Shahed, et al.
Pubblicazione: (2025)
di: Masoudian, Shahed, et al.
Pubblicazione: (2025)
ARN: Analogical Reasoning on Narratives
di: Sourati, Zhivar, et al.
Pubblicazione: (2023)
di: Sourati, Zhivar, et al.
Pubblicazione: (2023)
Composing or Not Composing? Towards Distributional Construction Grammars
di: Blache, Philippe, et al.
Pubblicazione: (2024)
di: Blache, Philippe, et al.
Pubblicazione: (2024)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
di: Raza, Shaina, et al.
Pubblicazione: (2025)
di: Raza, Shaina, et al.
Pubblicazione: (2025)
The Social Cost of Intelligence: Emergence, Propagation, and Amplification of Stereotypical Bias in Multi-Agent Systems
di: Nguyen, Thi-Nhung, et al.
Pubblicazione: (2025)
di: Nguyen, Thi-Nhung, et al.
Pubblicazione: (2025)
How Humans and LLMs Organize Conceptual Knowledge: Exploring Subordinate Categories in Italian
di: Pedrotti, Andrea, et al.
Pubblicazione: (2025)
di: Pedrotti, Andrea, et al.
Pubblicazione: (2025)
HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation
di: Deng, Zewei, et al.
Pubblicazione: (2026)
di: Deng, Zewei, et al.
Pubblicazione: (2026)
Analysing Differences in Persuasive Language in LLM-Generated Text: Uncovering Stereotypical Gender Patterns
di: Pauli, Amalie Brogaard, et al.
Pubblicazione: (2026)
di: Pauli, Amalie Brogaard, et al.
Pubblicazione: (2026)
Stereotype or Personalization? User Identity Biases Chatbot Recommendations
di: Kantharuban, Anjali, et al.
Pubblicazione: (2024)
di: Kantharuban, Anjali, et al.
Pubblicazione: (2024)
Hummus: A Dataset of Humorous Multimodal Metaphor Use
di: Tong, Xiaoyu, et al.
Pubblicazione: (2025)
di: Tong, Xiaoyu, et al.
Pubblicazione: (2025)
When Stereotypes GTG: The Impact of Predictive Text Suggestions on Gender Bias in Human-AI Co-Writing
di: Baumler, Connor, et al.
Pubblicazione: (2024)
di: Baumler, Connor, et al.
Pubblicazione: (2024)
Systematic Offensive Stereotyping (SOS) Bias in Language Models
di: Elsafoury, Fatma
Pubblicazione: (2023)
di: Elsafoury, Fatma
Pubblicazione: (2023)
The LLM Wears Prada: Analysing Gender Bias and Stereotypes through Online Shopping Data
di: Luca, Massimiliano, et al.
Pubblicazione: (2025)
di: Luca, Massimiliano, et al.
Pubblicazione: (2025)
Parallel LLM Reasoning for Bias-Resilient, Robust Conceptual Abstraction
di: Adeseye, Aisvarya, et al.
Pubblicazione: (2026)
di: Adeseye, Aisvarya, et al.
Pubblicazione: (2026)
Annotating Constructions with UD: the experience of the Italian Constructicon
di: Pannitto, Ludovica, et al.
Pubblicazione: (2024)
di: Pannitto, Ludovica, et al.
Pubblicazione: (2024)
Do Multilingual Large Language Models Mitigate Stereotype Bias?
di: Nie, Shangrui, et al.
Pubblicazione: (2024)
di: Nie, Shangrui, et al.
Pubblicazione: (2024)
Profiling Bias in LLMs: Stereotype Dimensions in Contextual Word Embeddings
di: Schuster, Carolin M., et al.
Pubblicazione: (2024)
di: Schuster, Carolin M., et al.
Pubblicazione: (2024)
When Hate Meets Facts: LLMs-in-the-Loop for Check-worthiness Detection in Hate Speech
di: Ocampo, Nicolás Benjamín, et al.
Pubblicazione: (2026)
di: Ocampo, Nicolás Benjamín, et al.
Pubblicazione: (2026)
Word Ladders: A Mobile Application for Semantic Data Collection
di: Bolognesi, Marianna Marcella, et al.
Pubblicazione: (2024)
di: Bolognesi, Marianna Marcella, et al.
Pubblicazione: (2024)
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
di: Liu, Yiran, et al.
Pubblicazione: (2024)
di: Liu, Yiran, et al.
Pubblicazione: (2024)
Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian
di: Kalhor, Ghazal, et al.
Pubblicazione: (2025)
di: Kalhor, Ghazal, et al.
Pubblicazione: (2025)
Characterizing Stereotypical Bias from Privacy-preserving Pre-Training
di: Arnold, Stefan, et al.
Pubblicazione: (2024)
di: Arnold, Stefan, et al.
Pubblicazione: (2024)
MaskSQL: Safeguarding Privacy for LLM-Based Text-to-SQL via Abstraction
di: Abedini, Sepideh, et al.
Pubblicazione: (2025)
di: Abedini, Sepideh, et al.
Pubblicazione: (2025)
FairI Tales: Evaluation of Fairness in Indian Contexts with a Focus on Bias and Stereotypes
di: Nawale, Janki Atul, et al.
Pubblicazione: (2025)
di: Nawale, Janki Atul, et al.
Pubblicazione: (2025)
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
di: Jha, Akshita, et al.
Pubblicazione: (2024)
di: Jha, Akshita, et al.
Pubblicazione: (2024)
Stereotype Detection as a Catalyst for Enhanced Bias Detection: A Multi-Task Learning Approach
di: Tomar, Aditya, et al.
Pubblicazione: (2025)
di: Tomar, Aditya, et al.
Pubblicazione: (2025)
DECASTE: Unveiling Caste Stereotypes in Large Language Models through Multi-Dimensional Bias Analysis
di: Vijayaraghavan, Prashanth, et al.
Pubblicazione: (2025)
di: Vijayaraghavan, Prashanth, et al.
Pubblicazione: (2025)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
di: Plaza-del-Arco, Flor Miriam, et al.
Pubblicazione: (2024)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
di: Robles, Melissa, et al.
Pubblicazione: (2025)
di: Robles, Melissa, et al.
Pubblicazione: (2025)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
di: Powers, Maximus, et al.
Pubblicazione: (2024)
di: Powers, Maximus, et al.
Pubblicazione: (2024)
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
di: Sarti, Gabriele, et al.
Pubblicazione: (2024)
di: Sarti, Gabriele, et al.
Pubblicazione: (2024)
Multilingual Text-to-Image Generation Magnifies Gender Stereotypes and Prompt Engineering May Not Help You
di: Friedrich, Felix, et al.
Pubblicazione: (2024)
di: Friedrich, Felix, et al.
Pubblicazione: (2024)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
HEARTS: A Holistic Framework for Explainable, Sustainable and Robust Text Stereotype Detection
di: King, Theo, et al.
Pubblicazione: (2024)
di: King, Theo, et al.
Pubblicazione: (2024)
Reading Between the Prompts: How Stereotypes Shape LLM's Implicit Personalization
di: Neplenbroek, Vera, et al.
Pubblicazione: (2025)
di: Neplenbroek, Vera, et al.
Pubblicazione: (2025)
Redirected, Not Removed: Task-Dependent Stereotyping Reveals the Limits of LLM Alignments
di: Kumar, Divyanshu, et al.
Pubblicazione: (2026)
di: Kumar, Divyanshu, et al.
Pubblicazione: (2026)
Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training
di: Rahmati, Elnaz, et al.
Pubblicazione: (2026)
di: Rahmati, Elnaz, et al.
Pubblicazione: (2026)
More Women, Same Stereotypes: Unpacking the Gender Bias Paradox in Large Language Models
di: Chen, Evan, et al.
Pubblicazione: (2025)
di: Chen, Evan, et al.
Pubblicazione: (2025)
Quantifying Stereotypes in Language
di: Liu, Yang
Pubblicazione: (2024)
di: Liu, Yang
Pubblicazione: (2024)
Documenti analoghi
-
Improving Causal Interventions in Amnesic Probing with Mean Projection or LEACE
di: Dobrzeniecka, Alicja, et al.
Pubblicazione: (2025) -
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
di: Masoudian, Shahed, et al.
Pubblicazione: (2025) -
ARN: Analogical Reasoning on Narratives
di: Sourati, Zhivar, et al.
Pubblicazione: (2023) -
Composing or Not Composing? Towards Distributional Construction Grammars
di: Blache, Philippe, et al.
Pubblicazione: (2024) -
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
di: Raza, Shaina, et al.
Pubblicazione: (2025)