Local Contrastive Editing of Gender Stereotypes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lutz, Marlene, Choenni, Rochelle, Strohmaier, Markus, Lauscher, Anne |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Echoes of Multilinguality: Tracing Cultural Value Shifts during LM Fine-tuning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
Self-Alignment: Improving Alignment of Cultural Values in LLMs via In-Context Learning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
von: Nejadgholi, Isar, et al.
Veröffentlicht: (2024)
von: Nejadgholi, Isar, et al.
Veröffentlicht: (2024)
Toxic comments reduce the activity of volunteer editors on Wikipedia
von: Smirnov, Ivan, et al.
Veröffentlicht: (2023)
von: Smirnov, Ivan, et al.
Veröffentlicht: (2023)
M-Wanda: Improving One-Shot Pruning for Multilingual LLMs
von: Choenni, Rochelle, et al.
Veröffentlicht: (2025)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2025)
Extracting Affect Aggregates from Longitudinal Social Media Data with Temporal Adapters for Large Language Models
von: Ahnert, Georg, et al.
Veröffentlicht: (2024)
von: Ahnert, Georg, et al.
Veröffentlicht: (2024)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
von: Vida, Karina, et al.
Veröffentlicht: (2024)
von: Vida, Karina, et al.
Veröffentlicht: (2024)
Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models
von: Ahnert, Georg, et al.
Veröffentlicht: (2025)
von: Ahnert, Georg, et al.
Veröffentlicht: (2025)
Persona-driven Simulation of Voting Behavior in the European Parliament with Large Language Models
von: Kreutner, Maximilian, et al.
Veröffentlicht: (2025)
von: Kreutner, Maximilian, et al.
Veröffentlicht: (2025)
Do Psychometric Tests Work for Large Language Models? Evaluation of Tests on Sexism, Racism, and Morality
von: Jung, Jana, et al.
Veröffentlicht: (2025)
von: Jung, Jana, et al.
Veröffentlicht: (2025)
Missing the Margins: A Systematic Literature Review on the Demographic Representativeness of LLMs
von: Sen, Indira, et al.
Veröffentlicht: (2025)
von: Sen, Indira, et al.
Veröffentlicht: (2025)
How do languages influence each other? Studying cross-lingual data sharing during LM fine-tuning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2023)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2023)
QSTN: A Modular Framework for Robust Questionnaire Inference with Large Language Models
von: Kreutner, Maximilian, et al.
Veröffentlicht: (2025)
von: Kreutner, Maximilian, et al.
Veröffentlicht: (2025)
Robust Pronoun Fidelity with English LLMs: Are they Reasoning, Repeating, or Just Biased?
von: Gautam, Vagrant, et al.
Veröffentlicht: (2024)
von: Gautam, Vagrant, et al.
Veröffentlicht: (2024)
Reviewing the Reviewer: Elevating Peer Review Quality through LLM-Guided Feedback
von: Purkayastha, Sukannya, et al.
Veröffentlicht: (2026)
von: Purkayastha, Sukannya, et al.
Veröffentlicht: (2026)
Prompt Perturbations Reveal Human-Like Biases in Large Language Model Survey Responses
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
Properties of Group Fairness Metrics for Rankings
von: Schumacher, Tobias, et al.
Veröffentlicht: (2022)
von: Schumacher, Tobias, et al.
Veröffentlicht: (2022)
German General Social Survey Personas: A Survey-Derived Persona Prompt Collection for Population-Aligned LLM Studies
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
Stop! In the Name of Flaws: Disentangling Personal Names and Sociodemographic Attributes in NLP
von: Gautam, Vagrant, et al.
Veröffentlicht: (2024)
von: Gautam, Vagrant, et al.
Veröffentlicht: (2024)
The LLM Wears Prada: Analysing Gender Bias and Stereotypes through Online Shopping Data
von: Luca, Massimiliano, et al.
Veröffentlicht: (2025)
von: Luca, Massimiliano, et al.
Veröffentlicht: (2025)
The Prompt Makes the Person(a): A Systematic Evaluation of Sociodemographic Persona Prompting for Large Language Models
von: Lutz, Marlene, et al.
Veröffentlicht: (2025)
von: Lutz, Marlene, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Multilingual Text-to-Image Generation Magnifies Gender Stereotypes and Prompt Engineering May Not Help You
von: Friedrich, Felix, et al.
Veröffentlicht: (2024)
von: Friedrich, Felix, et al.
Veröffentlicht: (2024)
Metaphor Understanding Challenge Dataset for LLMs
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
An Empirical Investigation of Gender Stereotype Representation in Large Language Models: The Italian Case
von: Giachino, Gioele, et al.
Veröffentlicht: (2025)
von: Giachino, Gioele, et al.
Veröffentlicht: (2025)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
von: Robles, Melissa, et al.
Veröffentlicht: (2025)
von: Robles, Melissa, et al.
Veröffentlicht: (2025)
A Survey on Stereotype Detection in Natural Language Processing
von: Cignarella, Alessandra Teresa, et al.
Veröffentlicht: (2025)
von: Cignarella, Alessandra Teresa, et al.
Veröffentlicht: (2025)
Ethical Concern Identification in NLP: A Corpus of ACL Anthology Ethics Statements
von: Karamolegkou, Antonia, et al.
Veröffentlicht: (2024)
von: Karamolegkou, Antonia, et al.
Veröffentlicht: (2024)
Leveraging Machine Learning to Identify Gendered Stereotypes and Body Image Concerns on Diet and Fitness Online Forums
von: Chu, Minh Duc, et al.
Veröffentlicht: (2024)
von: Chu, Minh Duc, et al.
Veröffentlicht: (2024)
Building Bridges: A Dataset for Evaluating Gender-Fair Machine Translation into German
von: Lardelli, Manuel, et al.
Veröffentlicht: (2024)
von: Lardelli, Manuel, et al.
Veröffentlicht: (2024)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
Best-of-L: Cross-Lingual Reward Modeling for Mathematical Reasoning
von: Rajaee, Sara, et al.
Veröffentlicht: (2025)
von: Rajaee, Sara, et al.
Veröffentlicht: (2025)
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
von: Nakanishi, Akito, et al.
Veröffentlicht: (2025)
von: Nakanishi, Akito, et al.
Veröffentlicht: (2025)
Who is better at math, Jenny or Jingzhen? Uncovering Stereotypes in Large Language Models
von: Siddique, Zara, et al.
Veröffentlicht: (2024)
von: Siddique, Zara, et al.
Veröffentlicht: (2024)
Surfacing Subtle Stereotypes: A Multilingual, Debate-Oriented Evaluation of Modern LLMs
von: Saeed, Muhammed, et al.
Veröffentlicht: (2025)
von: Saeed, Muhammed, et al.
Veröffentlicht: (2025)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
DECASTE: Unveiling Caste Stereotypes in Large Language Models through Multi-Dimensional Bias Analysis
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2025)
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2025)
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
von: Liu, Yiran, et al.
Veröffentlicht: (2024)
von: Liu, Yiran, et al.
Veröffentlicht: (2024)
Finding Culture-Sensitive Neurons in Vision-Language Models
von: Zhao, Xiutian, et al.
Veröffentlicht: (2025)
von: Zhao, Xiutian, et al.
Veröffentlicht: (2025)
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The Echoes of Multilinguality: Tracing Cultural Value Shifts during LM Fine-tuning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024) -
Self-Alignment: Improving Alignment of Cultural Values in LLMs via In-Context Learning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024) -
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
von: Nejadgholi, Isar, et al.
Veröffentlicht: (2024) -
Toxic comments reduce the activity of volunteer editors on Wikipedia
von: Smirnov, Ivan, et al.
Veröffentlicht: (2023) -
M-Wanda: Improving One-Shot Pruning for Multilingual LLMs
von: Choenni, Rochelle, et al.
Veröffentlicht: (2025)