Adaptable Moral Stances of Large Language Models on Sexist Content: Implications for Society and Gender Discourse
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Rongchen, Nejadgholi, Isar, Dawkins, Hillary, Fraser, Kathleen C., Kiritchenko, Svetlana |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fine-Tuning Lowers Safety and Disrupts Evaluation Consistency
by: Fraser, Kathleen C., et al.
Published: (2025)
by: Fraser, Kathleen C., et al.
Published: (2025)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
by: Nejadgholi, Isar, et al.
Published: (2024)
by: Nejadgholi, Isar, et al.
Published: (2024)
When Detection Fails: The Power of Fine-Tuned Models to Generate Human-Like Social Media Text
by: Dawkins, Hillary, et al.
Published: (2025)
by: Dawkins, Hillary, et al.
Published: (2025)
Detecting AI-Generated Text: Factors Influencing Detectability with Current Methods
by: Fraser, Kathleen C., et al.
Published: (2024)
by: Fraser, Kathleen C., et al.
Published: (2024)
The crime of being poor
by: Curto, Georgina, et al.
Published: (2023)
by: Curto, Georgina, et al.
Published: (2023)
Gender-Neutral Machine Translation Strategies in Practice
by: Dawkins, Hillary, et al.
Published: (2025)
by: Dawkins, Hillary, et al.
Published: (2025)
WMT24 Test Suite: Gender Resolution in Speaker-Listener Dialogue Roles
by: Dawkins, Hillary, et al.
Published: (2024)
by: Dawkins, Hillary, et al.
Published: (2024)
Projective Methods for Mitigating Gender Bias in Pre-trained Language Models
by: Dawkins, Hillary, et al.
Published: (2024)
by: Dawkins, Hillary, et al.
Published: (2024)
Examining Gender and Racial Bias in Large Vision-Language Models Using a Novel Dataset of Parallel Images
by: Fraser, Kathleen C., et al.
Published: (2024)
by: Fraser, Kathleen C., et al.
Published: (2024)
Tackling Social Bias against the Poor: A Dataset and Taxonomy on Aporophobia
by: Curto, Georgina, et al.
Published: (2025)
by: Curto, Georgina, et al.
Published: (2025)
From Perceived Effectiveness to Measured Impact: Identity-Aware Evaluation of Automated Counter-Stereotypes
by: Kiritchenko, Svetlana, et al.
Published: (2025)
by: Kiritchenko, Svetlana, et al.
Published: (2025)
Semantic Differentiation in Speech Emotion Recognition: Insights from Descriptive and Expressive Speech Roles
by: Guo, Rongchen, et al.
Published: (2025)
by: Guo, Rongchen, et al.
Published: (2025)
Socially Aware Synthetic Data Generation for Suicidal Ideation Detection Using Large Language Models
by: Ghanadian, Hamideh, et al.
Published: (2024)
by: Ghanadian, Hamideh, et al.
Published: (2024)
Uncovering Bias in Large Vision-Language Models at Scale with Counterfactuals
by: Howard, Phillip, et al.
Published: (2024)
by: Howard, Phillip, et al.
Published: (2024)
Uncovering Bias in Large Vision-Language Models with Counterfactuals
by: Howard, Phillip, et al.
Published: (2024)
by: Howard, Phillip, et al.
Published: (2024)
A Taxonomy for Design and Evaluation of Prompt-Based Natural Language Explanations
by: Nejadgholi, Isar, et al.
Published: (2025)
by: Nejadgholi, Isar, et al.
Published: (2025)
Defining Cultural Capabilities for AI Evaluation: A Taxonomy Grounded in Intercultural Communication Theory
by: Nejadgholi, Isar, et al.
Published: (2026)
by: Nejadgholi, Isar, et al.
Published: (2026)
LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories
by: Vishnubhotla, Krishnapriya, et al.
Published: (2026)
by: Vishnubhotla, Krishnapriya, et al.
Published: (2026)
Chain of Stance: Stance Detection with Large Language Models
by: Ma, Junxia, et al.
Published: (2024)
by: Ma, Junxia, et al.
Published: (2024)
Cross-Cultural Value Awareness in Large Vision-Language Models
by: Howard, Phillip, et al.
Published: (2026)
by: Howard, Phillip, et al.
Published: (2026)
Filipino Benchmarks for Measuring Sexist and Homophobic Bias in Multilingual Language Models from Southeast Asia
by: Gamboa, Lance Calvin Lim, et al.
Published: (2024)
by: Gamboa, Lance Calvin Lim, et al.
Published: (2024)
SafeSpeech: A Comprehensive and Interactive Tool for Analysing Sexist and Abusive Language in Conversations
by: Tan, Xingwei, et al.
Published: (2025)
by: Tan, Xingwei, et al.
Published: (2025)
CLA Guidelines on Non-Sexist Language
Published: (1976)
Published: (1976)
Adaptable Logical Control for Large Language Models
by: Zhang, Honghua, et al.
Published: (2024)
by: Zhang, Honghua, et al.
Published: (2024)
Can Large Language Models Address Open-Target Stance Detection?
by: Akash, Abu Ubaida, et al.
Published: (2024)
by: Akash, Abu Ubaida, et al.
Published: (2024)
Belief Revision: The Adaptability of Large Language Models Reasoning
by: Wilie, Bryan, et al.
Published: (2024)
by: Wilie, Bryan, et al.
Published: (2024)
Gender Bias in Large Language Models for Healthcare: Assignment Consistency and Clinical Implications
by: Liu, Mingxuan, et al.
Published: (2025)
by: Liu, Mingxuan, et al.
Published: (2025)
Reinforcement Tuning for Detecting Stances and Debunking Rumors Jointly with Large Language Models
by: Yang, Ruichao, et al.
Published: (2024)
by: Yang, Ruichao, et al.
Published: (2024)
Understanding and Mitigating Political Stance Cross-topic Generalization in Large Language Models
by: Zhang, Jiayi, et al.
Published: (2025)
by: Zhang, Jiayi, et al.
Published: (2025)
Mitigating Biases of Large Language Models in Stance Detection with Counterfactual Augmented Calibration
by: Li, Ang, et al.
Published: (2024)
by: Li, Ang, et al.
Published: (2024)
Masculine Defaults via Gendered Discourse in Podcasts and Large Language Models
by: Teleki, Maria, et al.
Published: (2025)
by: Teleki, Maria, et al.
Published: (2025)
Enhancing Stance Classification on Social Media Using Quantified Moral Foundations
by: Zhang, Hong, et al.
Published: (2023)
by: Zhang, Hong, et al.
Published: (2023)
Adaptable and Reliable Text Classification using Large Language Models
by: Wang, Zhiqiang, et al.
Published: (2024)
by: Wang, Zhiqiang, et al.
Published: (2024)
Predicting User Stances from Target-Agnostic Information using Large Language Models
by: Loh, Siyuan Brandon, et al.
Published: (2024)
by: Loh, Siyuan Brandon, et al.
Published: (2024)
Language Independent Stance Detection: Social Interaction-based Embeddings and Large Language Models
by: de Landa, Joseba Fernandez, et al.
Published: (2022)
by: de Landa, Joseba Fernandez, et al.
Published: (2022)
Slogans or Stance? A Label-Light Diagnostic for Entrepreneurial-Discourse Measurement on Chinese SOE Speeches
by: Gong, Ting, et al.
Published: (2026)
by: Gong, Ting, et al.
Published: (2026)
Stance Detection on Social Media with Fine-Tuned Large Language Models
by: Gül, İlker, et al.
Published: (2024)
by: Gül, İlker, et al.
Published: (2024)
Social and Ethical Risks Posed by General-Purpose LLMs for Settling Newcomers in Canada
by: Nejadgholi, Isar, et al.
Published: (2024)
by: Nejadgholi, Isar, et al.
Published: (2024)
Mind the Shift: Decoding Monetary Policy Stance from FOMC Statements with Large Language Models
by: Tang, Yixuan, et al.
Published: (2026)
by: Tang, Yixuan, et al.
Published: (2026)
Examining the Influence of Political Bias on Large Language Model Performance in Stance Classification
by: Ng, Lynnette Hui Xian, et al.
Published: (2024)
by: Ng, Lynnette Hui Xian, et al.
Published: (2024)
Similar Items
-
Fine-Tuning Lowers Safety and Disrupts Evaluation Consistency
by: Fraser, Kathleen C., et al.
Published: (2025) -
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
by: Nejadgholi, Isar, et al.
Published: (2024) -
When Detection Fails: The Power of Fine-Tuned Models to Generate Human-Like Social Media Text
by: Dawkins, Hillary, et al.
Published: (2025) -
Detecting AI-Generated Text: Factors Influencing Detectability with Current Methods
by: Fraser, Kathleen C., et al.
Published: (2024) -
The crime of being poor
by: Curto, Georgina, et al.
Published: (2023)