That is Unacceptable: the Moral Foundations of Canceling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lo, Soda Marem, Araque, Oscar, Sharma, Rajesh, Stranisci, Marco Antonio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Moral Foundations Reddit Corpus
von: Trager, Jackson, et al.
Veröffentlicht: (2022)
von: Trager, Jackson, et al.
Veröffentlicht: (2022)
What Are They Filtering Out? An Experimental Benchmark of Filtering Strategies for Harm Reduction in Pretraining Datasets
von: Stranisci, Marco Antonio, et al.
Veröffentlicht: (2025)
von: Stranisci, Marco Antonio, et al.
Veröffentlicht: (2025)
LML: A Novel Lexicon for the Moral Foundation of Liberty
von: Araque, Oscar, et al.
Veröffentlicht: (2024)
von: Araque, Oscar, et al.
Veröffentlicht: (2024)
Investigating Political and Demographic Associations in Large Language Models Through Moral Foundations Theory
von: Smith-Vaniz, Nicole, et al.
Veröffentlicht: (2025)
von: Smith-Vaniz, Nicole, et al.
Veröffentlicht: (2025)
Analyzing Toxicity in Deep Conversations: A Reddit Case Study
von: Shankaran, Vigneshwaran, et al.
Veröffentlicht: (2024)
von: Shankaran, Vigneshwaran, et al.
Veröffentlicht: (2024)
What Makes AI Applications Acceptable or Unacceptable? A Predictive Moral Framework
von: Eriksson, Kimmo, et al.
Veröffentlicht: (2025)
von: Eriksson, Kimmo, et al.
Veröffentlicht: (2025)
MoralBERT: A Fine-Tuned Language Model for Capturing Moral Values in Social Discussions
von: Preniqi, Vjosa, et al.
Veröffentlicht: (2024)
von: Preniqi, Vjosa, et al.
Veröffentlicht: (2024)
Are Language Models Sensitive to Morally Irrelevant Distractors?
von: Shaw, Andrew, et al.
Veröffentlicht: (2026)
von: Shaw, Andrew, et al.
Veröffentlicht: (2026)
Polarization and Morality: Lexical Analysis of Abortion Discourse on Reddit
von: Stanier, Tessa, et al.
Veröffentlicht: (2024)
von: Stanier, Tessa, et al.
Veröffentlicht: (2024)
MoVa: Towards Generalizable Classification of Human Morals and Values
von: Chen, Ziyu, et al.
Veröffentlicht: (2025)
von: Chen, Ziyu, et al.
Veröffentlicht: (2025)
Recognition Without Authorization: LLMs and the Moral Order of Online Advice
von: van Nuenen, Tom
Veröffentlicht: (2026)
von: van Nuenen, Tom
Veröffentlicht: (2026)
Are you sure? Measuring models bias in content moderation through uncertainty
von: Urbinati, Alessandra, et al.
Veröffentlicht: (2025)
von: Urbinati, Alessandra, et al.
Veröffentlicht: (2025)
Beyond the Headlines: Understanding Sentiments and Morals Impacting Female Employment in Spain
von: Araque, Oscar, et al.
Veröffentlicht: (2024)
von: Araque, Oscar, et al.
Veröffentlicht: (2024)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
von: Vida, Karina, et al.
Veröffentlicht: (2024)
von: Vida, Karina, et al.
Veröffentlicht: (2024)
CogniAlign: Survivability-Grounded Multi-Agent Moral Reasoning for Safe and Transparent AI
von: Ali, Hasin Jawad, et al.
Veröffentlicht: (2025)
von: Ali, Hasin Jawad, et al.
Veröffentlicht: (2025)
Moral Mazes in the Era of LLMs
von: Nguyen, Dang, et al.
Veröffentlicht: (2026)
von: Nguyen, Dang, et al.
Veröffentlicht: (2026)
The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas
von: Wu, Ya, et al.
Veröffentlicht: (2025)
von: Wu, Ya, et al.
Veröffentlicht: (2025)
Analysis of Socially Unacceptable Discourse with Zero-shot Learning
von: Ghilene, Rayane, et al.
Veröffentlicht: (2024)
von: Ghilene, Rayane, et al.
Veröffentlicht: (2024)
The Moral Machine Experiment on Large Language Models
von: Takemoto, Kazuhiro
Veröffentlicht: (2023)
von: Takemoto, Kazuhiro
Veröffentlicht: (2023)
Decoding Memes: Benchmarking Narrative Role Classification across Multilingual and Multimodal Models
von: Sharma, Shivam, et al.
Veröffentlicht: (2025)
von: Sharma, Shivam, et al.
Veröffentlicht: (2025)
How Well Do LLMs Imitate Human Writing Style?
von: Jemama, Rebira, et al.
Veröffentlicht: (2025)
von: Jemama, Rebira, et al.
Veröffentlicht: (2025)
A Survey on Moral Foundation Theory and Pre-Trained Language Models: Current Advances and Challenges
von: Zangari, Lorenzo, et al.
Veröffentlicht: (2024)
von: Zangari, Lorenzo, et al.
Veröffentlicht: (2024)
Dropouts in Confidence: Moral Uncertainty in Human-LLM Alignment
von: Kwon, Jea, et al.
Veröffentlicht: (2025)
von: Kwon, Jea, et al.
Veröffentlicht: (2025)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
von: Nandi, Palash, et al.
Veröffentlicht: (2024)
von: Nandi, Palash, et al.
Veröffentlicht: (2024)
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
Moral Sparks in Social Media Narratives
von: Xi, Ruijie, et al.
Veröffentlicht: (2023)
von: Xi, Ruijie, et al.
Veröffentlicht: (2023)
FoundationalASSIST: An Educational Dataset for Foundational Knowledge Tracing and Pedagogical Grounding of LLMs
von: Worden, Eamon, et al.
Veröffentlicht: (2026)
von: Worden, Eamon, et al.
Veröffentlicht: (2026)
Attributions toward Artificial Agents in a modified Moral Turing Test
von: Aharoni, Eyal, et al.
Veröffentlicht: (2024)
von: Aharoni, Eyal, et al.
Veröffentlicht: (2024)
MemeMQA: Multimodal Question Answering for Memes via Rationale-Based Inferencing
von: Agarwal, Siddhant, et al.
Veröffentlicht: (2024)
von: Agarwal, Siddhant, et al.
Veröffentlicht: (2024)
GPT-4's One-Dimensional Mapping of Morality: How the Accuracy of Country-Estimates Depends on Moral Domain
von: Strimling, Pontus, et al.
Veröffentlicht: (2024)
von: Strimling, Pontus, et al.
Veröffentlicht: (2024)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
von: Hu, Yueqing, et al.
Veröffentlicht: (2026)
von: Hu, Yueqing, et al.
Veröffentlicht: (2026)
Moral Susceptibility and Robustness under Persona Role-Play in Large Language Models
von: Costa, Davi Bastos, et al.
Veröffentlicht: (2025)
von: Costa, Davi Bastos, et al.
Veröffentlicht: (2025)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
von: Backmann, Steffen, et al.
Veröffentlicht: (2025)
von: Backmann, Steffen, et al.
Veröffentlicht: (2025)
"More Than Words": Linking Music Preferences and Moral Values Through Lyrics
von: Preniqi, Vjosa, et al.
Veröffentlicht: (2022)
von: Preniqi, Vjosa, et al.
Veröffentlicht: (2022)
The Moral Gap of Large Language Models
von: Skorski, Maciej, et al.
Veröffentlicht: (2025)
von: Skorski, Maciej, et al.
Veröffentlicht: (2025)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
Medical Hallucinations in Foundation Models and Their Impact on Healthcare
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
Whose Emotions and Moral Sentiments Do Language Models Reflect?
von: He, Zihao, et al.
Veröffentlicht: (2024)
von: He, Zihao, et al.
Veröffentlicht: (2024)
Affective Computing Has Changed: The Foundation Model Disruption
von: Schuller, Björn, et al.
Veröffentlicht: (2024)
von: Schuller, Björn, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The Moral Foundations Reddit Corpus
von: Trager, Jackson, et al.
Veröffentlicht: (2022) -
What Are They Filtering Out? An Experimental Benchmark of Filtering Strategies for Harm Reduction in Pretraining Datasets
von: Stranisci, Marco Antonio, et al.
Veröffentlicht: (2025) -
LML: A Novel Lexicon for the Moral Foundation of Liberty
von: Araque, Oscar, et al.
Veröffentlicht: (2024) -
Investigating Political and Demographic Associations in Large Language Models Through Moral Foundations Theory
von: Smith-Vaniz, Nicole, et al.
Veröffentlicht: (2025) -
Analyzing Toxicity in Deep Conversations: A Reddit Case Study
von: Shankaran, Vigneshwaran, et al.
Veröffentlicht: (2024)