A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance
Fuente:
arXiv
Saved in:
| Main Authors: | Melis, Matteo, Lapesa, Gabriella, Assenmacher, Dennis |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tell Me What You Know About Sexism: Expert-LLM Interaction Strategies and Co-Created Definitions for Zero-Shot Sexism Detection
by: Reuver, Myrthe, et al.
Published: (2025)
by: Reuver, Myrthe, et al.
Published: (2025)
Who and What? Using Linguistic Features and Annotator Characteristics to Analyze Annotation Variation
by: Maurer, Maximilian, et al.
Published: (2026)
by: Maurer, Maximilian, et al.
Published: (2026)
The Unseen Targets of Hate -- A Systematic Review of Hateful Communication Datasets
by: Yu, Zehui, et al.
Published: (2024)
by: Yu, Zehui, et al.
Published: (2024)
Towards Weakly-Supervised Hate Speech Classification Across Datasets
by: Jin, Yiping, et al.
Published: (2023)
by: Jin, Yiping, et al.
Published: (2023)
From Emotion to Expression: Theoretical Foundations and Resources for Fear Speech
by: Shankaran, Vigneshwaran, et al.
Published: (2026)
by: Shankaran, Vigneshwaran, et al.
Published: (2026)
Learning from Convenience Samples: A Case Study on Fine-Tuning LLMs for Survey Non-response in the German Longitudinal Election Study
by: Holtdirk, Tobias, et al.
Published: (2025)
by: Holtdirk, Tobias, et al.
Published: (2025)
The Enforcement and Feasibility of Hate Speech Moderation on Twitter
by: Tonneau, Manuel, et al.
Published: (2026)
by: Tonneau, Manuel, et al.
Published: (2026)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
by: Jin, Yiping, et al.
Published: (2024)
by: Jin, Yiping, et al.
Published: (2024)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
by: Yang, Xilin
Published: (2024)
by: Yang, Xilin
Published: (2024)
Can Prompting LLMs Unlock Hate Speech Detection across Languages? A Zero-shot and Few-shot Study
by: Ghorbanpour, Faeze, et al.
Published: (2025)
by: Ghorbanpour, Faeze, et al.
Published: (2025)
People Make Better Edits: Measuring the Efficacy of LLM-Generated Counterfactually Augmented Data for Harmful Language Detection
by: Sen, Indira, et al.
Published: (2023)
by: Sen, Indira, et al.
Published: (2023)
Few-shot Hate Speech Detection Based on the MindSpore Framework
by: Qin, Zhenkai, et al.
Published: (2025)
by: Qin, Zhenkai, et al.
Published: (2025)
Beyond Hate: Differentiating Uncivil and Intolerant Speech in Multimodal Content Moderation
by: Herrmann, Nils A., et al.
Published: (2026)
by: Herrmann, Nils A., et al.
Published: (2026)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
by: Nandi, Palash, et al.
Published: (2024)
by: Nandi, Palash, et al.
Published: (2024)
Towards Fairness Assessment of Dutch Hate Speech Detection
by: Bauer, Julie, et al.
Published: (2025)
by: Bauer, Julie, et al.
Published: (2025)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
by: Gajewska, Ewelina, et al.
Published: (2025)
by: Gajewska, Ewelina, et al.
Published: (2025)
Rethinking Hate Speech Detection on Social Media: Can LLMs Replace Traditional Models?
by: Singh, Daman Deep, et al.
Published: (2025)
by: Singh, Daman Deep, et al.
Published: (2025)
Zero-Shot Hierarchical Classification on the Common Procurement Vocabulary Taxonomy
by: Moiraghi, Federico, et al.
Published: (2024)
by: Moiraghi, Federico, et al.
Published: (2024)
Investigating Subjective Factors of Argument Strength: Storytelling, Emotions, and Hedging
by: Quensel, Carlotta, et al.
Published: (2025)
by: Quensel, Carlotta, et al.
Published: (2025)
Deep Learning Approaches for Detecting Adversarial Cyberbullying and Hate Speech in Social Networks
by: Azumah, Sylvia Worlali, et al.
Published: (2024)
by: Azumah, Sylvia Worlali, et al.
Published: (2024)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
Task-Dependent Evaluation of LLM Output Homogenization: A Taxonomy-Guided Framework
by: Jain, Shomik, et al.
Published: (2025)
by: Jain, Shomik, et al.
Published: (2025)
Fair Play in the Newsroom: Actor-Based Filtering Gender Discrimination in Text Corpora
by: Urchs, Stefanie, et al.
Published: (2025)
by: Urchs, Stefanie, et al.
Published: (2025)
Data-Efficient Hate Speech Detection via Cross-Lingual Nearest Neighbor Retrieval with Limited Labeled Data
by: Ghorbanpour, Faeze, et al.
Published: (2025)
by: Ghorbanpour, Faeze, et al.
Published: (2025)
Hate Personified: Investigating the role of LLMs in content moderation
by: Masud, Sarah, et al.
Published: (2024)
by: Masud, Sarah, et al.
Published: (2024)
ReZG: Retrieval-Augmented Zero-Shot Counter Narrative Generation for Hate Speech
by: Jiang, Shuyu, et al.
Published: (2023)
by: Jiang, Shuyu, et al.
Published: (2023)
Generalizing Hate Speech Detection Using Multi-Task Learning: A Case Study of Political Public Figures
by: Yuan, Lanqin, et al.
Published: (2022)
by: Yuan, Lanqin, et al.
Published: (2022)
Multilingualism, Transnationality, and K-pop in the Online #StopAsianHate Movement
by: Masis, Tessa, et al.
Published: (2025)
by: Masis, Tessa, et al.
Published: (2025)
Towards a Perspectivist Turn in Argument Quality Assessment
by: Romberg, Julia, et al.
Published: (2025)
by: Romberg, Julia, et al.
Published: (2025)
Toeing the Party Line: Election Manifestos as a Key to Understand Political Discourse on Twitter
by: Maurer, Maximilian, et al.
Published: (2024)
by: Maurer, Maximilian, et al.
Published: (2024)
Improving Hate Speech Classification with Cross-Taxonomy Dataset Integration
by: Fillies, Jan, et al.
Published: (2025)
by: Fillies, Jan, et al.
Published: (2025)
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
by: Masud, Sarah, et al.
Published: (2023)
by: Masud, Sarah, et al.
Published: (2023)
DefVerify: Do Hate Speech Models Reflect Their Dataset's Definition?
by: Khurana, Urja, et al.
Published: (2024)
by: Khurana, Urja, et al.
Published: (2024)
Hope vs. Hate: Understanding User Interactions with LGBTQ+ News Content in Mainstream US News Media through the Lens of Hope Speech
by: Pofcher, Jonathan, et al.
Published: (2025)
by: Pofcher, Jonathan, et al.
Published: (2025)
Large Language Models are Zero-Shot Next Location Predictors
by: Beneduce, Ciro, et al.
Published: (2024)
by: Beneduce, Ciro, et al.
Published: (2024)
Measuring Online Hate on 4chan using Pre-trained Deep Learning Models
by: Bermudez-Villalva, Adrian, et al.
Published: (2025)
by: Bermudez-Villalva, Adrian, et al.
Published: (2025)
Not My Voice! A Taxonomy of Ethical and Safety Harms of Speech Generators
by: Hutiri, Wiebke, et al.
Published: (2024)
by: Hutiri, Wiebke, et al.
Published: (2024)
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
by: Bhagat, Kirti, et al.
Published: (2025)
by: Bhagat, Kirti, et al.
Published: (2025)
Benchmarking Sociolinguistic Diversity in Swahili NLP: A Taxonomy-Guided Approach
by: Oketch, Kezia, et al.
Published: (2025)
by: Oketch, Kezia, et al.
Published: (2025)
Tackling Social Bias against the Poor: A Dataset and Taxonomy on Aporophobia
by: Curto, Georgina, et al.
Published: (2025)
by: Curto, Georgina, et al.
Published: (2025)
Similar Items
-
Tell Me What You Know About Sexism: Expert-LLM Interaction Strategies and Co-Created Definitions for Zero-Shot Sexism Detection
by: Reuver, Myrthe, et al.
Published: (2025) -
Who and What? Using Linguistic Features and Annotator Characteristics to Analyze Annotation Variation
by: Maurer, Maximilian, et al.
Published: (2026) -
The Unseen Targets of Hate -- A Systematic Review of Hateful Communication Datasets
by: Yu, Zehui, et al.
Published: (2024) -
Towards Weakly-Supervised Hate Speech Classification Across Datasets
by: Jin, Yiping, et al.
Published: (2023) -
From Emotion to Expression: Theoretical Foundations and Resources for Fear Speech
by: Shankaran, Vigneshwaran, et al.
Published: (2026)