The Unseen Targets of Hate -- A Systematic Review of Hateful Communication Datasets
Fuente:
arXiv
Salvato in:
| Autori principali: | Yu, Zehui, Sen, Indira, Assenmacher, Dennis, Samory, Mattia, Fröhling, Leon, Dahn, Christina, Nozza, Debora, Wagner, Claudia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
You are a Bot! -- Studying the Development of Bot Accusations on Twitter
di: Assenmacher, Dennis, et al.
Pubblicazione: (2023)
di: Assenmacher, Dennis, et al.
Pubblicazione: (2023)
People Make Better Edits: Measuring the Efficacy of LLM-Generated Counterfactually Augmented Data for Harmful Language Detection
di: Sen, Indira, et al.
Pubblicazione: (2023)
di: Sen, Indira, et al.
Pubblicazione: (2023)
A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance
di: Melis, Matteo, et al.
Pubblicazione: (2025)
di: Melis, Matteo, et al.
Pubblicazione: (2025)
Robustness and Confounders in the Demographic Alignment of LLMs with Human Perceptions of Offensiveness
di: Alipour, Shayan, et al.
Pubblicazione: (2024)
di: Alipour, Shayan, et al.
Pubblicazione: (2024)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
di: Fröhling, Leon, et al.
Pubblicazione: (2024)
di: Fröhling, Leon, et al.
Pubblicazione: (2024)
German General Social Survey Personas: A Survey-Derived Persona Prompt Collection for Population-Aligned LLM Studies
di: Rupprecht, Jens, et al.
Pubblicazione: (2025)
di: Rupprecht, Jens, et al.
Pubblicazione: (2025)
Assessing How Hate, Counterspeech, and Toxicity Affect Hate Group Newcomers
di: Hickey, Daniel, et al.
Pubblicazione: (2024)
di: Hickey, Daniel, et al.
Pubblicazione: (2024)
Learning from Convenience Samples: A Case Study on Fine-Tuning LLMs for Survey Non-response in the German Longitudinal Election Study
di: Holtdirk, Tobias, et al.
Pubblicazione: (2025)
di: Holtdirk, Tobias, et al.
Pubblicazione: (2025)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
di: Jin, Yiping, et al.
Pubblicazione: (2024)
di: Jin, Yiping, et al.
Pubblicazione: (2024)
Towards Weakly-Supervised Hate Speech Classification Across Datasets
di: Jin, Yiping, et al.
Pubblicazione: (2023)
di: Jin, Yiping, et al.
Pubblicazione: (2023)
A Multilingual Similarity Dataset for News Article Frame
di: Chen, Xi, et al.
Pubblicazione: (2024)
di: Chen, Xi, et al.
Pubblicazione: (2024)
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
di: Lee, Yejin, et al.
Pubblicazione: (2025)
di: Lee, Yejin, et al.
Pubblicazione: (2025)
The Enforcement and Feasibility of Hate Speech Moderation on Twitter
di: Tonneau, Manuel, et al.
Pubblicazione: (2026)
di: Tonneau, Manuel, et al.
Pubblicazione: (2026)
Hate Personified: Investigating the role of LLMs in content moderation
di: Masud, Sarah, et al.
Pubblicazione: (2024)
di: Masud, Sarah, et al.
Pubblicazione: (2024)
Missing the Margins: A Systematic Literature Review on the Demographic Representativeness of LLMs
di: Sen, Indira, et al.
Pubblicazione: (2025)
di: Sen, Indira, et al.
Pubblicazione: (2025)
What Is The Political Content in LLMs' Pre- and Post-Training Data?
di: Ceron, Tanise, et al.
Pubblicazione: (2025)
di: Ceron, Tanise, et al.
Pubblicazione: (2025)
Towards Fairness Assessment of Dutch Hate Speech Detection
di: Bauer, Julie, et al.
Pubblicazione: (2025)
di: Bauer, Julie, et al.
Pubblicazione: (2025)
Multilingualism, Transnationality, and K-pop in the Online #StopAsianHate Movement
di: Masis, Tessa, et al.
Pubblicazione: (2025)
di: Masis, Tessa, et al.
Pubblicazione: (2025)
Few-shot Hate Speech Detection Based on the MindSpore Framework
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
Beyond Hate: Differentiating Uncivil and Intolerant Speech in Multimodal Content Moderation
di: Herrmann, Nils A., et al.
Pubblicazione: (2026)
di: Herrmann, Nils A., et al.
Pubblicazione: (2026)
Deciphering Hate: Identifying Hateful Memes and Their Targets
di: Hossain, Eftekhar, et al.
Pubblicazione: (2024)
di: Hossain, Eftekhar, et al.
Pubblicazione: (2024)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
di: Nandi, Palash, et al.
Pubblicazione: (2024)
di: Nandi, Palash, et al.
Pubblicazione: (2024)
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
di: Masud, Sarah, et al.
Pubblicazione: (2023)
di: Masud, Sarah, et al.
Pubblicazione: (2023)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
di: Samory, Mattia, et al.
Pubblicazione: (2025)
di: Samory, Mattia, et al.
Pubblicazione: (2025)
The Peripatetic Hater: Predicting Movement Among Hate Subreddits
di: Hickey, Daniel, et al.
Pubblicazione: (2024)
di: Hickey, Daniel, et al.
Pubblicazione: (2024)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
Measuring Online Hate on 4chan using Pre-trained Deep Learning Models
di: Bermudez-Villalva, Adrian, et al.
Pubblicazione: (2025)
di: Bermudez-Villalva, Adrian, et al.
Pubblicazione: (2025)
Rethinking Hate Speech Detection on Social Media: Can LLMs Replace Traditional Models?
di: Singh, Daman Deep, et al.
Pubblicazione: (2025)
di: Singh, Daman Deep, et al.
Pubblicazione: (2025)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
di: Yang, Xilin
Pubblicazione: (2024)
di: Yang, Xilin
Pubblicazione: (2024)
Subjective $\textit{Isms}$? On the Danger of Conflating Hate and Offence in Abusive Language Detection
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
di: Curry, Amanda Cercas, et al.
Pubblicazione: (2024)
Deep Learning Approaches for Detecting Adversarial Cyberbullying and Hate Speech in Social Networks
di: Azumah, Sylvia Worlali, et al.
Pubblicazione: (2024)
di: Azumah, Sylvia Worlali, et al.
Pubblicazione: (2024)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
di: Ahmed, Ahmed Haj, et al.
Pubblicazione: (2024)
di: Ahmed, Ahmed Haj, et al.
Pubblicazione: (2024)
Toxic Synergy Between Hate Speech and Fake News Exposure
di: Kim, Munjung, et al.
Pubblicazione: (2024)
di: Kim, Munjung, et al.
Pubblicazione: (2024)
HateTinyLLM : Hate Speech Detection Using Tiny Large Language Models
di: Sen, Tanmay, et al.
Pubblicazione: (2024)
di: Sen, Tanmay, et al.
Pubblicazione: (2024)
A Civics-oriented Approach to Understanding Intersectionally Marginalized Users' Experience with Hate Speech Online
di: Sultana, Achhiya, et al.
Pubblicazione: (2024)
di: Sultana, Achhiya, et al.
Pubblicazione: (2024)
No Easy Way Out: the Effectiveness of Deplatforming an Extremist Forum to Suppress Hate and Harassment
di: Vu, Anh V., et al.
Pubblicazione: (2023)
di: Vu, Anh V., et al.
Pubblicazione: (2023)
From Measurement Instruments to Data: Leveraging Theory-Driven Synthetic Training Data for Classifying Social Constructs
di: Birkenmaier, Lukas, et al.
Pubblicazione: (2024)
di: Birkenmaier, Lukas, et al.
Pubblicazione: (2024)
Data-Efficient Hate Speech Detection via Cross-Lingual Nearest Neighbor Retrieval with Limited Labeled Data
di: Ghorbanpour, Faeze, et al.
Pubblicazione: (2025)
di: Ghorbanpour, Faeze, et al.
Pubblicazione: (2025)
Analyzing User Characteristics of Hate Speech Spreaders on Social Media
di: Geissler, Dominique, et al.
Pubblicazione: (2023)
di: Geissler, Dominique, et al.
Pubblicazione: (2023)
Labels or Input? Rethinking Augmentation in Multimodal Hate Detection
di: Singh, Sahajpreet, et al.
Pubblicazione: (2025)
di: Singh, Sahajpreet, et al.
Pubblicazione: (2025)
Documenti analoghi
-
You are a Bot! -- Studying the Development of Bot Accusations on Twitter
di: Assenmacher, Dennis, et al.
Pubblicazione: (2023) -
People Make Better Edits: Measuring the Efficacy of LLM-Generated Counterfactually Augmented Data for Harmful Language Detection
di: Sen, Indira, et al.
Pubblicazione: (2023) -
A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance
di: Melis, Matteo, et al.
Pubblicazione: (2025) -
Robustness and Confounders in the Demographic Alignment of LLMs with Human Perceptions of Offensiveness
di: Alipour, Shayan, et al.
Pubblicazione: (2024) -
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
di: Fröhling, Leon, et al.
Pubblicazione: (2024)