A Survey of Machine Learning Models and Datasets for the Multi-label Classification of Textual Hate Speech in English
Fuente:
arXiv
Salvato in:
| Autori principali: | Bäumler, Julian, Blöcher, Louis, Frey, Lars-Joel, Chen, Xian, Bayer, Markus, Reuter, Christian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios
di: Bayer, Markus, et al.
Pubblicazione: (2024)
di: Bayer, Markus, et al.
Pubblicazione: (2024)
BanTH: A Multi-label Hate Speech Detection Dataset for Transliterated Bangla
di: Haider, Fabiha, et al.
Pubblicazione: (2024)
di: Haider, Fabiha, et al.
Pubblicazione: (2024)
Automatic Textual Normalization for Hate Speech Detection
di: Nguyen, Anh Thi-Hoang, et al.
Pubblicazione: (2023)
di: Nguyen, Anh Thi-Hoang, et al.
Pubblicazione: (2023)
Towards Weakly-Supervised Hate Speech Classification Across Datasets
di: Jin, Yiping, et al.
Pubblicazione: (2023)
di: Jin, Yiping, et al.
Pubblicazione: (2023)
Web(er) of Hate: A Survey on How Hate Speech Is Typed
di: Wang, Luna, et al.
Pubblicazione: (2025)
di: Wang, Luna, et al.
Pubblicazione: (2025)
ThreatCrawl: A BERT-based Focused Crawler for the Cybersecurity Domain
di: Kuehn, Philipp, et al.
Pubblicazione: (2023)
di: Kuehn, Philipp, et al.
Pubblicazione: (2023)
Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech Detection
di: Brook, Joshua Wolfe, et al.
Pubblicazione: (2025)
di: Brook, Joshua Wolfe, et al.
Pubblicazione: (2025)
Code-Mixed Telugu-English Hate Speech Detection
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
Improving Hate Speech Classification with Cross-Taxonomy Dataset Integration
di: Fillies, Jan, et al.
Pubblicazione: (2025)
di: Fillies, Jan, et al.
Pubblicazione: (2025)
LLM in the Loop: Creating the ParaDeHate Dataset for Hate Speech Detoxification
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
Bangla Hate Speech Classification with Fine-tuned Transformer Models
di: Jafari, Yalda Keivan, et al.
Pubblicazione: (2025)
di: Jafari, Yalda Keivan, et al.
Pubblicazione: (2025)
BengaliSent140: A Large-Scale Bengali Binary Sentiment Dataset for Hate and Non-Hate Speech Classification
di: Islam, Akif, et al.
Pubblicazione: (2026)
di: Islam, Akif, et al.
Pubblicazione: (2026)
Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
Exploring Cross-Cultural Differences in English Hate Speech Annotations: From Dataset Construction to Analysis
di: Lee, Nayeon, et al.
Pubblicazione: (2023)
di: Lee, Nayeon, et al.
Pubblicazione: (2023)
MetaHate: A Dataset for Unifying Efforts on Hate Speech Detection
di: Piot, Paloma, et al.
Pubblicazione: (2024)
di: Piot, Paloma, et al.
Pubblicazione: (2024)
Hate Speech Detection and Classification in Amharic Text with Deep Learning
di: Gashe, Samuel Minale, et al.
Pubblicazione: (2024)
di: Gashe, Samuel Minale, et al.
Pubblicazione: (2024)
DefVerify: Do Hate Speech Models Reflect Their Dataset's Definition?
di: Khurana, Urja, et al.
Pubblicazione: (2024)
di: Khurana, Urja, et al.
Pubblicazione: (2024)
Empirical Evaluation of Public HateSpeech Datasets
di: Jaf, Sadar, et al.
Pubblicazione: (2024)
di: Jaf, Sadar, et al.
Pubblicazione: (2024)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
di: Yang, Xilin
Pubblicazione: (2024)
di: Yang, Xilin
Pubblicazione: (2024)
Multi3Hate: Multimodal, Multilingual, and Multicultural Hate Speech Detection with Vision-Language Models
di: Bui, Minh Duc, et al.
Pubblicazione: (2024)
di: Bui, Minh Duc, et al.
Pubblicazione: (2024)
AfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languages
di: Muhammad, Shamsuddeen Hassan, et al.
Pubblicazione: (2025)
di: Muhammad, Shamsuddeen Hassan, et al.
Pubblicazione: (2025)
Decoding Hate: Exploring Language Models' Reactions to Hate Speech
di: Piot, Paloma, et al.
Pubblicazione: (2024)
di: Piot, Paloma, et al.
Pubblicazione: (2024)
ProvocationProbe: Instigating Hate Speech Dataset from Twitter
di: Kumar, Abhay, et al.
Pubblicazione: (2024)
di: Kumar, Abhay, et al.
Pubblicazione: (2024)
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
HateDay: Insights from a Global Hate Speech Dataset Representative of a Day on Twitter
di: Tonneau, Manuel, et al.
Pubblicazione: (2024)
di: Tonneau, Manuel, et al.
Pubblicazione: (2024)
Challenger at MultiPRIDE: Is It Hate Speech or Reclaimed?
di: Tekanlou, Hadi Bayrami Asl, et al.
Pubblicazione: (2026)
di: Tekanlou, Hadi Bayrami Asl, et al.
Pubblicazione: (2026)
BIDWESH: A Bangla Regional Based Hate Speech Detection Dataset
di: Fayaz, Azizul Hakim, et al.
Pubblicazione: (2025)
di: Fayaz, Azizul Hakim, et al.
Pubblicazione: (2025)
HateDebias: On the Diversity and Variability of Hate Speech Debiasing
di: Wu, Hongyan, et al.
Pubblicazione: (2024)
di: Wu, Hongyan, et al.
Pubblicazione: (2024)
Disagreeing Rationales: Rethinking Classification and Explainability Evaluation in Hate Speech Detection
di: Muscato, Benedetta, et al.
Pubblicazione: (2026)
di: Muscato, Benedetta, et al.
Pubblicazione: (2026)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
di: Almohaimeed, Saad, et al.
Pubblicazione: (2025)
di: Almohaimeed, Saad, et al.
Pubblicazione: (2025)
Dealing with Annotator Disagreement in Hate Speech Classification
di: Dehghan, Somaiyeh, et al.
Pubblicazione: (2025)
di: Dehghan, Somaiyeh, et al.
Pubblicazione: (2025)
From Languages to Geographies: Towards Evaluating Cultural Bias in Hate Speech Datasets
di: Tonneau, Manuel, et al.
Pubblicazione: (2024)
di: Tonneau, Manuel, et al.
Pubblicazione: (2024)
Cracking the Code: Enhancing Implicit Hate Speech Detection through Coding Classification
di: Wei, Lu, et al.
Pubblicazione: (2025)
di: Wei, Lu, et al.
Pubblicazione: (2025)
Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content
di: Stepanov, Ihor, et al.
Pubblicazione: (2026)
di: Stepanov, Ihor, et al.
Pubblicazione: (2026)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
di: Chan, Fai Leui, et al.
Pubblicazione: (2024)
di: Chan, Fai Leui, et al.
Pubblicazione: (2024)
DEBATE: A Dataset for Disentangling Textual Ambiguity in Mandarin Through Speech
di: Guo, Haotian, et al.
Pubblicazione: (2025)
di: Guo, Haotian, et al.
Pubblicazione: (2025)
Specializing General-purpose LLM Embeddings for Implicit Hate Speech Detection across Datasets
di: Cheremetiev, Vassiliy, et al.
Pubblicazione: (2025)
di: Cheremetiev, Vassiliy, et al.
Pubblicazione: (2025)
MFTCXplain: A Multilingual Benchmark Dataset for Evaluating the Moral Reasoning of LLMs through Multi-hop Hate Speech Explanation
di: Trager, Jackson, et al.
Pubblicazione: (2025)
di: Trager, Jackson, et al.
Pubblicazione: (2025)
HateTinyLLM : Hate Speech Detection Using Tiny Large Language Models
di: Sen, Tanmay, et al.
Pubblicazione: (2024)
di: Sen, Tanmay, et al.
Pubblicazione: (2024)
Multilingual Hate Speech Detection and Counterspeech Generation: A Comprehensive Survey and Practical Guide
di: Fesaghandis, Zahra Safdari, et al.
Pubblicazione: (2026)
di: Fesaghandis, Zahra Safdari, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios
di: Bayer, Markus, et al.
Pubblicazione: (2024) -
BanTH: A Multi-label Hate Speech Detection Dataset for Transliterated Bangla
di: Haider, Fabiha, et al.
Pubblicazione: (2024) -
Automatic Textual Normalization for Hate Speech Detection
di: Nguyen, Anh Thi-Hoang, et al.
Pubblicazione: (2023) -
Towards Weakly-Supervised Hate Speech Classification Across Datasets
di: Jin, Yiping, et al.
Pubblicazione: (2023) -
Web(er) of Hate: A Survey on How Hate Speech Is Typed
di: Wang, Luna, et al.
Pubblicazione: (2025)