An Effective, Robust and Fairness-aware Hate Speech Detection Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Mou, Guanyi, Lee, Kyumin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection
by: Mou, Guanyi, et al.
Published: (2024)
by: Mou, Guanyi, et al.
Published: (2024)
Reducing and Exploiting Data Augmentation Noise through Meta Reweighting Contrastive Learning for Text Classification
by: Mou, Guanyi, et al.
Published: (2024)
by: Mou, Guanyi, et al.
Published: (2024)
Hate Speech Detection with Generalizable Target-aware Fairness
by: Chen, Tong, et al.
Published: (2024)
by: Chen, Tong, et al.
Published: (2024)
Contrastive Learning with Auxiliary User Detection for Identifying Activities
by: Ge, Wen, et al.
Published: (2024)
by: Ge, Wen, et al.
Published: (2024)
Heterogeneous Hyper-Graph Neural Networks for Context-aware Human Activity Recognition
by: Ge, Wen, et al.
Published: (2024)
by: Ge, Wen, et al.
Published: (2024)
Advancing Hate Speech Detection with Transformers: Insights from the MetaHate
by: Chapagain, Santosh, et al.
Published: (2025)
by: Chapagain, Santosh, et al.
Published: (2025)
HateTinyLLM : Hate Speech Detection Using Tiny Large Language Models
by: Sen, Tanmay, et al.
Published: (2024)
by: Sen, Tanmay, et al.
Published: (2024)
Hate Speech Detection and Classification in Amharic Text with Deep Learning
by: Gashe, Samuel Minale, et al.
Published: (2024)
by: Gashe, Samuel Minale, et al.
Published: (2024)
Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement
by: Sheth, Paras, et al.
Published: (2024)
by: Sheth, Paras, et al.
Published: (2024)
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
by: Eilertsen, Brage, et al.
Published: (2025)
by: Eilertsen, Brage, et al.
Published: (2025)
Deep Heterogeneous Contrastive Hyper-Graph Learning for In-the-Wild Context-Aware Human Activity Recognition
by: Ge, Wen, et al.
Published: (2024)
by: Ge, Wen, et al.
Published: (2024)
Semantically Encoding Activity Labels for Context-Aware Human Activity Recognition
by: Ge, Wen, et al.
Published: (2025)
by: Ge, Wen, et al.
Published: (2025)
Wildlife Product Trading in Online Social Networks: A Case Study on Ivory-Related Product Sales Promotion Posts
by: Mou, Guanyi, et al.
Published: (2024)
by: Mou, Guanyi, et al.
Published: (2024)
A Survey on Automatic Online Hate Speech Detection in Low-Resource Languages
by: Das, Susmita, et al.
Published: (2024)
by: Das, Susmita, et al.
Published: (2024)
BOISHOMMO: Holistic Approach for Bangla Hate Speech
by: Kafi, Md Abdullah Al, et al.
Published: (2025)
by: Kafi, Md Abdullah Al, et al.
Published: (2025)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
by: Almohaimeed, Saad, et al.
Published: (2025)
by: Almohaimeed, Saad, et al.
Published: (2025)
Deep Learning Approaches for Detecting Adversarial Cyberbullying and Hate Speech in Social Networks
by: Azumah, Sylvia Worlali, et al.
Published: (2024)
by: Azumah, Sylvia Worlali, et al.
Published: (2024)
On Importance of Code-Mixed Embeddings for Hate Speech Identification
by: Jagdale, Shruti, et al.
Published: (2024)
by: Jagdale, Shruti, et al.
Published: (2024)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
by: Das, Amit, et al.
Published: (2024)
by: Das, Amit, et al.
Published: (2024)
Causality Guided Representation Learning for Cross-Style Hate Speech Detection
by: Zhao, Chengshuai, et al.
Published: (2025)
by: Zhao, Chengshuai, et al.
Published: (2025)
Transformers and Ensemble methods: A solution for Hate Speech Detection in Arabic languages
by: de Paula, Angel Felipe Magnossão, et al.
Published: (2023)
by: de Paula, Angel Felipe Magnossão, et al.
Published: (2023)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
by: Nirmal, Ayushi, et al.
Published: (2024)
by: Nirmal, Ayushi, et al.
Published: (2024)
From BERT to Qwen: Hate Detection across architectures
by: Mon, Ariadna, et al.
Published: (2025)
by: Mon, Ariadna, et al.
Published: (2025)
Transfer Learning via Lexical Relatedness: A Sarcasm and Hate Speech Case Study
by: Cabrera, Angelly, et al.
Published: (2025)
by: Cabrera, Angelly, et al.
Published: (2025)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
by: Yang, Xilin
Published: (2024)
by: Yang, Xilin
Published: (2024)
Exploring the Plausibility of Hate and Counter Speech Detectors with Explainable AI
by: Böck, Adrian Jaques, et al.
Published: (2024)
by: Böck, Adrian Jaques, et al.
Published: (2024)
Multilingual Hate Speech Detection in Social Media Using Translation-Based Approaches with Large Language Models
by: Usman, Muhammad, et al.
Published: (2025)
by: Usman, Muhammad, et al.
Published: (2025)
Silencing Empowerment, Allowing Bigotry: Auditing the Moderation of Hate Speech on Twitch
by: Shukla, Prarabdh, et al.
Published: (2025)
by: Shukla, Prarabdh, et al.
Published: (2025)
An Investigation of Large Language Models for Real-World Hate Speech Detection
by: Guo, Keyan, et al.
Published: (2024)
by: Guo, Keyan, et al.
Published: (2024)
Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models
by: AlDahoul, Nouar, et al.
Published: (2025)
by: AlDahoul, Nouar, et al.
Published: (2025)
Enhancing Effectiveness and Robustness in a Low-Resource Regime via Decision-Boundary-aware Data Augmentation
by: Jin, Kyohoon, et al.
Published: (2024)
by: Jin, Kyohoon, et al.
Published: (2024)
1-800-SHARED-TASKS @ NLU of Devanagari Script Languages: Detection of Language, Hate Speech, and Targets using LLMs
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
Parameter-Efficient Fine-Tuning for Low-Resource Languages: A Comparative Study of LLMs for Bengali Hate Speech Detection
by: Islam, Akif, et al.
Published: (2025)
by: Islam, Akif, et al.
Published: (2025)
Comparison of Modern Multilingual Text Embedding Techniques for Hate Speech Detection Task
by: Vaiciukynas, Evaldas, et al.
Published: (2026)
by: Vaiciukynas, Evaldas, et al.
Published: (2026)
NoisyHate: Mining Online Human-Written Perturbations for Realistic Robustness Benchmarking of Content Moderation Models
by: Ye, Yiran, et al.
Published: (2023)
by: Ye, Yiran, et al.
Published: (2023)
OSPC: Artificial VLM Features for Hateful Meme Detection
by: Grönquist, Peter
Published: (2024)
by: Grönquist, Peter
Published: (2024)
Robust Adaptation of Large Multimodal Models for Retrieval Augmented Hateful Meme Detection
by: Mei, Jingbiao, et al.
Published: (2025)
by: Mei, Jingbiao, et al.
Published: (2025)
Attentive Fusion: A Transformer-based Approach to Multimodal Hate Speech Detection
by: Mandal, Atanu, et al.
Published: (2024)
by: Mandal, Atanu, et al.
Published: (2024)
Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content
by: Stepanov, Ihor, et al.
Published: (2026)
by: Stepanov, Ihor, et al.
Published: (2026)
Multi-Modal Discussion Transformer: Integrating Text, Images and Graph Transformers to Detect Hate Speech on Social Media
by: Hebert, Liam, et al.
Published: (2023)
by: Hebert, Liam, et al.
Published: (2023)
Similar Items
-
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection
by: Mou, Guanyi, et al.
Published: (2024) -
Reducing and Exploiting Data Augmentation Noise through Meta Reweighting Contrastive Learning for Text Classification
by: Mou, Guanyi, et al.
Published: (2024) -
Hate Speech Detection with Generalizable Target-aware Fairness
by: Chen, Tong, et al.
Published: (2024) -
Contrastive Learning with Auxiliary User Detection for Identifying Activities
by: Ge, Wen, et al.
Published: (2024) -
Heterogeneous Hyper-Graph Neural Networks for Context-aware Human Activity Recognition
by: Ge, Wen, et al.
Published: (2024)