An Effective, Robust and Fairness-aware Hate Speech Detection Framework
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mou, Guanyi, Lee, Kyumin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
Reducing and Exploiting Data Augmentation Noise through Meta Reweighting Contrastive Learning for Text Classification
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
Hate Speech Detection with Generalizable Target-aware Fairness
von: Chen, Tong, et al.
Veröffentlicht: (2024)
von: Chen, Tong, et al.
Veröffentlicht: (2024)
Contrastive Learning with Auxiliary User Detection for Identifying Activities
von: Ge, Wen, et al.
Veröffentlicht: (2024)
von: Ge, Wen, et al.
Veröffentlicht: (2024)
Heterogeneous Hyper-Graph Neural Networks for Context-aware Human Activity Recognition
von: Ge, Wen, et al.
Veröffentlicht: (2024)
von: Ge, Wen, et al.
Veröffentlicht: (2024)
Advancing Hate Speech Detection with Transformers: Insights from the MetaHate
von: Chapagain, Santosh, et al.
Veröffentlicht: (2025)
von: Chapagain, Santosh, et al.
Veröffentlicht: (2025)
HateTinyLLM : Hate Speech Detection Using Tiny Large Language Models
von: Sen, Tanmay, et al.
Veröffentlicht: (2024)
von: Sen, Tanmay, et al.
Veröffentlicht: (2024)
Hate Speech Detection and Classification in Amharic Text with Deep Learning
von: Gashe, Samuel Minale, et al.
Veröffentlicht: (2024)
von: Gashe, Samuel Minale, et al.
Veröffentlicht: (2024)
Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement
von: Sheth, Paras, et al.
Veröffentlicht: (2024)
von: Sheth, Paras, et al.
Veröffentlicht: (2024)
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
von: Eilertsen, Brage, et al.
Veröffentlicht: (2025)
von: Eilertsen, Brage, et al.
Veröffentlicht: (2025)
Deep Heterogeneous Contrastive Hyper-Graph Learning for In-the-Wild Context-Aware Human Activity Recognition
von: Ge, Wen, et al.
Veröffentlicht: (2024)
von: Ge, Wen, et al.
Veröffentlicht: (2024)
Semantically Encoding Activity Labels for Context-Aware Human Activity Recognition
von: Ge, Wen, et al.
Veröffentlicht: (2025)
von: Ge, Wen, et al.
Veröffentlicht: (2025)
Wildlife Product Trading in Online Social Networks: A Case Study on Ivory-Related Product Sales Promotion Posts
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
A Survey on Automatic Online Hate Speech Detection in Low-Resource Languages
von: Das, Susmita, et al.
Veröffentlicht: (2024)
von: Das, Susmita, et al.
Veröffentlicht: (2024)
BOISHOMMO: Holistic Approach for Bangla Hate Speech
von: Kafi, Md Abdullah Al, et al.
Veröffentlicht: (2025)
von: Kafi, Md Abdullah Al, et al.
Veröffentlicht: (2025)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
von: Almohaimeed, Saad, et al.
Veröffentlicht: (2025)
von: Almohaimeed, Saad, et al.
Veröffentlicht: (2025)
Deep Learning Approaches for Detecting Adversarial Cyberbullying and Hate Speech in Social Networks
von: Azumah, Sylvia Worlali, et al.
Veröffentlicht: (2024)
von: Azumah, Sylvia Worlali, et al.
Veröffentlicht: (2024)
On Importance of Code-Mixed Embeddings for Hate Speech Identification
von: Jagdale, Shruti, et al.
Veröffentlicht: (2024)
von: Jagdale, Shruti, et al.
Veröffentlicht: (2024)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
von: Das, Amit, et al.
Veröffentlicht: (2024)
von: Das, Amit, et al.
Veröffentlicht: (2024)
Causality Guided Representation Learning for Cross-Style Hate Speech Detection
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2025)
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2025)
Transformers and Ensemble methods: A solution for Hate Speech Detection in Arabic languages
von: de Paula, Angel Felipe Magnossão, et al.
Veröffentlicht: (2023)
von: de Paula, Angel Felipe Magnossão, et al.
Veröffentlicht: (2023)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
von: Nirmal, Ayushi, et al.
Veröffentlicht: (2024)
von: Nirmal, Ayushi, et al.
Veröffentlicht: (2024)
From BERT to Qwen: Hate Detection across architectures
von: Mon, Ariadna, et al.
Veröffentlicht: (2025)
von: Mon, Ariadna, et al.
Veröffentlicht: (2025)
Transfer Learning via Lexical Relatedness: A Sarcasm and Hate Speech Case Study
von: Cabrera, Angelly, et al.
Veröffentlicht: (2025)
von: Cabrera, Angelly, et al.
Veröffentlicht: (2025)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
von: Yang, Xilin
Veröffentlicht: (2024)
von: Yang, Xilin
Veröffentlicht: (2024)
Exploring the Plausibility of Hate and Counter Speech Detectors with Explainable AI
von: Böck, Adrian Jaques, et al.
Veröffentlicht: (2024)
von: Böck, Adrian Jaques, et al.
Veröffentlicht: (2024)
Multilingual Hate Speech Detection in Social Media Using Translation-Based Approaches with Large Language Models
von: Usman, Muhammad, et al.
Veröffentlicht: (2025)
von: Usman, Muhammad, et al.
Veröffentlicht: (2025)
Silencing Empowerment, Allowing Bigotry: Auditing the Moderation of Hate Speech on Twitch
von: Shukla, Prarabdh, et al.
Veröffentlicht: (2025)
von: Shukla, Prarabdh, et al.
Veröffentlicht: (2025)
An Investigation of Large Language Models for Real-World Hate Speech Detection
von: Guo, Keyan, et al.
Veröffentlicht: (2024)
von: Guo, Keyan, et al.
Veröffentlicht: (2024)
Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models
von: AlDahoul, Nouar, et al.
Veröffentlicht: (2025)
von: AlDahoul, Nouar, et al.
Veröffentlicht: (2025)
Enhancing Effectiveness and Robustness in a Low-Resource Regime via Decision-Boundary-aware Data Augmentation
von: Jin, Kyohoon, et al.
Veröffentlicht: (2024)
von: Jin, Kyohoon, et al.
Veröffentlicht: (2024)
1-800-SHARED-TASKS @ NLU of Devanagari Script Languages: Detection of Language, Hate Speech, and Targets using LLMs
von: Purbey, Jebish, et al.
Veröffentlicht: (2024)
von: Purbey, Jebish, et al.
Veröffentlicht: (2024)
Parameter-Efficient Fine-Tuning for Low-Resource Languages: A Comparative Study of LLMs for Bengali Hate Speech Detection
von: Islam, Akif, et al.
Veröffentlicht: (2025)
von: Islam, Akif, et al.
Veröffentlicht: (2025)
Comparison of Modern Multilingual Text Embedding Techniques for Hate Speech Detection Task
von: Vaiciukynas, Evaldas, et al.
Veröffentlicht: (2026)
von: Vaiciukynas, Evaldas, et al.
Veröffentlicht: (2026)
NoisyHate: Mining Online Human-Written Perturbations for Realistic Robustness Benchmarking of Content Moderation Models
von: Ye, Yiran, et al.
Veröffentlicht: (2023)
von: Ye, Yiran, et al.
Veröffentlicht: (2023)
OSPC: Artificial VLM Features for Hateful Meme Detection
von: Grönquist, Peter
Veröffentlicht: (2024)
von: Grönquist, Peter
Veröffentlicht: (2024)
Robust Adaptation of Large Multimodal Models for Retrieval Augmented Hateful Meme Detection
von: Mei, Jingbiao, et al.
Veröffentlicht: (2025)
von: Mei, Jingbiao, et al.
Veröffentlicht: (2025)
Attentive Fusion: A Transformer-based Approach to Multimodal Hate Speech Detection
von: Mandal, Atanu, et al.
Veröffentlicht: (2024)
von: Mandal, Atanu, et al.
Veröffentlicht: (2024)
Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content
von: Stepanov, Ihor, et al.
Veröffentlicht: (2026)
von: Stepanov, Ihor, et al.
Veröffentlicht: (2026)
Multi-Modal Discussion Transformer: Integrating Text, Images and Graph Transformers to Detect Hate Speech on Social Media
von: Hebert, Liam, et al.
Veröffentlicht: (2023)
von: Hebert, Liam, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection
von: Mou, Guanyi, et al.
Veröffentlicht: (2024) -
Reducing and Exploiting Data Augmentation Noise through Meta Reweighting Contrastive Learning for Text Classification
von: Mou, Guanyi, et al.
Veröffentlicht: (2024) -
Hate Speech Detection with Generalizable Target-aware Fairness
von: Chen, Tong, et al.
Veröffentlicht: (2024) -
Contrastive Learning with Auxiliary User Detection for Identifying Activities
von: Ge, Wen, et al.
Veröffentlicht: (2024) -
Heterogeneous Hyper-Graph Neural Networks for Context-aware Human Activity Recognition
von: Ge, Wen, et al.
Veröffentlicht: (2024)