An Investigation of Large Language Models for Real-World Hate Speech Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Keyan, Hu, Alexander, Mu, Jaden, Shi, Ziheng, Zhao, Ziming, Vishwamitra, Nishant, Hu, Hongxin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Moderating New Waves of Online Hate with Chain-of-Thought Reasoning in Large Language Models
by: Vishwamitra, Nishant, et al.
Published: (2023)
by: Vishwamitra, Nishant, et al.
Published: (2023)
Moderating Illicit Online Image Promotion for Unsafe User-Generated Content Games Using Large Vision-Language Models
by: Guo, Keyan, et al.
Published: (2024)
by: Guo, Keyan, et al.
Published: (2024)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
by: Jin, Yiping, et al.
Published: (2024)
by: Jin, Yiping, et al.
Published: (2024)
Generalizing Hate Speech Detection Using Multi-Task Learning: A Case Study of Political Public Figures
by: Yuan, Lanqin, et al.
Published: (2022)
by: Yuan, Lanqin, et al.
Published: (2022)
Let Silence Speak: Enhancing Fake News Detection with Generated Comments from Large Language Models
by: Nan, Qiong, et al.
Published: (2024)
by: Nan, Qiong, et al.
Published: (2024)
AI-Cybersecurity Education Through Designing AI-based Cyberharassment Detection Lab
by: Okpala, Ebuka, et al.
Published: (2024)
by: Okpala, Ebuka, et al.
Published: (2024)
Can Prompting LLMs Unlock Hate Speech Detection across Languages? A Zero-shot and Few-shot Study
by: Ghorbanpour, Faeze, et al.
Published: (2025)
by: Ghorbanpour, Faeze, et al.
Published: (2025)
Hate Personified: Investigating the role of LLMs in content moderation
by: Masud, Sarah, et al.
Published: (2024)
by: Masud, Sarah, et al.
Published: (2024)
Few-shot Hate Speech Detection Based on the MindSpore Framework
by: Qin, Zhenkai, et al.
Published: (2025)
by: Qin, Zhenkai, et al.
Published: (2025)
Large Language Models and Thematic Analysis: Human-AI Synergy in Researching Hate Speech on Social Media
by: Breazu, Petre, et al.
Published: (2024)
by: Breazu, Petre, et al.
Published: (2024)
Towards Fairness Assessment of Dutch Hate Speech Detection
by: Bauer, Julie, et al.
Published: (2025)
by: Bauer, Julie, et al.
Published: (2025)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
by: Nandi, Palash, et al.
Published: (2024)
by: Nandi, Palash, et al.
Published: (2024)
Data-Efficient Hate Speech Detection via Cross-Lingual Nearest Neighbor Retrieval with Limited Labeled Data
by: Ghorbanpour, Faeze, et al.
Published: (2025)
by: Ghorbanpour, Faeze, et al.
Published: (2025)
Hate Speech and Sentiment of YouTube Video Comments From Public and Private Sources Covering the Israel-Palestine Conflict
by: Hofmann, Simon, et al.
Published: (2025)
by: Hofmann, Simon, et al.
Published: (2025)
The Enforcement and Feasibility of Hate Speech Moderation on Twitter
by: Tonneau, Manuel, et al.
Published: (2026)
by: Tonneau, Manuel, et al.
Published: (2026)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
by: Gajewska, Ewelina, et al.
Published: (2025)
by: Gajewska, Ewelina, et al.
Published: (2025)
Rethinking Hate Speech Detection on Social Media: Can LLMs Replace Traditional Models?
by: Singh, Daman Deep, et al.
Published: (2025)
by: Singh, Daman Deep, et al.
Published: (2025)
MetaHate: A Dataset for Unifying Efforts on Hate Speech Detection
by: Piot, Paloma, et al.
Published: (2024)
by: Piot, Paloma, et al.
Published: (2024)
Towards Weakly-Supervised Hate Speech Classification Across Datasets
by: Jin, Yiping, et al.
Published: (2023)
by: Jin, Yiping, et al.
Published: (2023)
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
In-Group Love, Out-Group Hate: A Framework to Measure Affective Polarization via Contentious Online Discussions
by: Nettasinghe, Buddhika, et al.
Published: (2024)
by: Nettasinghe, Buddhika, et al.
Published: (2024)
Deep Learning Approaches for Detecting Adversarial Cyberbullying and Hate Speech in Social Networks
by: Azumah, Sylvia Worlali, et al.
Published: (2024)
by: Azumah, Sylvia Worlali, et al.
Published: (2024)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
Beyond Hate: Differentiating Uncivil and Intolerant Speech in Multimodal Content Moderation
by: Herrmann, Nils A., et al.
Published: (2026)
by: Herrmann, Nils A., et al.
Published: (2026)
Signal in the Noise: Decoding the Reality of Airline Service Quality with Large Language Models
by: Dawoud, Ahmed, et al.
Published: (2026)
by: Dawoud, Ahmed, et al.
Published: (2026)
Accuracy and Political Bias of News Source Credibility Ratings by Large Language Models
by: Yang, Kai-Cheng, et al.
Published: (2023)
by: Yang, Kai-Cheng, et al.
Published: (2023)
Evolving Hate Speech Online: An Adaptive Framework for Detection and Mitigation
by: Ali, Shiza, et al.
Published: (2025)
by: Ali, Shiza, et al.
Published: (2025)
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification
by: Faruk, Tanjim Bin
Published: (2024)
by: Faruk, Tanjim Bin
Published: (2024)
Large Language Models Require Curated Context for Reliable Political Fact-Checking -- Even with Reasoning and Web Search
by: DeVerna, Matthew R., et al.
Published: (2025)
by: DeVerna, Matthew R., et al.
Published: (2025)
Real-World Deployment and Evaluation of Kwame for Science, An AI Teaching Assistant for Science Education in West Africa
by: Boateng, George, et al.
Published: (2023)
by: Boateng, George, et al.
Published: (2023)
HateCOT: An Explanation-Enhanced Dataset for Generalizable Offensive Speech Detection via Large Language Models
by: Nghiem, Huy, et al.
Published: (2024)
by: Nghiem, Huy, et al.
Published: (2024)
LLM-Generated Fake News Induces Truth Decay in News Ecosystem: A Case Study on Neural News Recommendation
by: Hu, Beizhe, et al.
Published: (2025)
by: Hu, Beizhe, et al.
Published: (2025)
Can Reinforcement Learning Unlock the Hidden Dangers in Aligned Large Language Models?
by: Karkevandi, Mohammad Bahrami, et al.
Published: (2024)
by: Karkevandi, Mohammad Bahrami, et al.
Published: (2024)
Improving and Assessing the Fidelity of Large Language Models Alignment to Online Communities
by: Chu, Minh Duc, et al.
Published: (2024)
by: Chu, Minh Duc, et al.
Published: (2024)
The Unseen Targets of Hate -- A Systematic Review of Hateful Communication Datasets
by: Yu, Zehui, et al.
Published: (2024)
by: Yu, Zehui, et al.
Published: (2024)
Toxic Synergy Between Hate Speech and Fake News Exposure
by: Kim, Munjung, et al.
Published: (2024)
by: Kim, Munjung, et al.
Published: (2024)
A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance
by: Melis, Matteo, et al.
Published: (2025)
by: Melis, Matteo, et al.
Published: (2025)
Whose Emotions and Moral Sentiments Do Language Models Reflect?
by: He, Zihao, et al.
Published: (2024)
by: He, Zihao, et al.
Published: (2024)
Subjective $\textit{Isms}$? On the Danger of Conflating Hate and Offence in Abusive Language Detection
by: Curry, Amanda Cercas, et al.
Published: (2024)
by: Curry, Amanda Cercas, et al.
Published: (2024)
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
by: Masud, Sarah, et al.
Published: (2023)
by: Masud, Sarah, et al.
Published: (2023)
Similar Items
-
Moderating New Waves of Online Hate with Chain-of-Thought Reasoning in Large Language Models
by: Vishwamitra, Nishant, et al.
Published: (2023) -
Moderating Illicit Online Image Promotion for Unsafe User-Generated Content Games Using Large Vision-Language Models
by: Guo, Keyan, et al.
Published: (2024) -
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
by: Jin, Yiping, et al.
Published: (2024) -
Generalizing Hate Speech Detection Using Multi-Task Learning: A Case Study of Political Public Figures
by: Yuan, Lanqin, et al.
Published: (2022) -
Let Silence Speak: Enhancing Fake News Detection with Generated Comments from Large Language Models
by: Nan, Qiong, et al.
Published: (2024)