Detecting Anti-Semitic Hate Speech using Transformer-based Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Dengyi, Wang, Minghao, Catlin, Andrew G. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
HateTinyLLM : Hate Speech Detection Using Tiny Large Language Models
di: Sen, Tanmay, et al.
Pubblicazione: (2024)
di: Sen, Tanmay, et al.
Pubblicazione: (2024)
Exploring Large Language Models for Hate Speech Detection in Rioplatense Spanish
di: Pérez, Juan Manuel, et al.
Pubblicazione: (2024)
di: Pérez, Juan Manuel, et al.
Pubblicazione: (2024)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
di: Nirmal, Ayushi, et al.
Pubblicazione: (2024)
di: Nirmal, Ayushi, et al.
Pubblicazione: (2024)
Hate Speech Detection using Large Language Models with Data Augmentation and Feature Enhancement
di: Nge, Brian Jing Hong, et al.
Pubblicazione: (2026)
di: Nge, Brian Jing Hong, et al.
Pubblicazione: (2026)
cantnlp@LT-EDI-2024: Automatic Detection of Anti-LGBTQ+ Hate Speech in Under-resourced Languages
di: Wong, Sidney G. -J., et al.
Pubblicazione: (2024)
di: Wong, Sidney G. -J., et al.
Pubblicazione: (2024)
MasonPerplexity at Multimodal Hate Speech Event Detection 2024: Hate Speech and Target Detection Using Transformer Ensembles
di: Ganguly, Amrita, et al.
Pubblicazione: (2024)
di: Ganguly, Amrita, et al.
Pubblicazione: (2024)
Advancing Hate Speech Detection with Transformers: Insights from the MetaHate
di: Chapagain, Santosh, et al.
Pubblicazione: (2025)
di: Chapagain, Santosh, et al.
Pubblicazione: (2025)
Multi3Hate: Multimodal, Multilingual, and Multicultural Hate Speech Detection with Vision-Language Models
di: Bui, Minh Duc, et al.
Pubblicazione: (2024)
di: Bui, Minh Duc, et al.
Pubblicazione: (2024)
Personalisation or Prejudice? Addressing Geographic Bias in Hate Speech Detection using Debias Tuning in Large Language Models
di: Piot, Paloma, et al.
Pubblicazione: (2025)
di: Piot, Paloma, et al.
Pubblicazione: (2025)
Decoding Hate: Exploring Language Models' Reactions to Hate Speech
di: Piot, Paloma, et al.
Pubblicazione: (2024)
di: Piot, Paloma, et al.
Pubblicazione: (2024)
Outcome-Constrained Large Language Models for Countering Hate Speech
di: Hong, Lingzi, et al.
Pubblicazione: (2024)
di: Hong, Lingzi, et al.
Pubblicazione: (2024)
Conditioning Large Language Models on Legal Systems? Detecting Punishable Hate Speech
di: Ludwig, Florian, et al.
Pubblicazione: (2025)
di: Ludwig, Florian, et al.
Pubblicazione: (2025)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
di: Das, Amit, et al.
Pubblicazione: (2024)
di: Das, Amit, et al.
Pubblicazione: (2024)
ViHateT5: Enhancing Hate Speech Detection in Vietnamese With A Unified Text-to-Text Transformer Model
di: Nguyen, Luan Thanh
Pubblicazione: (2024)
di: Nguyen, Luan Thanh
Pubblicazione: (2024)
Web(er) of Hate: A Survey on How Hate Speech Is Typed
di: Wang, Luna, et al.
Pubblicazione: (2025)
di: Wang, Luna, et al.
Pubblicazione: (2025)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
An Investigation of Large Language Models for Real-World Hate Speech Detection
di: Guo, Keyan, et al.
Pubblicazione: (2024)
di: Guo, Keyan, et al.
Pubblicazione: (2024)
Self-Explaining Hate Speech Detection with Moral Rationales
di: Vargas, Francielle, et al.
Pubblicazione: (2026)
di: Vargas, Francielle, et al.
Pubblicazione: (2026)
NLPineers@ NLU of Devanagari Script Languages 2025: Hate Speech Detection using Ensembling of BERT-based models
di: Guragain, Anmol, et al.
Pubblicazione: (2024)
di: Guragain, Anmol, et al.
Pubblicazione: (2024)
Bangla Hate Speech Classification with Fine-tuned Transformer Models
di: Jafari, Yalda Keivan, et al.
Pubblicazione: (2025)
di: Jafari, Yalda Keivan, et al.
Pubblicazione: (2025)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
di: Chan, Fai Leui, et al.
Pubblicazione: (2024)
di: Chan, Fai Leui, et al.
Pubblicazione: (2024)
EkoHate: Abusive Language and Hate Speech Detection for Code-switched Political Discussions on Nigerian Twitter
di: Ilevbare, Comfort Eseohen, et al.
Pubblicazione: (2024)
di: Ilevbare, Comfort Eseohen, et al.
Pubblicazione: (2024)
Analyzing Bias in False Refusal Behavior of Large Language Models for Hate Speech Detoxification
di: Im, Kyuri, et al.
Pubblicazione: (2026)
di: Im, Kyuri, et al.
Pubblicazione: (2026)
Evaluating Simple Debiasing Techniques in RoBERTa-based Hate Speech Detection Models
di: Iftimie, Diana, et al.
Pubblicazione: (2025)
di: Iftimie, Diana, et al.
Pubblicazione: (2025)
Compositional Generalisation for Explainable Hate Speech Detection
di: Calabrese, Agostina, et al.
Pubblicazione: (2025)
di: Calabrese, Agostina, et al.
Pubblicazione: (2025)
Automatic Textual Normalization for Hate Speech Detection
di: Nguyen, Anh Thi-Hoang, et al.
Pubblicazione: (2023)
di: Nguyen, Anh Thi-Hoang, et al.
Pubblicazione: (2023)
HateDebias: On the Diversity and Variability of Hate Speech Debiasing
di: Wu, Hongyan, et al.
Pubblicazione: (2024)
di: Wu, Hongyan, et al.
Pubblicazione: (2024)
AfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languages
di: Muhammad, Shamsuddeen Hassan, et al.
Pubblicazione: (2025)
di: Muhammad, Shamsuddeen Hassan, et al.
Pubblicazione: (2025)
NaijaHate: Evaluating Hate Speech Detection on Nigerian Twitter Using Representative Data
di: Tonneau, Manuel, et al.
Pubblicazione: (2024)
di: Tonneau, Manuel, et al.
Pubblicazione: (2024)
When Hate Meets Facts: LLMs-in-the-Loop for Check-worthiness Detection in Hate Speech
di: Ocampo, Nicolás Benjamín, et al.
Pubblicazione: (2026)
di: Ocampo, Nicolás Benjamín, et al.
Pubblicazione: (2026)
Towards Efficient and Explainable Hate Speech Detection via Model Distillation
di: Piot, Paloma, et al.
Pubblicazione: (2024)
di: Piot, Paloma, et al.
Pubblicazione: (2024)
Recent Advances in Hate Speech Moderation: Multimodality and the Role of Large Models
di: Hee, Ming Shan, et al.
Pubblicazione: (2024)
di: Hee, Ming Shan, et al.
Pubblicazione: (2024)
Multilingual Hate Speech Detection in Social Media Using Translation-Based Approaches with Large Language Models
di: Usman, Muhammad, et al.
Pubblicazione: (2025)
di: Usman, Muhammad, et al.
Pubblicazione: (2025)
Hate Speech According to the Law: An Analysis for Effective Detection
di: Korre, Katerina, et al.
Pubblicazione: (2024)
di: Korre, Katerina, et al.
Pubblicazione: (2024)
Incorporating Human Explanations for Robust Hate Speech Detection
di: Chen, Jennifer L., et al.
Pubblicazione: (2024)
di: Chen, Jennifer L., et al.
Pubblicazione: (2024)
Code-Mixed Telugu-English Hate Speech Detection
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
MetaHate: A Dataset for Unifying Efforts on Hate Speech Detection
di: Piot, Paloma, et al.
Pubblicazione: (2024)
di: Piot, Paloma, et al.
Pubblicazione: (2024)
cantnlp@DravidianLangTech2025: A Bag-of-Sounds Approach to Multimodal Hate Speech Detection
di: Wong, Sidney, et al.
Pubblicazione: (2025)
di: Wong, Sidney, et al.
Pubblicazione: (2025)
Few-shot Hate Speech Detection Based on the MindSpore Framework
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025) -
HateTinyLLM : Hate Speech Detection Using Tiny Large Language Models
di: Sen, Tanmay, et al.
Pubblicazione: (2024) -
Exploring Large Language Models for Hate Speech Detection in Rioplatense Spanish
di: Pérez, Juan Manuel, et al.
Pubblicazione: (2024) -
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
di: Nirmal, Ayushi, et al.
Pubblicazione: (2024) -
Hate Speech Detection using Large Language Models with Data Augmentation and Feature Enhancement
di: Nge, Brian Jing Hong, et al.
Pubblicazione: (2026)