Can NLP Tackle Hate Speech in the Real World? Stakeholder-Informed Feedback and Survey on Counterspeech
Fuente:
arXiv
Saved in:
| Main Authors: | Dinkar, Tanvi, Jiang, Aiqi, Frenda, Simona, Gerrard-Abbott, Poppy, Gunson, Nancie, Abercrombie, Gavin, Konstas, Ioannis |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Erasing 'Ugly' from the Internet: Propagation of the Beauty Myth in Text-Image Models
by: Dinkar, Tanvi, et al.
Published: (2025)
by: Dinkar, Tanvi, et al.
Published: (2025)
Re-examining Sexism and Misogyny Classification with Annotator Attitudes
by: Jiang, Aiqi, et al.
Published: (2024)
by: Jiang, Aiqi, et al.
Published: (2024)
NLP for Counterspeech against Hate: A Survey and How-To Guide
by: Bonaldi, Helena, et al.
Published: (2024)
by: Bonaldi, Helena, et al.
Published: (2024)
Food Noise & False Safety: A Systematic Evaluation of How LLMs Fail to Adapt to Eating Disorder Queries with Clinician Feedback
by: Pucci, Giulia, et al.
Published: (2026)
by: Pucci, Giulia, et al.
Published: (2026)
Consistency is Key: Disentangling Label Variation in Natural Language Processing with Intra-Annotator Agreement
by: Abercrombie, Gavin, et al.
Published: (2023)
by: Abercrombie, Gavin, et al.
Published: (2023)
Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation
by: Martone, Genoveffa, et al.
Published: (2026)
by: Martone, Genoveffa, et al.
Published: (2026)
Multilingual Hate Speech Detection and Counterspeech Generation: A Comprehensive Survey and Practical Guide
by: Fesaghandis, Zahra Safdari, et al.
Published: (2026)
by: Fesaghandis, Zahra Safdari, et al.
Published: (2026)
Subjective $\textit{Isms}$? On the Danger of Conflating Hate and Offence in Abusive Language Detection
by: Curry, Amanda Cercas, et al.
Published: (2024)
by: Curry, Amanda Cercas, et al.
Published: (2024)
Assessing How Hate, Counterspeech, and Toxicity Affect Hate Group Newcomers
by: Hickey, Daniel, et al.
Published: (2024)
by: Hickey, Daniel, et al.
Published: (2024)
Will I Get Hate Speech Predicting the Volume of Abusive Replies before Posting in Social Media
by: Alharthi, Raneem, et al.
Published: (2025)
by: Alharthi, Raneem, et al.
Published: (2025)
NLP Systems That Can't Tell Use from Mention Censor Counterspeech, but Teaching the Distinction Helps
by: Gligoric, Kristina, et al.
Published: (2024)
by: Gligoric, Kristina, et al.
Published: (2024)
HatePRISM: Policies, Platforms, and Research Integration. Advancing NLP for Hate Speech Proactive Mitigation
by: Rizwan, Naquee, et al.
Published: (2025)
by: Rizwan, Naquee, et al.
Published: (2025)
Reasoning or a Semblance of it? A Diagnostic Study of Transitive Reasoning in LLMs
by: Mehrafarin, Houman, et al.
Published: (2024)
by: Mehrafarin, Houman, et al.
Published: (2024)
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
by: Mehrafarin, Houman, et al.
Published: (2026)
by: Mehrafarin, Houman, et al.
Published: (2026)
Voices in a Crowd: Searching for Clusters of Unique Perspectives
by: Vitsakis, Nikolas, et al.
Published: (2024)
by: Vitsakis, Nikolas, et al.
Published: (2024)
Beyond Hate Speech: NLP's Challenges and Opportunities in Uncovering Dehumanizing Language
by: Saffari, Hamidreza, et al.
Published: (2024)
by: Saffari, Hamidreza, et al.
Published: (2024)
The efficacy of facial skeletal treatment options in the management of obstructive sleep apnea
by: Michael J. Gunson
Published: (2025)
by: Michael J. Gunson
Published: (2025)
NLP Verification: Towards a General Methodology for Certifying Robustness
by: Casadio, Marco, et al.
Published: (2024)
by: Casadio, Marco, et al.
Published: (2024)
Are you sure? Measuring models bias in content moderation through uncertainty
by: Urbinati, Alessandra, et al.
Published: (2025)
by: Urbinati, Alessandra, et al.
Published: (2025)
Counterspeech
Published: (2025)
Published: (2025)
Explainable AI for Hate Speech Moderation: A Stakeholder‐Centered and Sociotechnical Review
by: Muhammad Deedahwar Mazhar Qureshi, et al.
Published: (2026)
by: Muhammad Deedahwar Mazhar Qureshi, et al.
Published: (2026)
Web(er) of Hate: A Survey on How Hate Speech Is Typed
by: Wang, Luna, et al.
Published: (2025)
by: Wang, Luna, et al.
Published: (2025)
MoRFI: Monotonic Sparse Autoencoder Feature Identification
by: Dimakopoulos, Dimitris, et al.
Published: (2026)
by: Dimakopoulos, Dimitris, et al.
Published: (2026)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
by: Gajewska, Ewelina, et al.
Published: (2025)
by: Gajewska, Ewelina, et al.
Published: (2025)
The Commercial Yellowtail Snapper fishery off Puerto Rico, 1983-2003
by: Cummings, Nancie J.
Published: (2007)
by: Cummings, Nancie J.
Published: (2007)
Investigating the Role of Instruction Variety and Task Difficulty in Robotic Manipulation Tasks
by: Parekh, Amit, et al.
Published: (2024)
by: Parekh, Amit, et al.
Published: (2024)
Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation
by: Nikandrou, Malvina, et al.
Published: (2024)
by: Nikandrou, Malvina, et al.
Published: (2024)
Retrievit: In-context Retrieval Capabilities of Transformers, State Space Models, and Hybrid Architectures
by: Pantazopoulos, Georgios, et al.
Published: (2026)
by: Pantazopoulos, Georgios, et al.
Published: (2026)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
by: Jin, Yiping, et al.
Published: (2024)
by: Jin, Yiping, et al.
Published: (2024)
Hate Speech
by: Guillén-Nieto, Victoria
Published: (2023)
by: Guillén-Nieto, Victoria
Published: (2023)
A Comprehensive Study on NLP Data Augmentation for Hate Speech Detection: Legacy Methods, BERT, and LLMs
by: Jahan, Md Saroar, et al.
Published: (2024)
by: Jahan, Md Saroar, et al.
Published: (2024)
Keeping and breeding Haaniella species successfully
by: Abercrombie, Ian
Published: (1993)
by: Abercrombie, Ian
Published: (1993)
The Ethnos, Histories, and Cultures of Ethnohistory: A view from the US Academy
by: Thomas Abercrombie
Published: (2012)
by: Thomas Abercrombie
Published: (2012)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
by: Kumar, Aswini, et al.
Published: (2025)
by: Kumar, Aswini, et al.
Published: (2025)
Digitale Hate Speech
Published: (2023)
Published: (2023)
Hate Speech Law
by: Brown, Alex
Published: (2025)
by: Brown, Alex
Published: (2025)
HateDebias: On the Diversity and Variability of Hate Speech Debiasing
by: Wu, Hongyan, et al.
Published: (2024)
by: Wu, Hongyan, et al.
Published: (2024)
Reproductive biology of the Big Brown Bat (Eptesicus fuscus) in Alberta
by: Schowalter, D. (Tim), et al.
Published: (1979)
by: Schowalter, D. (Tim), et al.
Published: (1979)
CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
by: Nikandrou, Malvina, et al.
Published: (2024)
by: Nikandrou, Malvina, et al.
Published: (2024)
Task Formulation Matters When Learning Continually: A Case Study in Visual Question Answering
by: Nikandrou, Mavina, et al.
Published: (2022)
by: Nikandrou, Mavina, et al.
Published: (2022)
Similar Items
-
Erasing 'Ugly' from the Internet: Propagation of the Beauty Myth in Text-Image Models
by: Dinkar, Tanvi, et al.
Published: (2025) -
Re-examining Sexism and Misogyny Classification with Annotator Attitudes
by: Jiang, Aiqi, et al.
Published: (2024) -
NLP for Counterspeech against Hate: A Survey and How-To Guide
by: Bonaldi, Helena, et al.
Published: (2024) -
Food Noise & False Safety: A Systematic Evaluation of How LLMs Fail to Adapt to Eating Disorder Queries with Clinician Feedback
by: Pucci, Giulia, et al.
Published: (2026) -
Consistency is Key: Disentangling Label Variation in Natural Language Processing with Intra-Annotator Agreement
by: Abercrombie, Gavin, et al.
Published: (2023)