Improving Implicit Hate Speech Detection via a Community-Driven Multi-Agent Framework
Fuente:
arXiv
Salvato in:
| Autori principali: | Gajewska, Ewelina, Budzynska, Katarzyna, Chudziak, Jarosław A |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
Who Decides What Is Harmful? Content Moderation Policy Through A Multi-Agent Personalised Inference Framework
di: Gajewska, Ewelina, et al.
Pubblicazione: (2026)
di: Gajewska, Ewelina, et al.
Pubblicazione: (2026)
Leveraging a Multi-Agent LLM-Based System to Educate Teachers in Hate Incidents Management
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
On Verifiable Legal Reasoning: A Multi-Agent Framework with Formalized Knowledge Representations
di: Sadowski, Albert, et al.
Pubblicazione: (2025)
di: Sadowski, Albert, et al.
Pubblicazione: (2025)
A Natural Language Agentic Approach to Study Affective Polarization
di: Malvicini, Stephanie Anneris, et al.
Pubblicazione: (2026)
di: Malvicini, Stephanie Anneris, et al.
Pubblicazione: (2026)
Multi-Agent Dialectical Refinement for Enhanced Argument Classification
di: Bąba, Jakub, et al.
Pubblicazione: (2026)
di: Bąba, Jakub, et al.
Pubblicazione: (2026)
Predicting the winner of the US 2024 elections using trust analytics
di: Budzynska, Katarzyna, et al.
Pubblicazione: (2024)
di: Budzynska, Katarzyna, et al.
Pubblicazione: (2024)
Selective Demonstration Retrieval for Improved Implicit Hate Speech Detection
di: Kim, Yumin, et al.
Pubblicazione: (2025)
di: Kim, Yumin, et al.
Pubblicazione: (2025)
Explainable Rule Application via Structured Prompting: A Neural-Symbolic Approach
di: Sadowski, Albert, et al.
Pubblicazione: (2025)
di: Sadowski, Albert, et al.
Pubblicazione: (2025)
HatePrototypes: Interpretable and Transferable Representations for Implicit and Explicit Hate Speech Detection
di: Proskurina, Irina, et al.
Pubblicazione: (2025)
di: Proskurina, Irina, et al.
Pubblicazione: (2025)
GAIus: Combining Genai with Legal Clauses Retrieval for Knowledge-based Assistant
di: Matak, Michał, et al.
Pubblicazione: (2025)
di: Matak, Michał, et al.
Pubblicazione: (2025)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
di: Almohaimeed, Saad, et al.
Pubblicazione: (2025)
di: Almohaimeed, Saad, et al.
Pubblicazione: (2025)
Heterogeneous Debate Engine: Identity-Grounded Cognitive Architecture for Resilient LLM-Based Ethical Tutoring
di: Masłowski, Jakub, et al.
Pubblicazione: (2026)
di: Masłowski, Jakub, et al.
Pubblicazione: (2026)
xList-Hate: A Checklist-Based Framework for Interpretable and Generalizable Hate Speech Detection
di: Girón, Adrián, et al.
Pubblicazione: (2026)
di: Girón, Adrián, et al.
Pubblicazione: (2026)
Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory
di: Sadowski, Albert, et al.
Pubblicazione: (2026)
di: Sadowski, Albert, et al.
Pubblicazione: (2026)
RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection
di: Lee, Yejin, et al.
Pubblicazione: (2025)
di: Lee, Yejin, et al.
Pubblicazione: (2025)
TACLA: An LLM-Based Multi-Agent Tool for Transactional Analysis Training in Education
di: Zamojska, Monika, et al.
Pubblicazione: (2025)
di: Zamojska, Monika, et al.
Pubblicazione: (2025)
A Federated Approach to Few-Shot Hate Speech Detection for Marginalized Communities
di: Ye, Haotian, et al.
Pubblicazione: (2024)
di: Ye, Haotian, et al.
Pubblicazione: (2024)
Ethos and Pathos in Online Group Discussions: Corpora for Polarisation Issues in Social Media
di: Gajewska, Ewelina, et al.
Pubblicazione: (2024)
di: Gajewska, Ewelina, et al.
Pubblicazione: (2024)
Hierarchical Sentiment Analysis Framework for Hate Speech Detection: Implementing Binary and Multiclass Classification Strategy
di: Naznin, Faria, et al.
Pubblicazione: (2024)
di: Naznin, Faria, et al.
Pubblicazione: (2024)
Towards Fairness Assessment of Dutch Hate Speech Detection
di: Bauer, Julie, et al.
Pubblicazione: (2025)
di: Bauer, Julie, et al.
Pubblicazione: (2025)
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
di: Lee, Yejin, et al.
Pubblicazione: (2025)
di: Lee, Yejin, et al.
Pubblicazione: (2025)
Games Agents Play: Towards Transactional Analysis in LLM-based Multi-Agent Systems
di: Zamojska, Monika, et al.
Pubblicazione: (2025)
di: Zamojska, Monika, et al.
Pubblicazione: (2025)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
ToxSyn: Reducing Bias in Hate Speech Detection via Synthetic Minority Data in Brazilian Portuguese
di: Brito, Iago Alves, et al.
Pubblicazione: (2025)
di: Brito, Iago Alves, et al.
Pubblicazione: (2025)
Conditioning Large Language Models on Legal Systems? Detecting Punishable Hate Speech
di: Ludwig, Florian, et al.
Pubblicazione: (2025)
di: Ludwig, Florian, et al.
Pubblicazione: (2025)
Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages
di: Prome, Ruhina Tabasshum, et al.
Pubblicazione: (2025)
di: Prome, Ruhina Tabasshum, et al.
Pubblicazione: (2025)
HateXScore: A Metric Suite for Evaluating Reasoning Quality in Hate Speech Explanations
di: Hu, Yujia, et al.
Pubblicazione: (2026)
di: Hu, Yujia, et al.
Pubblicazione: (2026)
Evaluating Theory of Mind and Internal Beliefs in LLM-Based Multi-Agent Systems
di: Kostka, Adam, et al.
Pubblicazione: (2026)
di: Kostka, Adam, et al.
Pubblicazione: (2026)
Empirical Evaluation of Public HateSpeech Datasets
di: Jaf, Sadar, et al.
Pubblicazione: (2024)
di: Jaf, Sadar, et al.
Pubblicazione: (2024)
Hate Speech Detection using Large Language Models with Data Augmentation and Feature Enhancement
di: Nge, Brian Jing Hong, et al.
Pubblicazione: (2026)
di: Nge, Brian Jing Hong, et al.
Pubblicazione: (2026)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
di: Ahmed, Ahmed Haj, et al.
Pubblicazione: (2024)
di: Ahmed, Ahmed Haj, et al.
Pubblicazione: (2024)
An Investigation Into Explainable Audio Hate Speech Detection
di: An, Jinmyeong, et al.
Pubblicazione: (2024)
di: An, Jinmyeong, et al.
Pubblicazione: (2024)
More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection
di: Sun, Runze, et al.
Pubblicazione: (2026)
di: Sun, Runze, et al.
Pubblicazione: (2026)
SEAHateCheck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of Southeast Asia
di: Ng, Ri Chi, et al.
Pubblicazione: (2026)
di: Ng, Ri Chi, et al.
Pubblicazione: (2026)
Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
di: Piot, Paloma, et al.
Pubblicazione: (2025)
di: Piot, Paloma, et al.
Pubblicazione: (2025)
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
Feature Selection Empowered BERT for Detection of Hate Speech with Vocabulary Augmentation
di: Desai, Pritish N., et al.
Pubblicazione: (2025)
di: Desai, Pritish N., et al.
Pubblicazione: (2025)
MultiHedge: Adaptive Coordination via Retrieval-Augmented Control
di: Bańka, Feliks, et al.
Pubblicazione: (2026)
di: Bańka, Feliks, et al.
Pubblicazione: (2026)
System Report for CCL25-Eval Task 10: Prompt-Driven Large Language Model Merge for Fine-Grained Chinese Hate Speech Detection
di: Wu, Binglin, et al.
Pubblicazione: (2025)
di: Wu, Binglin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025) -
Who Decides What Is Harmful? Content Moderation Policy Through A Multi-Agent Personalised Inference Framework
di: Gajewska, Ewelina, et al.
Pubblicazione: (2026) -
Leveraging a Multi-Agent LLM-Based System to Educate Teachers in Hate Incidents Management
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025) -
On Verifiable Legal Reasoning: A Multi-Agent Framework with Formalized Knowledge Representations
di: Sadowski, Albert, et al.
Pubblicazione: (2025) -
A Natural Language Agentic Approach to Study Affective Polarization
di: Malvicini, Stephanie Anneris, et al.
Pubblicazione: (2026)