Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yadav, Neemesh, Masud, Sarah, Goyal, Vikram, Akhtar, Md Shad, Chakraborty, Tanmoy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Probing Critical Learning Dynamics of PLMs for Hate Speech Detection
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
QUENCH: Measuring the gap between Indic and Non-Indic Contextual General Reasoning in LLMs
von: Khan, Mohammad Aflah, et al.
Veröffentlicht: (2024)
von: Khan, Mohammad Aflah, et al.
Veröffentlicht: (2024)
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
von: Masud, Sarah, et al.
Veröffentlicht: (2023)
von: Masud, Sarah, et al.
Veröffentlicht: (2023)
Persona-aware Generative Model for Code-mixed Language
von: Sengupta, Ayan, et al.
Veröffentlicht: (2023)
von: Sengupta, Ayan, et al.
Veröffentlicht: (2023)
Knowledge Planning in Large Language Models for Domain-Aligned Counseling Summarization
von: Srivastava, Aseem, et al.
Veröffentlicht: (2024)
von: Srivastava, Aseem, et al.
Veröffentlicht: (2024)
Crowd Intelligence for Early Misinformation Prediction on Social Media
von: Sundriyal, Megha, et al.
Veröffentlicht: (2024)
von: Sundriyal, Megha, et al.
Veröffentlicht: (2024)
From Images to Words: Efficient Cross-Modal Knowledge Distillation to Language Models from Black-box Teachers
von: Sengupta, Ayan, et al.
Veröffentlicht: (2026)
von: Sengupta, Ayan, et al.
Veröffentlicht: (2026)
Hate Personified: Investigating the role of LLMs in content moderation
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
Sentiment-guided Commonsense-aware Response Generation for Mental Health Counseling
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
SemEval 2024 -- Task 10: Emotion Discovery and Reasoning its Flip in Conversation (EDiReF)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
Trust Modeling in Counseling Conversations: A Benchmark Study
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
Synthetic Data Generation and Joint Learning for Robust Code-Mixed Translation
von: Kartik, Kartik, et al.
Veröffentlicht: (2024)
von: Kartik, Kartik, et al.
Veröffentlicht: (2024)
Target-Augmented Shared Fusion-based Multimodal Sarcasm Explanation Generation
von: Goel, Palaash, et al.
Veröffentlicht: (2025)
von: Goel, Palaash, et al.
Veröffentlicht: (2025)
SymTax: Symbiotic Relationship and Taxonomy Fusion for Effective Citation Recommendation
von: Goyal, Karan, et al.
Veröffentlicht: (2024)
von: Goyal, Karan, et al.
Veröffentlicht: (2024)
LifeTox: Unveiling Implicit Toxicity in Life Advice
von: Kim, Minbeom, et al.
Veröffentlicht: (2023)
von: Kim, Minbeom, et al.
Veröffentlicht: (2023)
Public Profile Matters: A Scalable Integrated Approach to Recommend Citations in the Wild
von: Goyal, Karan, et al.
Veröffentlicht: (2026)
von: Goyal, Karan, et al.
Veröffentlicht: (2026)
HateMirage: An Explainable Multi-Dimensional Dataset for Decoding Faux Hate and Subtle Online Abuse
von: Kasu, Sai Kartheek Reddy, et al.
Veröffentlicht: (2026)
von: Kasu, Sai Kartheek Reddy, et al.
Veröffentlicht: (2026)
Information Anxiety in Large Language Models
von: Bajpai, Prasoon, et al.
Veröffentlicht: (2024)
von: Bajpai, Prasoon, et al.
Veröffentlicht: (2024)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
von: Nandi, Palash, et al.
Veröffentlicht: (2024)
von: Nandi, Palash, et al.
Veröffentlicht: (2024)
Redefining Experts: Interpretable Decomposition of Language Models for Toxicity Mitigation
von: Shaik, Zuhair Hasan, et al.
Veröffentlicht: (2025)
von: Shaik, Zuhair Hasan, et al.
Veröffentlicht: (2025)
No perspective, no perception!! Perspective-aware Healthcare Answer Summarization
von: Naik, Gauri, et al.
Veröffentlicht: (2024)
von: Naik, Gauri, et al.
Veröffentlicht: (2024)
IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages
von: Dawar, Aviral, et al.
Veröffentlicht: (2026)
von: Dawar, Aviral, et al.
Veröffentlicht: (2026)
Temporally Consistent Factuality Probing for Large Language Models
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech Detection
von: Brook, Joshua Wolfe, et al.
Veröffentlicht: (2025)
von: Brook, Joshua Wolfe, et al.
Veröffentlicht: (2025)
Intent-conditioned and Non-toxic Counterspeech Generation using Multi-Task Instruction Tuning with RLAIF
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
Emotion-Aware Multimodal Fusion for Meme Emotion Detection
von: Sharma, Shivam, et al.
Veröffentlicht: (2024)
von: Sharma, Shivam, et al.
Veröffentlicht: (2024)
Incongruence Identification in Eyewitness Testimony
von: Nair, Akshara, et al.
Veröffentlicht: (2025)
von: Nair, Akshara, et al.
Veröffentlicht: (2025)
Recent Advances in Hate Speech Moderation: Multimodality and the Role of Large Models
von: Hee, Ming Shan, et al.
Veröffentlicht: (2024)
von: Hee, Ming Shan, et al.
Veröffentlicht: (2024)
ToxSyn: Reducing Bias in Hate Speech Detection via Synthetic Minority Data in Brazilian Portuguese
von: Brito, Iago Alves, et al.
Veröffentlicht: (2025)
von: Brito, Iago Alves, et al.
Veröffentlicht: (2025)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
von: Kumar, Aswini, et al.
Veröffentlicht: (2025)
von: Kumar, Aswini, et al.
Veröffentlicht: (2025)
ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues
von: Vedula, Bhaskara Hanuma, et al.
Veröffentlicht: (2026)
von: Vedula, Bhaskara Hanuma, et al.
Veröffentlicht: (2026)
Incorporating Human Explanations for Robust Hate Speech Detection
von: Chen, Jennifer L., et al.
Veröffentlicht: (2024)
von: Chen, Jennifer L., et al.
Veröffentlicht: (2024)
MAMA-Memeia! Multi-Aspect Multi-Agent Collaboration for Depressive Symptoms Identification in Memes
von: Agarwal, Siddhant, et al.
Veröffentlicht: (2025)
von: Agarwal, Siddhant, et al.
Veröffentlicht: (2025)
HatePrototypes: Interpretable and Transferable Representations for Implicit and Explicit Hate Speech Detection
von: Proskurina, Irina, et al.
Veröffentlicht: (2025)
von: Proskurina, Irina, et al.
Veröffentlicht: (2025)
SyntaxMind at BLP-2025 Task 1: Leveraging Attention Fusion of CNN and GRU for Hate Speech Detection
von: Riad, Md. Shihab Uddin
Veröffentlicht: (2026)
von: Riad, Md. Shihab Uddin
Veröffentlicht: (2026)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
von: Almohaimeed, Saad, et al.
Veröffentlicht: (2025)
von: Almohaimeed, Saad, et al.
Veröffentlicht: (2025)
Assess and Prompt: A Generative RL Framework for Improving Engagement in Online Mental Health Communities
von: Gaur, Bhagesh, et al.
Veröffentlicht: (2025)
von: Gaur, Bhagesh, et al.
Veröffentlicht: (2025)
Specializing General-purpose LLM Embeddings for Implicit Hate Speech Detection across Datasets
von: Cheremetiev, Vassiliy, et al.
Veröffentlicht: (2025)
von: Cheremetiev, Vassiliy, et al.
Veröffentlicht: (2025)
HateXScore: A Metric Suite for Evaluating Reasoning Quality in Hate Speech Explanations
von: Hu, Yujia, et al.
Veröffentlicht: (2026)
von: Hu, Yujia, et al.
Veröffentlicht: (2026)
PERCS: Persona-Guided Controllable Biomedical Summarization Dataset
von: Salvi, Rohan Charudatt, et al.
Veröffentlicht: (2025)
von: Salvi, Rohan Charudatt, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Probing Critical Learning Dynamics of PLMs for Hate Speech Detection
von: Masud, Sarah, et al.
Veröffentlicht: (2024) -
QUENCH: Measuring the gap between Indic and Non-Indic Contextual General Reasoning in LLMs
von: Khan, Mohammad Aflah, et al.
Veröffentlicht: (2024) -
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
von: Masud, Sarah, et al.
Veröffentlicht: (2023) -
Persona-aware Generative Model for Code-mixed Language
von: Sengupta, Ayan, et al.
Veröffentlicht: (2023) -
Knowledge Planning in Large Language Models for Domain-Aligned Counseling Summarization
von: Srivastava, Aseem, et al.
Veröffentlicht: (2024)