Intent-conditioned and Non-toxic Counterspeech Generation using Multi-Task Instruction Tuning with RLAIF
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hengle, Amey, Kumar, Aswini, Singh, Sahajpreet, Bandhakavi, Anil, Akhtar, Md Shad, Chakroborty, Tanmoy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CSEval: Towards Automated, Multi-Dimensional, and Reference-Free Counterspeech Evaluation using Auto-Calibrated LLMs
von: Hengle, Amey, et al.
Veröffentlicht: (2025)
von: Hengle, Amey, et al.
Veröffentlicht: (2025)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
von: Kumar, Aswini, et al.
Veröffentlicht: (2025)
von: Kumar, Aswini, et al.
Veröffentlicht: (2025)
SemEval 2024 -- Task 10: Emotion Discovery and Reasoning its Flip in Conversation (EDiReF)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
Persona-aware Generative Model for Code-mixed Language
von: Sengupta, Ayan, et al.
Veröffentlicht: (2023)
von: Sengupta, Ayan, et al.
Veröffentlicht: (2023)
Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacks
von: Hengle, Amey, et al.
Veröffentlicht: (2025)
von: Hengle, Amey, et al.
Veröffentlicht: (2025)
Emotion-Aware Multimodal Fusion for Meme Emotion Detection
von: Sharma, Shivam, et al.
Veröffentlicht: (2024)
von: Sharma, Shivam, et al.
Veröffentlicht: (2024)
Independent fact-checking organizations exhibit a departure from political neutrality
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2024)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2024)
From Images to Words: Efficient Cross-Modal Knowledge Distillation to Language Models from Black-box Teachers
von: Sengupta, Ayan, et al.
Veröffentlicht: (2026)
von: Sengupta, Ayan, et al.
Veröffentlicht: (2026)
Knowledge Planning in Large Language Models for Domain-Aligned Counseling Summarization
von: Srivastava, Aseem, et al.
Veröffentlicht: (2024)
von: Srivastava, Aseem, et al.
Veröffentlicht: (2024)
Crowd Intelligence for Early Misinformation Prediction on Social Media
von: Sundriyal, Megha, et al.
Veröffentlicht: (2024)
von: Sundriyal, Megha, et al.
Veröffentlicht: (2024)
Sentiment-guided Commonsense-aware Response Generation for Mental Health Counseling
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
Synthetic Data Generation and Joint Learning for Robust Code-Mixed Translation
von: Kartik, Kartik, et al.
Veröffentlicht: (2024)
von: Kartik, Kartik, et al.
Veröffentlicht: (2024)
Target-Augmented Shared Fusion-based Multimodal Sarcasm Explanation Generation
von: Goel, Palaash, et al.
Veröffentlicht: (2025)
von: Goel, Palaash, et al.
Veröffentlicht: (2025)
Trust Modeling in Counseling Conversations: A Benchmark Study
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
von: Srivastava, Aseem, et al.
Veröffentlicht: (2025)
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech
von: Yadav, Neemesh, et al.
Veröffentlicht: (2024)
von: Yadav, Neemesh, et al.
Veröffentlicht: (2024)
Diversity Augmentation of Dynamic User Preference Data for Boosting Personalized Text Summarizers
von: Chatterjee, Parthiv, et al.
Veröffentlicht: (2025)
von: Chatterjee, Parthiv, et al.
Veröffentlicht: (2025)
QUENCH: Measuring the gap between Indic and Non-Indic Contextual General Reasoning in LLMs
von: Khan, Mohammad Aflah, et al.
Veröffentlicht: (2024)
von: Khan, Mohammad Aflah, et al.
Veröffentlicht: (2024)
Probing Critical Learning Dynamics of PLMs for Hate Speech Detection
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
Hate Personified: Investigating the role of LLMs in content moderation
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
von: Masud, Sarah, et al.
Veröffentlicht: (2024)
EROS: Entity-Driven Controlled Policy Document Summarization
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
Incongruence Identification in Eyewitness Testimony
von: Nair, Akshara, et al.
Veröffentlicht: (2025)
von: Nair, Akshara, et al.
Veröffentlicht: (2025)
GitSearch: Enhancing Community Notes Generation with Gap-Informed Targeted Search
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
Counterspeech
Veröffentlicht: (2025)
Veröffentlicht: (2025)
HateMirage: An Explainable Multi-Dimensional Dataset for Decoding Faux Hate and Subtle Online Abuse
von: Kasu, Sai Kartheek Reddy, et al.
Veröffentlicht: (2026)
von: Kasu, Sai Kartheek Reddy, et al.
Veröffentlicht: (2026)
On the Effect of Instruction Tuning Loss on Generalization
von: Chatterjee, Anwoy, et al.
Veröffentlicht: (2025)
von: Chatterjee, Anwoy, et al.
Veröffentlicht: (2025)
Image denoising as a conditional expectation
von: Chakroborty, Sajal, et al.
Veröffentlicht: (2025)
von: Chakroborty, Sajal, et al.
Veröffentlicht: (2025)
Assess and Prompt: A Generative RL Framework for Improving Engagement in Online Mental Health Communities
von: Gaur, Bhagesh, et al.
Veröffentlicht: (2025)
von: Gaur, Bhagesh, et al.
Veröffentlicht: (2025)
Labels or Input? Rethinking Augmentation in Multimodal Hate Detection
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2025)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2025)
No perspective, no perception!! Perspective-aware Healthcare Answer Summarization
von: Naik, Gauri, et al.
Veröffentlicht: (2024)
von: Naik, Gauri, et al.
Veröffentlicht: (2024)
MAMA-Memeia! Multi-Aspect Multi-Agent Collaboration for Depressive Symptoms Identification in Memes
von: Agarwal, Siddhant, et al.
Veröffentlicht: (2025)
von: Agarwal, Siddhant, et al.
Veröffentlicht: (2025)
AI Across Borders: Exploring Perceptions and Interactions in Higher Education
von: Gerard, Juliana, et al.
Veröffentlicht: (2024)
von: Gerard, Juliana, et al.
Veröffentlicht: (2024)
Making Effective Statistical Inferences: From Significance Testing to the Open Science Inference Ecosystem (2016-2026)
von: Patra, Aswini Kumar
Veröffentlicht: (2026)
von: Patra, Aswini Kumar
Veröffentlicht: (2026)
On Zero-Shot Counterspeech Generation by LLMs
von: Saha, Punyajoy, et al.
Veröffentlicht: (2024)
von: Saha, Punyajoy, et al.
Veröffentlicht: (2024)
CommunityFact: A Dynamic, Multilingual, Multi-domain Benchmark for Misinformation Detection in the Wild
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
CODEOFCONDUCT at Multilingual Counterspeech Generation: A Context-Aware Model for Robust Counterspeech Generation in Low-Resource Languages
von: Bennie, Michael, et al.
Veröffentlicht: (2025)
von: Bennie, Michael, et al.
Veröffentlicht: (2025)
Redefining Experts: Interpretable Decomposition of Language Models for Toxicity Mitigation
von: Shaik, Zuhair Hasan, et al.
Veröffentlicht: (2025)
von: Shaik, Zuhair Hasan, et al.
Veröffentlicht: (2025)
The study of hadronic rescattering on $K^{*0}$ resonance yield in baryon-rich QCD matter
von: Sahoo, Aswini Kumar, et al.
Veröffentlicht: (2023)
von: Sahoo, Aswini Kumar, et al.
Veröffentlicht: (2023)
The study of $K^{*0}$ meson production using a multi-phase transport model at RHIC BES energies
von: Barik, Pranjal, et al.
Veröffentlicht: (2026)
von: Barik, Pranjal, et al.
Veröffentlicht: (2026)
Scrotal Hirudiniasis: The Mysterious Migration of a Leech Into the Scrotum: A Case Report
von: Pranto Chakroborty, et al.
Veröffentlicht: (2026)
von: Pranto Chakroborty, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CSEval: Towards Automated, Multi-Dimensional, and Reference-Free Counterspeech Evaluation using Auto-Calibrated LLMs
von: Hengle, Amey, et al.
Veröffentlicht: (2025) -
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
von: Kumar, Aswini, et al.
Veröffentlicht: (2025) -
SemEval 2024 -- Task 10: Emotion Discovery and Reasoning its Flip in Conversation (EDiReF)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024) -
Persona-aware Generative Model for Code-mixed Language
von: Sengupta, Ayan, et al.
Veröffentlicht: (2023) -
Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models
von: Hengle, Amey, et al.
Veröffentlicht: (2024)