On Zero-Shot Counterspeech Generation by LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Saha, Punyajoy, Agrawal, Aalok, Jana, Abhik, Biemann, Chris, Mukherjee, Animesh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
di: Saha, Punyajoy, et al.
Pubblicazione: (2024)
di: Saha, Punyajoy, et al.
Pubblicazione: (2024)
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
di: Das, Mithun, et al.
Pubblicazione: (2024)
di: Das, Mithun, et al.
Pubblicazione: (2024)
Demarked: A Strategy for Enhanced Abusive Speech Moderation through Counterspeech, Detoxification, and Message Management
di: Yimam, Seid Muhie, et al.
Pubblicazione: (2024)
di: Yimam, Seid Muhie, et al.
Pubblicazione: (2024)
Exploring the Limits of Zero Shot Vision Language Models for Hate Meme Detection: The Vulnerabilities and their Interpretations
di: Rizwan, Naquee, et al.
Pubblicazione: (2024)
di: Rizwan, Naquee, et al.
Pubblicazione: (2024)
InfFeed: Influence Functions as a Feedback to Improve the Performance of Subjective Tasks
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
HatePRISM: Policies, Platforms, and Research Integration. Advancing NLP for Hate Speech Proactive Mitigation
di: Rizwan, Naquee, et al.
Pubblicazione: (2025)
di: Rizwan, Naquee, et al.
Pubblicazione: (2025)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
di: Kumar, Aswini, et al.
Pubblicazione: (2025)
di: Kumar, Aswini, et al.
Pubblicazione: (2025)
Perspectives - Interactive Document Clustering in the Discourse Analysis Tool Suite
di: Fischer, Tim, et al.
Pubblicazione: (2026)
di: Fischer, Tim, et al.
Pubblicazione: (2026)
CODEOFCONDUCT at Multilingual Counterspeech Generation: A Context-Aware Model for Robust Counterspeech Generation in Low-Resource Languages
di: Bennie, Michael, et al.
Pubblicazione: (2025)
di: Bennie, Michael, et al.
Pubblicazione: (2025)
Speaking at the Right Level: Literacy-Controlled Counterspeech Generation with RAG-RL
di: Song, Xiaoying, et al.
Pubblicazione: (2025)
di: Song, Xiaoying, et al.
Pubblicazione: (2025)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
di: Adak, Sayantan, et al.
Pubblicazione: (2024)
di: Adak, Sayantan, et al.
Pubblicazione: (2024)
Dataset of Quotation Attribution in German News Articles
di: Petersen-Frey, Fynn, et al.
Pubblicazione: (2024)
di: Petersen-Frey, Fynn, et al.
Pubblicazione: (2024)
Assessing the Human Likeness of AI-Generated Counterspeech
di: Song, Xiaoying, et al.
Pubblicazione: (2024)
di: Song, Xiaoying, et al.
Pubblicazione: (2024)
Can Safety Emerge from Weak Supervision? A Systematic Analysis of Small Language Models
di: Saha, Punyajoy, et al.
Pubblicazione: (2026)
di: Saha, Punyajoy, et al.
Pubblicazione: (2026)
Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text
di: Hans, Abhimanyu, et al.
Pubblicazione: (2024)
di: Hans, Abhimanyu, et al.
Pubblicazione: (2024)
Multi-Agent Retrieval-Augmented Framework for Evidence-Based Counterspeech Against Health Misinformation
di: Anik, Anirban Saha, et al.
Pubblicazione: (2025)
di: Anik, Anirban Saha, et al.
Pubblicazione: (2025)
DistALANER: Distantly Supervised Active Learning Augmented Named Entity Recognition in the Open Source Software Ecosystem
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
di: Nag, Arijit, et al.
Pubblicazione: (2024)
di: Nag, Arijit, et al.
Pubblicazione: (2024)
Self-Taught Self-Correction for Small Language Models
di: Moskvoretskii, Viktor, et al.
Pubblicazione: (2025)
di: Moskvoretskii, Viktor, et al.
Pubblicazione: (2025)
RA-MTR: A Retrieval Augmented Multi-Task Reader based Approach for Inspirational Quote Extraction from Long Documents
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
Echoes of Discord: Forecasting Hater Reactions to Counterspeech
di: Song, Xiaoying, et al.
Pubblicazione: (2025)
di: Song, Xiaoying, et al.
Pubblicazione: (2025)
Zero-Shot Stance Detection using Contextual Data Generation with LLMs
di: Mahmoudi, Ghazaleh, et al.
Pubblicazione: (2024)
di: Mahmoudi, Ghazaleh, et al.
Pubblicazione: (2024)
Efficient Continual Pre-training of LLMs for Low-resource Languages
di: Nag, Arijit, et al.
Pubblicazione: (2024)
di: Nag, Arijit, et al.
Pubblicazione: (2024)
Probing Large Language Models from A Human Behavioral Perspective
di: Wang, Xintong, et al.
Pubblicazione: (2023)
di: Wang, Xintong, et al.
Pubblicazione: (2023)
GIMMICK -- Globally Inclusive Multimodal Multitask Cultural Knowledge Benchmarking
di: Schneider, Florian, et al.
Pubblicazione: (2025)
di: Schneider, Florian, et al.
Pubblicazione: (2025)
Are LLMs Good Zero-Shot Fallacy Classifiers?
di: Pan, Fengjun, et al.
Pubblicazione: (2024)
di: Pan, Fengjun, et al.
Pubblicazione: (2024)
Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation
di: Martone, Genoveffa, et al.
Pubblicazione: (2026)
di: Martone, Genoveffa, et al.
Pubblicazione: (2026)
Just Use XML: Revisiting Joint Translation and Label Projection
di: DK, Thennal, et al.
Pubblicazione: (2026)
di: DK, Thennal, et al.
Pubblicazione: (2026)
Investigating Prompting Techniques for Zero- and Few-Shot Visual Question Answering
di: Awal, Rabiul, et al.
Pubblicazione: (2023)
di: Awal, Rabiul, et al.
Pubblicazione: (2023)
Multilingual Hate Speech Detection and Counterspeech Generation: A Comprehensive Survey and Practical Guide
di: Fesaghandis, Zahra Safdari, et al.
Pubblicazione: (2026)
di: Fesaghandis, Zahra Safdari, et al.
Pubblicazione: (2026)
How (un)ethical are instruction-centric responses of LLMs? Unveiling the vulnerabilities of safety guardrails to harmful queries
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
MVL-SIB: A Massively Multilingual Vision-Language Benchmark for Cross-Modal Topical Matching
di: Schmidt, Fabian David, et al.
Pubblicazione: (2025)
di: Schmidt, Fabian David, et al.
Pubblicazione: (2025)
Towards Zero-Shot, Controllable Dialog Planning with LLMs
di: Väth, Dirk, et al.
Pubblicazione: (2024)
di: Väth, Dirk, et al.
Pubblicazione: (2024)
LLMs Are Zero-Shot Context-Aware Simultaneous Translators
di: Koshkin, Roman, et al.
Pubblicazione: (2024)
di: Koshkin, Roman, et al.
Pubblicazione: (2024)
Better Benchmarking LLMs for Zero-Shot Dependency Parsing
di: Ezquerro, Ana, et al.
Pubblicazione: (2025)
di: Ezquerro, Ana, et al.
Pubblicazione: (2025)
Zero-Shot Belief: A Hard Problem for LLMs
di: Murzaku, John, et al.
Pubblicazione: (2025)
di: Murzaku, John, et al.
Pubblicazione: (2025)
NLP for Counterspeech against Hate: A Survey and How-To Guide
di: Bonaldi, Helena, et al.
Pubblicazione: (2024)
di: Bonaldi, Helena, et al.
Pubblicazione: (2024)
Language and Task Arithmetic with Parameter-Efficient Layers for Zero-Shot Summarization
di: Chronopoulou, Alexandra, et al.
Pubblicazione: (2023)
di: Chronopoulou, Alexandra, et al.
Pubblicazione: (2023)
Evaluating the Ebb and Flow: An In-depth Analysis of Question-Answering Trends across Diverse Platforms
di: Hazra, Rima, et al.
Pubblicazione: (2023)
di: Hazra, Rima, et al.
Pubblicazione: (2023)
Text Takes Over: A Study of Modality Bias in Multimodal Intent Detection
di: Mullick, Ankan, et al.
Pubblicazione: (2025)
di: Mullick, Ankan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
di: Saha, Punyajoy, et al.
Pubblicazione: (2024) -
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
di: Das, Mithun, et al.
Pubblicazione: (2024) -
Demarked: A Strategy for Enhanced Abusive Speech Moderation through Counterspeech, Detoxification, and Message Management
di: Yimam, Seid Muhie, et al.
Pubblicazione: (2024) -
Exploring the Limits of Zero Shot Vision Language Models for Hate Meme Detection: The Vulnerabilities and their Interpretations
di: Rizwan, Naquee, et al.
Pubblicazione: (2024) -
InfFeed: Influence Functions as a Feedback to Improve the Performance of Subjective Tasks
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)