Beating Harmful Stereotypes Through Facts: RAG-based Counter-speech Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Damo, Greta, Cabrio, Elena, Villata, Serena |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Effectiveness of Counter-Speech against Abusive Content: A Multidimensional Annotation and Classification Study
by: Damo, Greta, et al.
Published: (2025)
by: Damo, Greta, et al.
Published: (2025)
PEACE 2.0: Grounded Explanations and Counter-Speech for Combating Hate Expressions
by: Damo, Greta, et al.
Published: (2026)
by: Damo, Greta, et al.
Published: (2026)
Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech Countering
by: Bonaldi, Helena, et al.
Published: (2024)
by: Bonaldi, Helena, et al.
Published: (2024)
RooseBERT: A New Deal For Political Language Modelling
by: Dore, Deborah, et al.
Published: (2025)
by: Dore, Deborah, et al.
Published: (2025)
Compact Prompting in Instruction-tuned LLMs for Joint Argumentative Component Detection
by: Elguendouze, Sofiane, et al.
Published: (2026)
by: Elguendouze, Sofiane, et al.
Published: (2026)
CasiMedicos-Arg: A Medical Question Answering Dataset Annotated with Explanatory Argumentative Structures
by: Sviridova, Ekaterina, et al.
Published: (2024)
by: Sviridova, Ekaterina, et al.
Published: (2024)
Argument Quality Assessment in the Age of Instruction-Following Large Language Models
by: Wachsmuth, Henning, et al.
Published: (2024)
by: Wachsmuth, Henning, et al.
Published: (2024)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
by: Nejadgholi, Isar, et al.
Published: (2024)
by: Nejadgholi, Isar, et al.
Published: (2024)
Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking
by: Russo, Daniel, et al.
Published: (2024)
by: Russo, Daniel, et al.
Published: (2024)
How Are LLMs Mitigating Stereotyping Harms? Learning from Search Engine Studies
by: Leidinger, Alina, et al.
Published: (2024)
by: Leidinger, Alina, et al.
Published: (2024)
Show Me the Work: Fact-Checkers' Requirements for Explainable Automated Fact-Checking
by: Warren, Greta, et al.
Published: (2025)
by: Warren, Greta, et al.
Published: (2025)
Stakeholder Suite: A Unified AI Framework for Mapping Actors, Topics and Arguments in Public Debates
by: Chenene, Mohamed, et al.
Published: (2025)
by: Chenene, Mohamed, et al.
Published: (2025)
Explaining Sources of Uncertainty in Automated Fact-Checking
by: Sun, Jingyi, et al.
Published: (2025)
by: Sun, Jingyi, et al.
Published: (2025)
Beyond Fact Retrieval: Episodic Memory for RAG with Generative Semantic Workspaces
by: Rajesh, Shreyas, et al.
Published: (2025)
by: Rajesh, Shreyas, et al.
Published: (2025)
FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation
by: Zhang, Qinggang, et al.
Published: (2025)
by: Zhang, Qinggang, et al.
Published: (2025)
HyperRAG: Reasoning N-ary Facts over Hypergraphs for Retrieval Augmented Generation
by: Lien, Wen-Sheng, et al.
Published: (2026)
by: Lien, Wen-Sheng, et al.
Published: (2026)
Can Community Notes Replace Professional Fact-Checkers?
by: Borenstein, Nadav, et al.
Published: (2025)
by: Borenstein, Nadav, et al.
Published: (2025)
VaccineRAG: Boosting Multimodal Large Language Models' Immunity to Harmful RAG Samples
by: Sun, Qixin, et al.
Published: (2025)
by: Sun, Qixin, et al.
Published: (2025)
CommunityKG-RAG: Leveraging Community Structures in Knowledge Graphs for Advanced Retrieval-Augmented Generation in Fact-Checking
by: Chang, Rong-Ching, et al.
Published: (2024)
by: Chang, Rong-Ching, et al.
Published: (2024)
Medical mT5: An Open-Source Multilingual Text-to-Text LLM for The Medical Domain
by: García-Ferrero, Iker, et al.
Published: (2024)
by: García-Ferrero, Iker, et al.
Published: (2024)
ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking
by: Wu, Yutao, et al.
Published: (2025)
by: Wu, Yutao, et al.
Published: (2025)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
by: Robles, Melissa, et al.
Published: (2025)
by: Robles, Melissa, et al.
Published: (2025)
ChronoFact: Timeline-based Temporal Fact Verification
by: Barik, Anab Maulana, et al.
Published: (2024)
by: Barik, Anab Maulana, et al.
Published: (2024)
Quantifying Stereotypes in Language
by: Liu, Yang
Published: (2024)
by: Liu, Yang
Published: (2024)
The Psychosocial Impacts of Generative AI Harms
by: Vassel, Faye-Marie, et al.
Published: (2024)
by: Vassel, Faye-Marie, et al.
Published: (2024)
Robust Claim Verification Through Fact Detection
by: Jafari, Nazanin, et al.
Published: (2024)
by: Jafari, Nazanin, et al.
Published: (2024)
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
by: Pradeep, Ronak, et al.
Published: (2025)
by: Pradeep, Ronak, et al.
Published: (2025)
IndexRAG: Bridging Facts for Cross-Document Reasoning at Index Time
by: Bao, Zhenghua, et al.
Published: (2026)
by: Bao, Zhenghua, et al.
Published: (2026)
Evidence-backed Fact Checking using RAG and Few-Shot In-Context Learning with LLMs
by: Singhal, Ronit, et al.
Published: (2024)
by: Singhal, Ronit, et al.
Published: (2024)
Simulating Identity, Propagating Bias: Abstraction and Stereotypes in LLM-Generated Text
by: Sommerauer, Pia, et al.
Published: (2025)
by: Sommerauer, Pia, et al.
Published: (2025)
MBBQ: A Dataset for Cross-Lingual Comparison of Stereotypes in Generative LLMs
by: Neplenbroek, Vera, et al.
Published: (2024)
by: Neplenbroek, Vera, et al.
Published: (2024)
Enhancing Democratic Deliberations with AI: Insights from the ORBIS Co-creation Journey
by: MARIANI, ILARIA, et al.
Published: (2026)
by: MARIANI, ILARIA, et al.
Published: (2026)
Classification is a RAG problem: A case study on hate speech detection
by: Willats, Richard, et al.
Published: (2025)
by: Willats, Richard, et al.
Published: (2025)
Basque and Spanish Counter Narrative Generation: Data Creation and Evaluation
by: Bengoetxea, Jaione, et al.
Published: (2024)
by: Bengoetxea, Jaione, et al.
Published: (2024)
Local Contrastive Editing of Gender Stereotypes
by: Lutz, Marlene, et al.
Published: (2024)
by: Lutz, Marlene, et al.
Published: (2024)
When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering
by: Astaraki, Mahdi, et al.
Published: (2026)
by: Astaraki, Mahdi, et al.
Published: (2026)
FactSim: Fact-Checking for Opinion Summarization
by: Anghinoni, Leandro, et al.
Published: (2026)
by: Anghinoni, Leandro, et al.
Published: (2026)
LLM-based Semantic Augmentation for Harmful Content Detection
by: Meguellati, Elyas, et al.
Published: (2025)
by: Meguellati, Elyas, et al.
Published: (2025)
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments
by: Verma, Preetika, et al.
Published: (2024)
by: Verma, Preetika, et al.
Published: (2024)
LLMs Reproduce Stereotypes of Sexual and Gender Minorities
by: Ostrow, Ruby, et al.
Published: (2025)
by: Ostrow, Ruby, et al.
Published: (2025)
Similar Items
-
Effectiveness of Counter-Speech against Abusive Content: A Multidimensional Annotation and Classification Study
by: Damo, Greta, et al.
Published: (2025) -
PEACE 2.0: Grounded Explanations and Counter-Speech for Combating Hate Expressions
by: Damo, Greta, et al.
Published: (2026) -
Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech Countering
by: Bonaldi, Helena, et al.
Published: (2024) -
RooseBERT: A New Deal For Political Language Modelling
by: Dore, Deborah, et al.
Published: (2025) -
Compact Prompting in Instruction-tuned LLMs for Joint Argumentative Component Detection
by: Elguendouze, Sofiane, et al.
Published: (2026)