People Make Better Edits: Measuring the Efficacy of LLM-Generated Counterfactually Augmented Data for Harmful Language Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Sen, Indira, Assenmacher, Dennis, Samory, Mattia, Augenstein, Isabelle, van der Aalst, Wil, Wagner, Claudia |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Unseen Targets of Hate -- A Systematic Review of Hateful Communication Datasets
by: Yu, Zehui, et al.
Published: (2024)
by: Yu, Zehui, et al.
Published: (2024)
Robustness and Confounders in the Demographic Alignment of LLMs with Human Perceptions of Offensiveness
by: Alipour, Shayan, et al.
Published: (2024)
by: Alipour, Shayan, et al.
Published: (2024)
PM-LLM-Benchmark: Evaluating Large Language Models on Process Mining Tasks
by: Berti, Alessandro, et al.
Published: (2024)
by: Berti, Alessandro, et al.
Published: (2024)
Translating Workflow Nets into the Partially Ordered Workflow Language
by: Kourani, Humam, et al.
Published: (2025)
by: Kourani, Humam, et al.
Published: (2025)
Evaluation of Study Plans using Partial Orders
by: Rennert, Christian, et al.
Published: (2024)
by: Rennert, Christian, et al.
Published: (2024)
Bridging Domain Knowledge and Process Discovery Using Large Language Models
by: Norouzifar, Ali, et al.
Published: (2024)
by: Norouzifar, Ali, et al.
Published: (2024)
You are a Bot! -- Studying the Development of Bot Accusations on Twitter
by: Assenmacher, Dennis, et al.
Published: (2023)
by: Assenmacher, Dennis, et al.
Published: (2023)
Learning from Convenience Samples: A Case Study on Fine-Tuning LLMs for Survey Non-response in the German Longitudinal Election Study
by: Holtdirk, Tobias, et al.
Published: (2025)
by: Holtdirk, Tobias, et al.
Published: (2025)
Understanding the Interplay between LLMs' Utilisation of Parametric and Contextual Knowledge: A keynote at ECIR 2025
by: Augenstein, Isabelle
Published: (2026)
by: Augenstein, Isabelle
Published: (2026)
ProMoAI: Process Modeling with Generative AI
by: Kourani, Humam, et al.
Published: (2024)
by: Kourani, Humam, et al.
Published: (2024)
Semantic Sensitivities and Inconsistent Predictions: Measuring the Fragility of NLI Models
by: Arakelyan, Erik, et al.
Published: (2024)
by: Arakelyan, Erik, et al.
Published: (2024)
A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance
by: Melis, Matteo, et al.
Published: (2025)
by: Melis, Matteo, et al.
Published: (2025)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
by: Samory, Mattia, et al.
Published: (2025)
by: Samory, Mattia, et al.
Published: (2025)
Neuro-Symbolic Process Anomaly Detection
by: Gaikwad, Devashish, et al.
Published: (2026)
by: Gaikwad, Devashish, et al.
Published: (2026)
Sexism Detection on a Data Diet
by: Bandyopadhyay, Rabiraj, et al.
Published: (2024)
by: Bandyopadhyay, Rabiraj, et al.
Published: (2024)
Beyond the Explicit: A Bilingual Dataset for Dehumanization Detection in Social Media
by: Assenmacher, Dennis, et al.
Published: (2025)
by: Assenmacher, Dennis, et al.
Published: (2025)
Aggregating Soft Labels from Crowd Annotations Improves Uncertainty Estimation Under Distribution Shift
by: Wright, Dustin, et al.
Published: (2022)
by: Wright, Dustin, et al.
Published: (2022)
Discriminative Rule Learning for Outcome-Guided Process Model Discovery
by: Norouzifar, Ali, et al.
Published: (2025)
by: Norouzifar, Ali, et al.
Published: (2025)
No AI Without PI! Object-Centric Process Mining as the Enabler for Generative, Predictive, and Prescriptive Artificial Intelligence
by: van der Aalst, Wil M. P.
Published: (2025)
by: van der Aalst, Wil M. P.
Published: (2025)
How to Write Beautiful Process-and-Data-Science Papers?
by: van der Aalst, Wil M. P.
Published: (2022)
by: van der Aalst, Wil M. P.
Published: (2022)
Measuring and Benchmarking Large Language Models' Capabilities to Generate Persuasive Language
by: Pauli, Amalie Brogaard, et al.
Published: (2024)
by: Pauli, Amalie Brogaard, et al.
Published: (2024)
All Eyes on the Workflow: Automated and Efficient Event Discovery from Video Streams
by: Pegoraro, Marco, et al.
Published: (2026)
by: Pegoraro, Marco, et al.
Published: (2026)
Quantifying Gender Biases Towards Politicians on Reddit
by: Marjanovic, Sara, et al.
Published: (2021)
by: Marjanovic, Sara, et al.
Published: (2021)
From Measurement Instruments to Data: Leveraging Theory-Driven Synthetic Training Data for Classifying Social Constructs
by: Birkenmaier, Lukas, et al.
Published: (2024)
by: Birkenmaier, Lukas, et al.
Published: (2024)
Show Me the Work: Fact-Checkers' Requirements for Explainable Automated Fact-Checking
by: Warren, Greta, et al.
Published: (2025)
by: Warren, Greta, et al.
Published: (2025)
Efficiency and Effectiveness of LLM-Based Summarization of Evidence in Crowdsourced Fact-Checking
by: Roitero, Kevin, et al.
Published: (2025)
by: Roitero, Kevin, et al.
Published: (2025)
LLM-based Semantic Augmentation for Harmful Content Detection
by: Meguellati, Elyas, et al.
Published: (2025)
by: Meguellati, Elyas, et al.
Published: (2025)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
by: Fröhling, Leon, et al.
Published: (2024)
by: Fröhling, Leon, et al.
Published: (2024)
Expanding Computation Spaces of LLMs at Inference Time
by: Jang, Yoonna, et al.
Published: (2025)
by: Jang, Yoonna, et al.
Published: (2025)
Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor Discussions
by: Kaffee, Lucie-Aimée, et al.
Published: (2023)
by: Kaffee, Lucie-Aimée, et al.
Published: (2023)
A Multilingual Similarity Dataset for News Article Frame
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Not What, But How: A Communicative Audit of LLM Response Framing
by: Pawar, Siddhesh Milind, et al.
Published: (2026)
by: Pawar, Siddhesh Milind, et al.
Published: (2026)
Modeling Public Perceptions of Science in Media
by: Pei, Jiaxin, et al.
Published: (2025)
by: Pei, Jiaxin, et al.
Published: (2025)
Counterfactual Edits for Generative Evaluation
by: Lymperaiou, Maria, et al.
Published: (2023)
by: Lymperaiou, Maria, et al.
Published: (2023)
Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection
by: Hakimi, Ahmad Dawar, et al.
Published: (2026)
by: Hakimi, Ahmad Dawar, et al.
Published: (2026)
Probing Pre-Trained Language Models for Cross-Cultural Differences in Values
by: Arora, Arnav, et al.
Published: (2022)
by: Arora, Arnav, et al.
Published: (2022)
Mind the Style Gap: Meta-Evaluation of Style and Attribute Transfer Metrics
by: Pauli, Amalie Brogaard, et al.
Published: (2025)
by: Pauli, Amalie Brogaard, et al.
Published: (2025)
Presumed Cultural Identity: How Names Shape LLM Responses
by: Pawar, Siddhesh, et al.
Published: (2025)
by: Pawar, Siddhesh, et al.
Published: (2025)
Counterfactual LLM-based Framework for Measuring Rhetorical Style
by: Qiu, Jingyi, et al.
Published: (2025)
by: Qiu, Jingyi, et al.
Published: (2025)
OCPQ: Object-Centric Process Querying & Constraints
by: Küsters, Aaron, et al.
Published: (2025)
by: Küsters, Aaron, et al.
Published: (2025)
Similar Items
-
The Unseen Targets of Hate -- A Systematic Review of Hateful Communication Datasets
by: Yu, Zehui, et al.
Published: (2024) -
Robustness and Confounders in the Demographic Alignment of LLMs with Human Perceptions of Offensiveness
by: Alipour, Shayan, et al.
Published: (2024) -
PM-LLM-Benchmark: Evaluating Large Language Models on Process Mining Tasks
by: Berti, Alessandro, et al.
Published: (2024) -
Translating Workflow Nets into the Partially Ordered Workflow Language
by: Kourani, Humam, et al.
Published: (2025) -
Evaluation of Study Plans using Partial Orders
by: Rennert, Christian, et al.
Published: (2024)