Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
Fuente:
arXiv
Saved in:
| Main Authors: | Weerasooriya, Tharindu Cyril, Dutta, Sujan, Ranasinghe, Tharindu, Zampieri, Marcos, Homan, Christopher M., KhudaBukhsh, Ashiqur R. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rater Cohesion and Quality from a Vicarious Perspective
by: Pandita, Deepak, et al.
Published: (2024)
by: Pandita, Deepak, et al.
Published: (2024)
ARTICLE: Annotator Reliability Through In-Context Learning
by: Dutta, Sujan, et al.
Published: (2024)
by: Dutta, Sujan, et al.
Published: (2024)
Subasa - Adapting Language Models for Low-resourced Offensive Language Detection in Sinhala
by: Haturusinghe, Shanilka, et al.
Published: (2025)
by: Haturusinghe, Shanilka, et al.
Published: (2025)
A Federated Learning Approach to Privacy Preserving Offensive Language Identification
by: Zampieri, Marcos, et al.
Published: (2024)
by: Zampieri, Marcos, et al.
Published: (2024)
Towards Generalized Offensive Language Identification
by: Dmonte, Alphaeus, et al.
Published: (2024)
by: Dmonte, Alphaeus, et al.
Published: (2024)
Down the Toxicity Rabbit Hole: A Novel Framework to Bias Audit Large Language Models
by: Dutta, Arka, et al.
Published: (2023)
by: Dutta, Arka, et al.
Published: (2023)
Gender Representation and Bias in Indian Civil Service Mock Interviews
by: Banerjee, Somonnoy, et al.
Published: (2024)
by: Banerjee, Somonnoy, et al.
Published: (2024)
SOLD: Sinhala Offensive Language Dataset
by: Ranasinghe, Tharindu, et al.
Published: (2022)
by: Ranasinghe, Tharindu, et al.
Published: (2022)
Hope vs. Hate: Understanding User Interactions with LGBTQ+ News Content in Mainstream US News Media through the Lens of Hope Speech
by: Pofcher, Jonathan, et al.
Published: (2025)
by: Pofcher, Jonathan, et al.
Published: (2025)
What About the Scene with the Hitler Reference? HAUNT: A Framework to Probe LLMs' Self-consistency Via Adversarial Nudge
by: Dutta, Arka, et al.
Published: (2025)
by: Dutta, Arka, et al.
Published: (2025)
ALEXSIS-PT: A New Resource for Portuguese Lexical Simplification
by: North, Kai, et al.
Published: (2022)
by: North, Kai, et al.
Published: (2022)
Community Needs and Assets: A Computational Analysis of Community Conversations
by: Chowdhury, Md Towhidul Absar, et al.
Published: (2024)
by: Chowdhury, Md Towhidul Absar, et al.
Published: (2024)
Investigating Vaccine Buyer's Remorse: Post-Vaccination Decision Regret in COVID-19 Social Media Using Politically Diverse Human Annotation
by: Stanley, Miles, et al.
Published: (2026)
by: Stanley, Miles, et al.
Published: (2026)
MultiLS: A Multi-task Lexical Simplification Framework
by: North, Kai, et al.
Published: (2024)
by: North, Kai, et al.
Published: (2024)
On the State of NLP Approaches to Modeling Depression in Social Media: A Post-COVID-19 Outlook
by: Bucur, Ana-Maria, et al.
Published: (2024)
by: Bucur, Ana-Maria, et al.
Published: (2024)
Datasets for Depression Modeling in Social Media: An Overview
by: Bucur, Ana-Maria, et al.
Published: (2025)
by: Bucur, Ana-Maria, et al.
Published: (2025)
Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM
by: Shetty, Samay U., et al.
Published: (2026)
by: Shetty, Samay U., et al.
Published: (2026)
Navigating the Rabbit Hole: Emergent Biases in LLM-Generated Attack Narratives Targeting Mental Health Groups
by: Magu, Rijul, et al.
Published: (2025)
by: Magu, Rijul, et al.
Published: (2025)
A Survey on Multilingual Mental Disorders Detection from Social Media Data
by: Bucur, Ana-Maria, et al.
Published: (2025)
by: Bucur, Ana-Maria, et al.
Published: (2025)
Infrastructure Ombudsman: Mining Future Failure Concerns from Structural Disaster Response
by: Chowdhury, Md Towhidul Absar, et al.
Published: (2024)
by: Chowdhury, Md Towhidul Absar, et al.
Published: (2024)
A Survey of Multimodal Sarcasm Detection
by: Farabi, Shafkat, et al.
Published: (2024)
by: Farabi, Shafkat, et al.
Published: (2024)
When Neutral Summaries are not that Neutral: Quantifying Political Neutrality in LLM-Generated News Summaries
by: Vijay, Supriti, et al.
Published: (2024)
by: Vijay, Supriti, et al.
Published: (2024)
Exploring the Performance of Large Language Models on Subjective Span Identification Tasks
by: Dmonte, Alphaeus, et al.
Published: (2026)
by: Dmonte, Alphaeus, et al.
Published: (2026)
LPI-RIT at LeWiDi-2025: Improving Distributional Predictions via Metadata and Loss Reweighting with DisCo
by: Sawkar, Mandira, et al.
Published: (2025)
by: Sawkar, Mandira, et al.
Published: (2025)
A Neuro-Symbolic Multi-Agent Approach to Legal-Cybersecurity Knowledge Integration
by: Bonfanti, Chiara, et al.
Published: (2025)
by: Bonfanti, Chiara, et al.
Published: (2025)
MUNIChus: Multilingual News Image Captioning Benchmark
by: Chen, Yuji, et al.
Published: (2026)
by: Chen, Yuji, et al.
Published: (2026)
Text Generation Models for Luxembourgish with Limited Data: A Balanced Multilingual Strategy
by: Plum, Alistair, et al.
Published: (2024)
by: Plum, Alistair, et al.
Published: (2024)
CSEPrompts: A Benchmark of Introductory Computer Science Prompts
by: Raihan, Nishat, et al.
Published: (2024)
by: Raihan, Nishat, et al.
Published: (2024)
Do LLMs Judge Distantly Supervised Named Entity Labels Well? Constructing the JudgeWEL Dataset
by: Plum, Alistair, et al.
Published: (2026)
by: Plum, Alistair, et al.
Published: (2026)
Guided Distant Supervision for Multilingual Relation Extraction Data: Adapting to a New Language
by: Plum, Alistair, et al.
Published: (2024)
by: Plum, Alistair, et al.
Published: (2024)
LLM-based Embedders for Prior Case Retrieval
by: Premasiri, Damith, et al.
Published: (2025)
by: Premasiri, Damith, et al.
Published: (2025)
D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and Evaluation
by: Davani, Aida Mostafazadeh, et al.
Published: (2024)
by: Davani, Aida Mostafazadeh, et al.
Published: (2024)
MultiGA: Leveraging Multi-Source Seeding in Genetic Algorithms
by: Ng, Isabelle Diana May-Xin, et al.
Published: (2025)
by: Ng, Isabelle Diana May-Xin, et al.
Published: (2025)
OffensiveLang: A Community Based Implicit Offensive Language Dataset
by: Das, Amit, et al.
Published: (2024)
by: Das, Amit, et al.
Published: (2024)
ProRefine: Inference-Time Prompt Refinement with Textual Feedback
by: Pandita, Deepak, et al.
Published: (2025)
by: Pandita, Deepak, et al.
Published: (2025)
Is LLM an Overconfident Judge? Unveiling the Capabilities of LLMs in Detecting Offensive Language with Annotation Disagreement
by: Lu, Junyu, et al.
Published: (2025)
by: Lu, Junyu, et al.
Published: (2025)
Statistical inference for a multiscale stochastic model of enzyme kinetics via propagation of chaos
by: Ganguly, Arnab, et al.
Published: (2024)
by: Ganguly, Arnab, et al.
Published: (2024)
Asymptotic Analysis of the Total Quasi-Steady State Approximation for the Michaelis--Menten Enzyme Kinetic Reactions
by: Ganguly, Arnab, et al.
Published: (2025)
by: Ganguly, Arnab, et al.
Published: (2025)
Local times of self-intersection and sample path properties of Volterra Gaussian processes
by: Izyumtseva, Olga, et al.
Published: (2024)
by: Izyumtseva, Olga, et al.
Published: (2024)
Mixing time for an epidemic model on graphs with external sources of infection
by: KhudaBukhsh, Wasiur R., et al.
Published: (2025)
by: KhudaBukhsh, Wasiur R., et al.
Published: (2025)
Similar Items
-
Rater Cohesion and Quality from a Vicarious Perspective
by: Pandita, Deepak, et al.
Published: (2024) -
ARTICLE: Annotator Reliability Through In-Context Learning
by: Dutta, Sujan, et al.
Published: (2024) -
Subasa - Adapting Language Models for Low-resourced Offensive Language Detection in Sinhala
by: Haturusinghe, Shanilka, et al.
Published: (2025) -
A Federated Learning Approach to Privacy Preserving Offensive Language Identification
by: Zampieri, Marcos, et al.
Published: (2024) -
Towards Generalized Offensive Language Identification
by: Dmonte, Alphaeus, et al.
Published: (2024)