D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and Evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Davani, Aida Mostafazadeh, Díaz, Mark, Baker, Dylan, Prabhakaran, Vinodkumar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GeniL: A Multilingual Dataset on Generalizing Language
por: Davani, Aida Mostafazadeh, et al.
Publicado: (2024)
por: Davani, Aida Mostafazadeh, et al.
Publicado: (2024)
GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives
por: Prabhakaran, Vinodkumar, et al.
Publicado: (2023)
por: Prabhakaran, Vinodkumar, et al.
Publicado: (2023)
Risks of Cultural Erasure in Large Language Models
por: Qadri, Rida, et al.
Publicado: (2025)
por: Qadri, Rida, et al.
Publicado: (2025)
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
por: Davani, Aida, et al.
Publicado: (2025)
por: Davani, Aida, et al.
Publicado: (2025)
Humanlike AI Design Increases Anthropomorphism but Yields Divergent Outcomes on Engagement and Trust Globally
por: Schimmelpfennig, Robin, et al.
Publicado: (2025)
por: Schimmelpfennig, Robin, et al.
Publicado: (2025)
Insights on Disagreement Patterns in Multimodal Safety Perception across Diverse Rater Groups
por: Rastogi, Charvi, et al.
Publicado: (2024)
por: Rastogi, Charvi, et al.
Publicado: (2024)
Taxonomy of User Needs and Actions
por: Shelby, Renee, et al.
Publicado: (2025)
por: Shelby, Renee, et al.
Publicado: (2025)
SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes
por: Bhutani, Mukul, et al.
Publicado: (2024)
por: Bhutani, Mukul, et al.
Publicado: (2024)
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
por: Weerasooriya, Tharindu Cyril, et al.
Publicado: (2023)
por: Weerasooriya, Tharindu Cyril, et al.
Publicado: (2023)
Towards Geo-Culturally Grounded LLM Generations
por: Lertvittayakumjorn, Piyawat, et al.
Publicado: (2025)
por: Lertvittayakumjorn, Piyawat, et al.
Publicado: (2025)
Is LLM an Overconfident Judge? Unveiling the Capabilities of LLMs in Detecting Offensive Language with Annotation Disagreement
por: Lu, Junyu, et al.
Publicado: (2025)
por: Lu, Junyu, et al.
Publicado: (2025)
Cultural Authenticity: Comparing LLM Cultural Representations to Native Human Expectations
por: van Liemt, Erin MacMurray, et al.
Publicado: (2026)
por: van Liemt, Erin MacMurray, et al.
Publicado: (2026)
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
por: Jin, Jiho, et al.
Publicado: (2026)
por: Jin, Jiho, et al.
Publicado: (2026)
Cultural Compass: Predicting Transfer Learning Success in Offensive Language Detection with Cultural Features
por: Zhou, Li, et al.
Publicado: (2023)
por: Zhou, Li, et al.
Publicado: (2023)
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
por: Cheng, Myra, et al.
Publicado: (2026)
por: Cheng, Myra, et al.
Publicado: (2026)
A Unified Framework to Quantify Cultural Intelligence of AI
por: Dev, Sunipa, et al.
Publicado: (2026)
por: Dev, Sunipa, et al.
Publicado: (2026)
Guardians of Discourse: Evaluating LLMs on Multilingual Offensive Language Detection
por: He, Jianfei, et al.
Publicado: (2024)
por: He, Jianfei, et al.
Publicado: (2024)
Language, Culture, and Ideology: Personalizing Offensiveness Detection in Political Tweets with Reasoning LLMs
por: Pihulski, Dzmitry, et al.
Publicado: (2025)
por: Pihulski, Dzmitry, et al.
Publicado: (2025)
Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models
por: Rastogi, Charvi, et al.
Publicado: (2025)
por: Rastogi, Charvi, et al.
Publicado: (2025)
SAFARI: A Community-Engaged Approach and Dataset of Stereotype Resources in the Sub-Saharan African Context
por: Verma, Aishwarya, et al.
Publicado: (2026)
por: Verma, Aishwarya, et al.
Publicado: (2026)
ToxiCloakCN: Evaluating Robustness of Offensive Language Detection in Chinese with Cloaking Perturbations
por: Xiao, Yunze, et al.
Publicado: (2024)
por: Xiao, Yunze, et al.
Publicado: (2024)
Detection and Analysis of Offensive Online Content in Hausa Language
por: Adam, Fatima Muhammad, et al.
Publicado: (2023)
por: Adam, Fatima Muhammad, et al.
Publicado: (2023)
Mind the Gesture: Evaluating AI Sensitivity to Culturally Offensive Non-Verbal Gestures
por: Yerukola, Akhila, et al.
Publicado: (2025)
por: Yerukola, Akhila, et al.
Publicado: (2025)
OffensiveLang: A Community Based Implicit Offensive Language Dataset
por: Das, Amit, et al.
Publicado: (2024)
por: Das, Amit, et al.
Publicado: (2024)
Investigating the Impact of Semi-Supervised Methods with Data Augmentation on Offensive Language Detection in Romanian Language
por: Nicola, Elena-Beatrice, et al.
Publicado: (2024)
por: Nicola, Elena-Beatrice, et al.
Publicado: (2024)
Enhancing Romanian Offensive Language Detection through Knowledge Distillation, Multi-Task Learning, and Data Augmentation
por: Matei, Vlad-Cristian, et al.
Publicado: (2024)
por: Matei, Vlad-Cristian, et al.
Publicado: (2024)
Offensive Language Detection on Social Media Using XLNet
por: Alothman, Reem, et al.
Publicado: (2025)
por: Alothman, Reem, et al.
Publicado: (2025)
Multi-task Learning with Active Learning for Arabic Offensive Speech Detection
por: Alansari, Aisha, et al.
Publicado: (2025)
por: Alansari, Aisha, et al.
Publicado: (2025)
From Disagreement to Understanding: The Case for Ambiguity Detection in NLI
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
Disentangling Language and Culture for Evaluating Multilingual Large Language Models
por: Ying, Jiahao, et al.
Publicado: (2025)
por: Ying, Jiahao, et al.
Publicado: (2025)
Subasa - Adapting Language Models for Low-resourced Offensive Language Detection in Sinhala
por: Haturusinghe, Shanilka, et al.
Publicado: (2025)
por: Haturusinghe, Shanilka, et al.
Publicado: (2025)
Lost in Pronunciation: Detecting Chinese Offensive Language Disguised by Phonetic Cloaking Replacement
por: Guo, Haotan, et al.
Publicado: (2025)
por: Guo, Haotan, et al.
Publicado: (2025)
"Just a strange pic": Evaluating 'safety' in GenAI Image safety annotation tasks from diverse annotators' perspectives
por: Wang, Ding, et al.
Publicado: (2025)
por: Wang, Ding, et al.
Publicado: (2025)
Beyond Consensus: Perspectivist Modeling and Evaluation of Annotator Disagreement in NLP
por: Xu, Yinuo, et al.
Publicado: (2026)
por: Xu, Yinuo, et al.
Publicado: (2026)
Chinese Offensive Language Detection:Current Status and Future Directions
por: Xiao, Yunze, et al.
Publicado: (2024)
por: Xiao, Yunze, et al.
Publicado: (2024)
Mixed Signals: Understanding Model Disagreement in Multimodal Empathy Detection
por: Srikanth, Maya, et al.
Publicado: (2025)
por: Srikanth, Maya, et al.
Publicado: (2025)
Leveraging Sentiment for Offensive Text Classification
por: Islam, Khondoker Ittehadul
Publicado: (2024)
por: Islam, Khondoker Ittehadul
Publicado: (2024)
Disagreement as Data: Reasoning Trace Analytics in Multi-Agent Systems
por: Tajik, Elham, et al.
Publicado: (2026)
por: Tajik, Elham, et al.
Publicado: (2026)
Training-Free Cultural Alignment of Large Language Models via Persona Disagreement
por: Kiet, Huynh Trung, et al.
Publicado: (2026)
por: Kiet, Huynh Trung, et al.
Publicado: (2026)
A Decomposition-Based Approach for Evaluating and Analyzing Inter-Annotator Disagreement
por: Levi, Effi, et al.
Publicado: (2022)
por: Levi, Effi, et al.
Publicado: (2022)
Ejemplares similares
-
GeniL: A Multilingual Dataset on Generalizing Language
por: Davani, Aida Mostafazadeh, et al.
Publicado: (2024) -
GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives
por: Prabhakaran, Vinodkumar, et al.
Publicado: (2023) -
Risks of Cultural Erasure in Large Language Models
por: Qadri, Rida, et al.
Publicado: (2025) -
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
por: Davani, Aida, et al.
Publicado: (2025) -
Humanlike AI Design Increases Anthropomorphism but Yields Divergent Outcomes on Engagement and Trust Globally
por: Schimmelpfennig, Robin, et al.
Publicado: (2025)