Evaluating Simple Debiasing Techniques in RoBERTa-based Hate Speech Detection Models
Fuente:
arXiv
Saved in:
| Main Authors: | Iftimie, Diana, Zinn, Erik |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multiclass Hate Speech Detection with RoBERTa-OTA: Integrating Transformer Attention and Graph Convolutional Networks
by: Abusaqer, Mahmoud, et al.
Published: (2026)
by: Abusaqer, Mahmoud, et al.
Published: (2026)
RoBERTurk: Adjusting RoBERTa for Turkish
by: Tas, Nuri
Published: (2024)
by: Tas, Nuri
Published: (2024)
Performance Evaluation of Emotion Classification in Japanese Using RoBERTa and DeBERTa
by: Takenaka, Yoichi
Published: (2025)
by: Takenaka, Yoichi
Published: (2025)
NER- RoBERTa: Fine-Tuning RoBERTa for Named Entity Recognition (NER) within low-resource languages
by: Abdullah, Abdulhady Abas, et al.
Published: (2024)
by: Abdullah, Abdulhady Abas, et al.
Published: (2024)
A RoBERTa-Based Functional Syntax Annotation Model for Chinese Texts
by: Xiaohui, Han, et al.
Published: (2025)
by: Xiaohui, Han, et al.
Published: (2025)
Multilingual Hope Speech Detection: A Comparative Study of Logistic Regression, mBERT, and XLM-RoBERTa with Active Learning
by: Abiola, T. O., et al.
Published: (2025)
by: Abiola, T. O., et al.
Published: (2025)
RoBERTa-BiLSTM: A Context-Aware Hybrid Model for Sentiment Analysis
by: Rahman, Md. Mostafizer, et al.
Published: (2024)
by: Rahman, Md. Mostafizer, et al.
Published: (2024)
Decoding Fake Narratives in Spreading Hateful Stories: A Dual-Head RoBERTa Model with Multi-Task Learning
by: Bhaskar, Yash, et al.
Published: (2025)
by: Bhaskar, Yash, et al.
Published: (2025)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
by: Hua, Wenjie, et al.
Published: (2025)
by: Hua, Wenjie, et al.
Published: (2025)
Data Quality Matters: Suicide Intention Detection on Social Media Posts Using RoBERTa-CNN
by: Lin, Emily, et al.
Published: (2024)
by: Lin, Emily, et al.
Published: (2024)
EfficientQA : a RoBERTa Based Phrase-Indexed Question-Answering System
by: Chaybouti, Sofian, et al.
Published: (2021)
by: Chaybouti, Sofian, et al.
Published: (2021)
TartuNLP @ SIGTYP 2024 Shared Task: Adapting XLM-RoBERTa for Ancient and Historical Languages
by: Dorkin, Aleksei, et al.
Published: (2024)
by: Dorkin, Aleksei, et al.
Published: (2024)
Does RoBERTa Perform Better than BERT in Continual Learning: An Attention Sink Perspective
by: Bai, Xueying, et al.
Published: (2024)
by: Bai, Xueying, et al.
Published: (2024)
"AGI" team at SHROOM-CAP: Data-Centric Approach to Multilingual Hallucination Detection using XLM-RoBERTa
by: Rathva, Harsh, et al.
Published: (2025)
by: Rathva, Harsh, et al.
Published: (2025)
SemEval-2024 Task 8: Weighted Layer Averaging RoBERTa for Black-Box Machine-Generated Text Detection
by: Datta, Ayan, et al.
Published: (2024)
by: Datta, Ayan, et al.
Published: (2024)
Bilingual Sexism Classification: Fine-Tuned XLM-RoBERTa and GPT-3.5 Few-Shot Learning
by: Azadi, AmirMohammad, et al.
Published: (2024)
by: Azadi, AmirMohammad, et al.
Published: (2024)
Why Antiwork: A RoBERTa-Based System for Work-Related Stress Identification and Leading Factor Analysis
by: Lu, Tao, et al.
Published: (2024)
by: Lu, Tao, et al.
Published: (2024)
NCL-BU at SemEval-2026 Task 3: Fine-tuning XLM-RoBERTa for Multilingual Dimensional Sentiment Regression
by: Wu, Tong, et al.
Published: (2026)
by: Wu, Tong, et al.
Published: (2026)
HateDebias: On the Diversity and Variability of Hate Speech Debiasing
by: Wu, Hongyan, et al.
Published: (2024)
by: Wu, Hongyan, et al.
Published: (2024)
Fine-tuning RoBERTa for CVE-to-CWE Classification: A 125M Parameter Model Competitive with LLMs
by: Mosievskiy, Nikita
Published: (2026)
by: Mosievskiy, Nikita
Published: (2026)
EDU-NER-2025: Named Entity Recognition in Urdu Educational Texts using XLM-RoBERTa with X (formerly Twitter)
by: Ullah, Fida, et al.
Published: (2025)
by: Ullah, Fida, et al.
Published: (2025)
The Role of Model Architecture and Scale in Predicting Molecular Properties: Insights from Fine-Tuning RoBERTa, BART, and LLaMA
by: Youngmin, Lee, et al.
Published: (2024)
by: Youngmin, Lee, et al.
Published: (2024)
Mast Kalandar at SemEval-2024 Task 8: On the Trail of Textual Origins: RoBERTa-BiLSTM Approach to Detect AI-Generated Text
by: Bafna, Jainit Sushil, et al.
Published: (2024)
by: Bafna, Jainit Sushil, et al.
Published: (2024)
SG-UniBuc-NLP at SemEval-2026 Task 6: Multi-Head RoBERTa with Chunking for Long-Context Evasion Detection
by: Stefan, Gabriel, et al.
Published: (2026)
by: Stefan, Gabriel, et al.
Published: (2026)
Antibody Foundational Model : Ab-RoBERTa
by: Huh, Eunna, et al.
Published: (2025)
by: Huh, Eunna, et al.
Published: (2025)
QuadAI at SemEval-2026 Task 3: Ensemble Learning of Hybrid RoBERTa and LLMs for Dimensional Aspect-Based Sentiment Analysis
by: de Vink, A. J. W., et al.
Published: (2026)
by: de Vink, A. J. W., et al.
Published: (2026)
MaLei at the PLABA Track of TREC 2024: RoBERTa for Term Replacement -- LLaMA3.1 and GPT-4o for Complete Abstract Adaptation
by: Ling, Zhidong, et al.
Published: (2024)
by: Ling, Zhidong, et al.
Published: (2024)
The Large Language Model GreekLegalRoBERTa
by: Saketos, Vasileios, et al.
Published: (2024)
by: Saketos, Vasileios, et al.
Published: (2024)
RoBERTa-Augmented Synthesis for Detecting Malicious API Requests
by: Aharon, Udi, et al.
Published: (2024)
by: Aharon, Udi, et al.
Published: (2024)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
by: Chan, Fai Leui, et al.
Published: (2024)
by: Chan, Fai Leui, et al.
Published: (2024)
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
NaijaHate: Evaluating Hate Speech Detection on Nigerian Twitter Using Representative Data
by: Tonneau, Manuel, et al.
Published: (2024)
by: Tonneau, Manuel, et al.
Published: (2024)
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
by: Alqahtani, Amal, et al.
Published: (2025)
by: Alqahtani, Amal, et al.
Published: (2025)
VLAI: A RoBERTa-Based Model for Automated Vulnerability Severity Classification
by: Bonhomme, Cédric, et al.
Published: (2025)
by: Bonhomme, Cédric, et al.
Published: (2025)
Disagreeing Rationales: Rethinking Classification and Explainability Evaluation in Hate Speech Detection
by: Muscato, Benedetta, et al.
Published: (2026)
by: Muscato, Benedetta, et al.
Published: (2026)
Multi3Hate: Multimodal, Multilingual, and Multicultural Hate Speech Detection with Vision-Language Models
by: Bui, Minh Duc, et al.
Published: (2024)
by: Bui, Minh Duc, et al.
Published: (2024)
MasonPerplexity at Multimodal Hate Speech Event Detection 2024: Hate Speech and Target Detection Using Transformer Ensembles
by: Ganguly, Amrita, et al.
Published: (2024)
by: Ganguly, Amrita, et al.
Published: (2024)
Decoding Hate: Exploring Language Models' Reactions to Hate Speech
by: Piot, Paloma, et al.
Published: (2024)
by: Piot, Paloma, et al.
Published: (2024)
Detecting Anti-Semitic Hate Speech using Transformer-based Large Language Models
by: Liu, Dengyi, et al.
Published: (2024)
by: Liu, Dengyi, et al.
Published: (2024)
Compositional Generalisation for Explainable Hate Speech Detection
by: Calabrese, Agostina, et al.
Published: (2025)
by: Calabrese, Agostina, et al.
Published: (2025)
Similar Items
-
Multiclass Hate Speech Detection with RoBERTa-OTA: Integrating Transformer Attention and Graph Convolutional Networks
by: Abusaqer, Mahmoud, et al.
Published: (2026) -
RoBERTurk: Adjusting RoBERTa for Turkish
by: Tas, Nuri
Published: (2024) -
Performance Evaluation of Emotion Classification in Japanese Using RoBERTa and DeBERTa
by: Takenaka, Yoichi
Published: (2025) -
NER- RoBERTa: Fine-Tuning RoBERTa for Named Entity Recognition (NER) within low-resource languages
by: Abdullah, Abdulhady Abas, et al.
Published: (2024) -
A RoBERTa-Based Functional Syntax Annotation Model for Chinese Texts
by: Xiaohui, Han, et al.
Published: (2025)