ToxSyn: Reducing Bias in Hate Speech Detection via Synthetic Minority Data in Brazilian Portuguese
Fuente:
arXiv
Saved in:
| Main Authors: | Brito, Iago Alves, Dollis, Julia Soares, Färber, Fernanda Bufon, Silva, Diogo Fernandes Costa, Filho, Arlindo Rodrigues Galvão |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MedPT: A Massive Medical Question Answering Dataset for Brazilian-Portuguese Speakers
by: Färber, Fernanda Bufon, et al.
Published: (2025)
by: Färber, Fernanda Bufon, et al.
Published: (2025)
Integrating Personality into Digital Humans: A Review of LLM-Driven Approaches for Virtual Reality
by: Brito, Iago Alves, et al.
Published: (2025)
by: Brito, Iago Alves, et al.
Published: (2025)
Safety Is Not Universal: The Selective Safety Trap in LLM Alignment
by: Brito, Iago Alves, et al.
Published: (2026)
by: Brito, Iago Alves, et al.
Published: (2026)
When Avatars Have Personality: Effects on Engagement and Communication in Immersive Medical Training
by: Dollis, Julia S., et al.
Published: (2025)
by: Dollis, Julia S., et al.
Published: (2025)
SynHate: Detecting Hate Speech in Synthetic Deepfake Audio
by: Ranjan, Rishabh, et al.
Published: (2025)
by: Ranjan, Rishabh, et al.
Published: (2025)
Semi-automated Fact-checking in Portuguese: Corpora Enrichment using Retrieval with Claim extraction
by: Gomes, Juliana Resplande Sant'anna, et al.
Published: (2025)
by: Gomes, Juliana Resplande Sant'anna, et al.
Published: (2025)
Analyzing Bias in False Refusal Behavior of Large Language Models for Hate Speech Detoxification
by: Im, Kyuri, et al.
Published: (2026)
by: Im, Kyuri, et al.
Published: (2026)
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech
by: Yadav, Neemesh, et al.
Published: (2024)
by: Yadav, Neemesh, et al.
Published: (2024)
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
Attention Guidance through Video Script: A Case Study of Object Focusing on 360° VR Video Tours
by: Silva, Paulo Vitor Santana, et al.
Published: (2026)
by: Silva, Paulo Vitor Santana, et al.
Published: (2026)
Tagarela - A Portuguese speech dataset from podcasts
by: de Oliveira, Frederico Santos, et al.
Published: (2026)
by: de Oliveira, Frederico Santos, et al.
Published: (2026)
LLM in the Loop: Creating the ParaDeHate Dataset for Hate Speech Detoxification
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
Brazilian Portuguese version of the Spence Children’s Anxiety Scale (SCAS-Brasil)
by: Diogo A. DeSousa,
Published: (2012)
by: Diogo A. DeSousa,
Published: (2012)
Estimativa do tempo de vida útil de represa de pequeno porte
by: André Gustavo Mazzini Bufon
Published: (2009)
by: André Gustavo Mazzini Bufon
Published: (2009)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
by: Das, Amit, et al.
Published: (2024)
by: Das, Amit, et al.
Published: (2024)
Hate Speech
by: Guillén-Nieto, Victoria
Published: (2023)
by: Guillén-Nieto, Victoria
Published: (2023)
CAPACITAÇÃO DE TÉCNICOS DE SAÚDE: UMA EXPERIÊNCIA PIONEIRA NO ESTADO DO TOCANTINS, BRASIL
by: Arlindo Serpa Filho
Published: (2010)
by: Arlindo Serpa Filho
Published: (2010)
ALBA: A European Portuguese Benchmark for Evaluating Language and Linguistic Dimensions in Generative LLMs
by: Vieira, Inês, et al.
Published: (2026)
by: Vieira, Inês, et al.
Published: (2026)
Enhancing Speech Emotion Recognition with Graph-Based Multimodal Fusion and Prosodic Features for the Speech Emotion Recognition in Naturalistic Conditions Challenge at Interspeech 2025
by: Ferreira, Alef Iury Siqueira, et al.
Published: (2025)
by: Ferreira, Alef Iury Siqueira, et al.
Published: (2025)
From Languages to Geographies: Towards Evaluating Cultural Bias in Hate Speech Datasets
by: Tonneau, Manuel, et al.
Published: (2024)
by: Tonneau, Manuel, et al.
Published: (2024)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
by: Ahmed, Ahmed Haj, et al.
Published: (2024)
Antibacterial activity and chemical composition of essential oil of Lippia microphylla Cham
by: Fabiola Fernandes Galvão Rodrigues
Published: (2011)
by: Fabiola Fernandes Galvão Rodrigues
Published: (2011)
Digitale Hate Speech
Published: (2023)
Published: (2023)
Hate Speech Law
by: Brown, Alex
Published: (2025)
by: Brown, Alex
Published: (2025)
Possible Readings of 'Simplesmente (simply)' in Brazilian Portuguese
by: Lopes Rodrigues, Adriano, et al.
Published: (2025)
by: Lopes Rodrigues, Adriano, et al.
Published: (2025)
Smart ETL and LLM-based contents classification: the European Smart Tourism Tools Observatory experience
by: Cosme, Diogo, et al.
Published: (2024)
by: Cosme, Diogo, et al.
Published: (2024)
HateDebias: On the Diversity and Variability of Hate Speech Debiasing
by: Wu, Hongyan, et al.
Published: (2024)
by: Wu, Hongyan, et al.
Published: (2024)
Confident, Calibrated, or Complicit: Safety Alignment and Ideological Bias in LLM Hate Speech Detection
by: Selvaganapathy, Sanjeeevan, et al.
Published: (2025)
by: Selvaganapathy, Sanjeeevan, et al.
Published: (2025)
Cross-cultural adaptation of the Cyberchondria Severity Scale for Brazilian Portuguese
by: Fernanda Gonçalves da Silva
Published: (2016)
by: Fernanda Gonçalves da Silva
Published: (2016)
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
by: Shen, Xinyue, et al.
Published: (2025)
by: Shen, Xinyue, et al.
Published: (2025)
Seeing Hate Differently: Hate Subspace Modeling for Culture-Aware Hate Speech Detection
by: Cai, Weibin, et al.
Published: (2025)
by: Cai, Weibin, et al.
Published: (2025)
Self-Explaining Hate Speech Detection with Moral Rationales
by: Vargas, Francielle, et al.
Published: (2026)
by: Vargas, Francielle, et al.
Published: (2026)
Decoding Hate: Exploring Language Models' Reactions to Hate Speech
by: Piot, Paloma, et al.
Published: (2024)
by: Piot, Paloma, et al.
Published: (2024)
An Investigation of Hate Speech in Italian
Published: (2025)
Published: (2025)
Focus360: Guiding User Attention in Immersive Videos for VR
by: Silva, Paulo Vitor S., et al.
Published: (2026)
by: Silva, Paulo Vitor S., et al.
Published: (2026)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
by: Chan, Fai Leui, et al.
Published: (2024)
by: Chan, Fai Leui, et al.
Published: (2024)
Advancing Hate Speech Detection with Transformers: Insights from the MetaHate
by: Chapagain, Santosh, et al.
Published: (2025)
by: Chapagain, Santosh, et al.
Published: (2025)
Web(er) of Hate: A Survey on How Hate Speech Is Typed
by: Wang, Luna, et al.
Published: (2025)
by: Wang, Luna, et al.
Published: (2025)
MetaHate: A Dataset for Unifying Efforts on Hate Speech Detection
by: Piot, Paloma, et al.
Published: (2024)
by: Piot, Paloma, et al.
Published: (2024)
FairSSD: Understanding Bias in Synthetic Speech Detectors
by: Yadav, Amit Kumar Singh, et al.
Published: (2024)
by: Yadav, Amit Kumar Singh, et al.
Published: (2024)
Similar Items
-
MedPT: A Massive Medical Question Answering Dataset for Brazilian-Portuguese Speakers
by: Färber, Fernanda Bufon, et al.
Published: (2025) -
Integrating Personality into Digital Humans: A Review of LLM-Driven Approaches for Virtual Reality
by: Brito, Iago Alves, et al.
Published: (2025) -
Safety Is Not Universal: The Selective Safety Trap in LLM Alignment
by: Brito, Iago Alves, et al.
Published: (2026) -
When Avatars Have Personality: Effects on Engagement and Communication in Immersive Medical Training
by: Dollis, Julia S., et al.
Published: (2025) -
SynHate: Detecting Hate Speech in Synthetic Deepfake Audio
by: Ranjan, Rishabh, et al.
Published: (2025)