STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Zewen, Yin, Shengdi, Lu, Junyu, Zeng, Jingjie, Zhu, Haohao, Sun, Yuanyuan, Yang, Liang, Lin, Hongfei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fine-Grained Chinese Hate Speech Understanding: Span-Level Resources, Coded Term Lexicon, and Enhanced Detection Frameworks
von: Bai, Zewen, et al.
Veröffentlicht: (2025)
von: Bai, Zewen, et al.
Veröffentlicht: (2025)
ToxiTrace: Gradient-Aligned Training for Explainable Chinese Toxicity Detection
von: Li, Boyang, et al.
Veröffentlicht: (2026)
von: Li, Boyang, et al.
Veröffentlicht: (2026)
ToxiCloakCN: Evaluating Robustness of Offensive Language Detection in Chinese with Cloaking Perturbations
von: Xiao, Yunze, et al.
Veröffentlicht: (2024)
von: Xiao, Yunze, et al.
Veröffentlicht: (2024)
Commonality and Individuality! Integrating Humor Commonality with Speaker Individuality for Humor Recognition
von: Zhu, Haohao, et al.
Veröffentlicht: (2025)
von: Zhu, Haohao, et al.
Veröffentlicht: (2025)
Hate Speech Detection with Generalizable Target-aware Fairness
von: Chen, Tong, et al.
Veröffentlicht: (2024)
von: Chen, Tong, et al.
Veröffentlicht: (2024)
Towards Comprehensive Detection of Chinese Harmful Memes
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
von: Delaval, Axel, et al.
Veröffentlicht: (2025)
von: Delaval, Axel, et al.
Veröffentlicht: (2025)
Integrating Multi-view Analysis: Multi-view Mixture-of-Expert for Textual Personality Detection
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
Enhancing Textual Personality Detection toward Social Media: Integrating Long-term and Short-term Perspectives
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
Distinguishing Right from Wrong in Debates: Attribution Analysis of Chinese Harmful Memes
von: Wang, Weiming, et al.
Veröffentlicht: (2026)
von: Wang, Weiming, et al.
Veröffentlicht: (2026)
A Target-Aware Analysis of Data Augmentation for Hate Speech Detection
von: Casula, Camilla, et al.
Veröffentlicht: (2024)
von: Casula, Camilla, et al.
Veröffentlicht: (2024)
Seeing Hate Differently: Hate Subspace Modeling for Culture-Aware Hate Speech Detection
von: Cai, Weibin, et al.
Veröffentlicht: (2025)
von: Cai, Weibin, et al.
Veröffentlicht: (2025)
Radio Frequency Interference Detection Using Swin Transformer Embedding U 2 ‐Net
von: Shengdi Chen, et al.
Veröffentlicht: (2025)
von: Shengdi Chen, et al.
Veröffentlicht: (2025)
MasonPerplexity at Multimodal Hate Speech Event Detection 2024: Hate Speech and Target Detection Using Transformer Ensembles
von: Ganguly, Amrita, et al.
Veröffentlicht: (2024)
von: Ganguly, Amrita, et al.
Veröffentlicht: (2024)
The Straight and Narrow: Do LLMs Possess an Internal Moral Path?
von: Hu, Luoming, et al.
Veröffentlicht: (2026)
von: Hu, Luoming, et al.
Veröffentlicht: (2026)
ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances
von: Do, Huy Ba, et al.
Veröffentlicht: (2025)
von: Do, Huy Ba, et al.
Veröffentlicht: (2025)
ToxiGAN: Toxic Data Augmentation via LLM-Guided Directional Adversarial Generation
von: Li, Peiran, et al.
Veröffentlicht: (2026)
von: Li, Peiran, et al.
Veröffentlicht: (2026)
ToxiShield: Promoting Inclusive Developer Communication through Real-Time Toxicity Filtering
von: Anindya, MD Awsaf Alam, et al.
Veröffentlicht: (2026)
von: Anindya, MD Awsaf Alam, et al.
Veröffentlicht: (2026)
Toxic Synergy Between Hate Speech and Fake News Exposure
von: Kim, Munjung, et al.
Veröffentlicht: (2024)
von: Kim, Munjung, et al.
Veröffentlicht: (2024)
Mapping the Italian Telegram Ecosystem: Communities, Toxicity, and Hate Speech
von: Alvisi, Lorenzo, et al.
Veröffentlicht: (2025)
von: Alvisi, Lorenzo, et al.
Veröffentlicht: (2025)
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
von: Shen, Xinyue, et al.
Veröffentlicht: (2025)
von: Shen, Xinyue, et al.
Veröffentlicht: (2025)
Harder to Defend: Towards Chinese Toxicity Attacks via Implicit Enhancement and Obfuscation Rewriting
von: Kang, Jingyi, et al.
Veröffentlicht: (2026)
von: Kang, Jingyi, et al.
Veröffentlicht: (2026)
ToxiLab: How Well Do Open-Source LLMs Generate Synthetic Toxicity Data?
von: Hui, Zheng, et al.
Veröffentlicht: (2024)
von: Hui, Zheng, et al.
Veröffentlicht: (2024)
ParsCN: A Persian Dataset for Counter-Narrative Generation to Combat Online Hate Speech
von: Fesaghandis, Zahra Safdari, et al.
Veröffentlicht: (2026)
von: Fesaghandis, Zahra Safdari, et al.
Veröffentlicht: (2026)
Take its Essence, Discard its Dross! Debiasing for Toxic Language Detection via Counterfactual Causal Effect
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
SynHate: Detecting Hate Speech in Synthetic Deepfake Audio
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads
von: Li, Guojing, et al.
Veröffentlicht: (2026)
von: Li, Guojing, et al.
Veröffentlicht: (2026)
Causality‐Guided Data Augmentation for Cross‐Platform Hate Speech Detection
von: Tianming Jiang, et al.
Veröffentlicht: (2026)
von: Tianming Jiang, et al.
Veröffentlicht: (2026)
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech
von: Yadav, Neemesh, et al.
Veröffentlicht: (2024)
von: Yadav, Neemesh, et al.
Veröffentlicht: (2024)
Something Just Like TRuST : Toxicity Recognition of Span and Target
von: Atil, Berk, et al.
Veröffentlicht: (2025)
von: Atil, Berk, et al.
Veröffentlicht: (2025)
RealTalk-CN: A Realistic Chinese Speech-Text Dialogue Benchmark With Cross-Modal Interaction Analysis
von: Wang, Enzhi, et al.
Veröffentlicht: (2025)
von: Wang, Enzhi, et al.
Veröffentlicht: (2025)
ToxiTwitch: Toward Emote-Aware Hybrid Moderation for Live Streaming Platforms
von: Ansari, Baktash, et al.
Veröffentlicht: (2026)
von: Ansari, Baktash, et al.
Veröffentlicht: (2026)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
von: Chan, Fai Leui, et al.
Veröffentlicht: (2024)
von: Chan, Fai Leui, et al.
Veröffentlicht: (2024)
Advancing Hate Speech Detection with Transformers: Insights from the MetaHate
von: Chapagain, Santosh, et al.
Veröffentlicht: (2025)
von: Chapagain, Santosh, et al.
Veröffentlicht: (2025)
MetaHate: A Dataset for Unifying Efforts on Hate Speech Detection
von: Piot, Paloma, et al.
Veröffentlicht: (2024)
von: Piot, Paloma, et al.
Veröffentlicht: (2024)
Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech Detection
von: Brook, Joshua Wolfe, et al.
Veröffentlicht: (2025)
von: Brook, Joshua Wolfe, et al.
Veröffentlicht: (2025)
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
von: Hasan, Md Arid, et al.
Veröffentlicht: (2025)
von: Hasan, Md Arid, et al.
Veröffentlicht: (2025)
Automatic Landmark Detection for Preoperative Planning of High Tibial Osteotomy Using Traditional Feature Extraction and Deep Learning Methods
von: Jiaqi Han, et al.
Veröffentlicht: (2024)
von: Jiaqi Han, et al.
Veröffentlicht: (2024)
ImpliHateVid: A Benchmark Dataset and Two-stage Contrastive Learning Framework for Implicit Hate Speech Detection in Videos
von: Rehman, Mohammad Zia Ur, et al.
Veröffentlicht: (2025)
von: Rehman, Mohammad Zia Ur, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Fine-Grained Chinese Hate Speech Understanding: Span-Level Resources, Coded Term Lexicon, and Enhanced Detection Frameworks
von: Bai, Zewen, et al.
Veröffentlicht: (2025) -
ToxiTrace: Gradient-Aligned Training for Explainable Chinese Toxicity Detection
von: Li, Boyang, et al.
Veröffentlicht: (2026) -
ToxiCloakCN: Evaluating Robustness of Offensive Language Detection in Chinese with Cloaking Perturbations
von: Xiao, Yunze, et al.
Veröffentlicht: (2024) -
Commonality and Individuality! Integrating Humor Commonality with Speaker Individuality for Humor Recognition
von: Zhu, Haohao, et al.
Veröffentlicht: (2025) -
Hate Speech Detection with Generalizable Target-aware Fairness
von: Chen, Tong, et al.
Veröffentlicht: (2024)