ToxiLab: How Well Do Open-Source LLMs Generate Synthetic Toxicity Data?
Fuente:
arXiv
Saved in:
| Main Authors: | Hui, Zheng, Guo, Zhaoxiao, Zhao, Hang, Duan, Juanyong, Ai, Lin, Li, Yinheng, Hirschberg, Julia, Huang, Congrui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ToxiCraft: A Novel Framework for Synthetic Generation of Harmful Information
by: Hui, Zheng, et al.
Published: (2024)
by: Hui, Zheng, et al.
Published: (2024)
QASE Enhanced PLMs: Improved Control in Text Generation for MRC
by: Ai, Lin, et al.
Published: (2024)
by: Ai, Lin, et al.
Published: (2024)
Enhancing Pre-Trained Generative Language Models with Question Attended Span Extraction on Machine Reading Comprehension
by: Ai, Lin, et al.
Published: (2024)
by: Ai, Lin, et al.
Published: (2024)
Detecting Mental Manipulation in Speech via Synthetic Multi-Speaker Dialogue
by: Chen, Run, et al.
Published: (2026)
by: Chen, Run, et al.
Published: (2026)
ToxiGAN: Toxic Data Augmentation via LLM-Guided Directional Adversarial Generation
by: Li, Peiran, et al.
Published: (2026)
by: Li, Peiran, et al.
Published: (2026)
ToxiTrace: Gradient-Aligned Training for Explainable Chinese Toxicity Detection
by: Li, Boyang, et al.
Published: (2026)
by: Li, Boyang, et al.
Published: (2026)
Data Processing Techniques for Modern Multimodal Models
by: Li, Yinheng, et al.
Published: (2024)
by: Li, Yinheng, et al.
Published: (2024)
A Review of Incorporating Psychological Theories in LLMs
by: Liu, Zizhou, et al.
Published: (2025)
by: Liu, Zizhou, et al.
Published: (2025)
Beyond Silent Letters: Amplifying LLMs in Emotion Recognition with Vocal Nuances
by: Wu, Zehui, et al.
Published: (2024)
by: Wu, Zehui, et al.
Published: (2024)
ToxiShield: Promoting Inclusive Developer Communication through Real-Time Toxicity Filtering
by: Anindya, MD Awsaf Alam, et al.
Published: (2026)
by: Anindya, MD Awsaf Alam, et al.
Published: (2026)
An abundance-type result for the tangent bundles of smooth Fano varieties
by: Wang, Juanyong
Published: (2024)
by: Wang, Juanyong
Published: (2024)
What Makes A Video Radicalizing? Identifying Sources of Influence in QAnon Videos
by: Ai, Lin, et al.
Published: (2024)
by: Ai, Lin, et al.
Published: (2024)
STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection
by: Bai, Zewen, et al.
Published: (2025)
by: Bai, Zewen, et al.
Published: (2025)
ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
by: Delaval, Axel, et al.
Published: (2025)
by: Delaval, Axel, et al.
Published: (2025)
Open Source in Lab Management
by: Cohen-Adad, Julien
Published: (2024)
by: Cohen-Adad, Julien
Published: (2024)
SYNTHEMPATHY: A Scalable Empathy Corpus Generated Using LLMs Without Any Crowdsourcing
by: Chen, Run, et al.
Published: (2025)
by: Chen, Run, et al.
Published: (2025)
How Well Do LLMs Understand Tunisian Arabic?
by: Mahdi, Mohamed
Published: (2025)
by: Mahdi, Mohamed
Published: (2025)
An Interactive Paradigm for Deep Research
by: Ai, Lin, et al.
Published: (2026)
by: Ai, Lin, et al.
Published: (2026)
Adaptable Non-parametric Approach for Speech-based Symptom Assessment: Isolating Private Medical Data in a Retrieval Datastore
by: Chen, Yu-Wen, et al.
Published: (2025)
by: Chen, Yu-Wen, et al.
Published: (2025)
Why Do Open-Source LLMs Struggle with Data Analysis? A Systematic Empirical Study
by: Zhu, Yuqi, et al.
Published: (2025)
by: Zhu, Yuqi, et al.
Published: (2025)
Read to Hear: A Zero-Shot Pronunciation Assessment Using Textual Descriptions and LLMs
by: Chen, Yu-Wen, et al.
Published: (2025)
by: Chen, Yu-Wen, et al.
Published: (2025)
MultiPA: A Multi-task Speech Pronunciation Assessment Model for Open Response Scenarios
by: Chen, Yu-Wen, et al.
Published: (2023)
by: Chen, Yu-Wen, et al.
Published: (2023)
How Well Do LLMs Identify Cultural Unity in Diversity?
by: Li, Jialin, et al.
Published: (2024)
by: Li, Jialin, et al.
Published: (2024)
How Well Do LLMs Imitate Human Writing Style?
by: Jemama, Rebira, et al.
Published: (2025)
by: Jemama, Rebira, et al.
Published: (2025)
Toxicity of the Commons: Curating Open-Source Pre-Training Data
by: Arnett, Catherine, et al.
Published: (2024)
by: Arnett, Catherine, et al.
Published: (2024)
ExaRanker-Open: Synthetic Explanation for IR using Open-Source LLMs
by: Ferraretto, Fernando, et al.
Published: (2024)
by: Ferraretto, Fernando, et al.
Published: (2024)
VoxRAG: A Step Toward Transcription-Free RAG Systems in Spoken Question Answering
by: Rackauckas, Zackary, et al.
Published: (2025)
by: Rackauckas, Zackary, et al.
Published: (2025)
Comparative Evaluation of Expressive Japanese Character Text-to-Speech with VITS and Style-BERT-VITS2
by: Rackauckas, Zackary, et al.
Published: (2025)
by: Rackauckas, Zackary, et al.
Published: (2025)
Animating Language Practice: Engagement with Stylized Conversational Agents in Japanese Learning
by: Rackauckas, Zackary, et al.
Published: (2025)
by: Rackauckas, Zackary, et al.
Published: (2025)
"Estado actual de las Toxi-infecciones alimentarias"
by: Jose GIL SANCHEZ
Published: (2009)
by: Jose GIL SANCHEZ
Published: (2009)
Strictly nef divisors on singular threefolds
by: Wang, Juanyong, et al.
Published: (2021)
by: Wang, Juanyong, et al.
Published: (2021)
Detecting Localized Deepfakes: How Well Do Synthetic Image Detectors Handle Inpainting?
by: Pandolfini, Serafino, et al.
Published: (2025)
by: Pandolfini, Serafino, et al.
Published: (2025)
SAGDA: Open-Source Synthetic Agriculture Data for Africa
by: Belgaid, Abdelghani, et al.
Published: (2025)
by: Belgaid, Abdelghani, et al.
Published: (2025)
PropaInsight: Toward Deeper Understanding of Propaganda in Terms of Techniques, Appeals, and Intent
by: Liu, Jiateng, et al.
Published: (2024)
by: Liu, Jiateng, et al.
Published: (2024)
Large Language Models in Finance: A Survey
by: Li, Yinheng, et al.
Published: (2023)
by: Li, Yinheng, et al.
Published: (2023)
Chatterbox TTS: Open Source vs. ElevenLabs
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
OpenDataLab: Empowering General Artificial Intelligence with Open Datasets
by: He, Conghui, et al.
Published: (2024)
by: He, Conghui, et al.
Published: (2024)
How Well Do LLMs Represent Values Across Cultures? Empirical Analysis of LLM Responses Based on Hofstede Cultural Dimensions
by: Kharchenko, Julia, et al.
Published: (2024)
by: Kharchenko, Julia, et al.
Published: (2024)
SPIN-Bench: How Well Do LLMs Plan Strategically and Reason Socially?
by: Yao, Jianzhu, et al.
Published: (2025)
by: Yao, Jianzhu, et al.
Published: (2025)
RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?
by: Dai, Yuyang, et al.
Published: (2026)
by: Dai, Yuyang, et al.
Published: (2026)
Similar Items
-
ToxiCraft: A Novel Framework for Synthetic Generation of Harmful Information
by: Hui, Zheng, et al.
Published: (2024) -
QASE Enhanced PLMs: Improved Control in Text Generation for MRC
by: Ai, Lin, et al.
Published: (2024) -
Enhancing Pre-Trained Generative Language Models with Question Attended Span Extraction on Machine Reading Comprehension
by: Ai, Lin, et al.
Published: (2024) -
Detecting Mental Manipulation in Speech via Synthetic Multi-Speaker Dialogue
by: Chen, Run, et al.
Published: (2026) -
ToxiGAN: Toxic Data Augmentation via LLM-Guided Directional Adversarial Generation
by: Li, Peiran, et al.
Published: (2026)