ylmmcl at Multilingual Text Detoxification 2025: Lexicon-Guided Detoxification and Classifier-Gated Rewriting
Fuente:
arXiv
Salvato in:
| Autori principali: | Lai-Lopez, Nicole, Wang, Lusha, Yuan, Su, Zhang, Liza |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Text Detoxification: Data Efficiency, Semantic Preservation and Model Generalization
di: Yu, Jing, et al.
Pubblicazione: (2025)
di: Yu, Jing, et al.
Pubblicazione: (2025)
DetoxLLM: A Framework for Detoxification with Explanations
di: Khondaker, Md Tawkat Islam, et al.
Pubblicazione: (2024)
di: Khondaker, Md Tawkat Islam, et al.
Pubblicazione: (2024)
Multilingual and Explainable Text Detoxification with Parallel Corpora
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
Test-Time Detoxification without Training or Learning Anything
di: Saglam, Baturay, et al.
Pubblicazione: (2026)
di: Saglam, Baturay, et al.
Pubblicazione: (2026)
DAPI: Domain Adaptive Toxicity Probe Vector Intervention for Fine-Grained Detoxification
di: Hyeonsu, Cho, et al.
Pubblicazione: (2025)
di: Hyeonsu, Cho, et al.
Pubblicazione: (2025)
UniDetox: Universal Detoxification of Large Language Models via Dataset Distillation
di: Lu, Huimin, et al.
Pubblicazione: (2025)
di: Lu, Huimin, et al.
Pubblicazione: (2025)
SmurfCat at PAN 2024 TextDetox: Alignment of Multilingual Transformers for Text Detoxification
di: Rykov, Elisei, et al.
Pubblicazione: (2024)
di: Rykov, Elisei, et al.
Pubblicazione: (2024)
Text Detoxification as Style Transfer in English and Hindi
di: Mukherjee, Sourabrata, et al.
Pubblicazione: (2024)
di: Mukherjee, Sourabrata, et al.
Pubblicazione: (2024)
GemDetox at TextDetox CLEF 2025: Enhancing a Massively Multilingual Model for Text Detoxification on Low-resource Languages
di: Dang, Trung Duc Anh, et al.
Pubblicazione: (2025)
di: Dang, Trung Duc Anh, et al.
Pubblicazione: (2025)
Cross-Lingual Transfer of Debiasing and Detoxification in Multilingual LLMs: An Extensive Investigation
di: Neplenbroek, Vera, et al.
Pubblicazione: (2024)
di: Neplenbroek, Vera, et al.
Pubblicazione: (2024)
IF-GUIDE: Influence Function-Guided Detoxification of LLMs
di: Coalson, Zachary, et al.
Pubblicazione: (2025)
di: Coalson, Zachary, et al.
Pubblicazione: (2025)
On the Robustness of Knowledge Editing for Detoxification
di: Dong, Ming, et al.
Pubblicazione: (2026)
di: Dong, Ming, et al.
Pubblicazione: (2026)
Detoxification for LLM: From Dataset Itself
di: Shao, Wei, et al.
Pubblicazione: (2026)
di: Shao, Wei, et al.
Pubblicazione: (2026)
GPT-DETOX: An In-Context Learning-Based Paraphraser for Text Detoxification
di: Pesaranghader, Ali, et al.
Pubblicazione: (2024)
di: Pesaranghader, Ali, et al.
Pubblicazione: (2024)
Evaluating Text Style Transfer: A Nine-Language Benchmark for Text Detoxification
di: Protasov, Vitaly, et al.
Pubblicazione: (2025)
di: Protasov, Vitaly, et al.
Pubblicazione: (2025)
RAZOR: Sharpening Knowledge by Cutting Bias with Unsupervised Text Rewriting
di: Yang, Shuo, et al.
Pubblicazione: (2024)
di: Yang, Shuo, et al.
Pubblicazione: (2024)
Parameter-Efficient Detoxification with Contrastive Decoding
di: Niu, Tong, et al.
Pubblicazione: (2024)
di: Niu, Tong, et al.
Pubblicazione: (2024)
Short-form Text Rewriting with Phi Silica
di: Tadimeti, Divya, et al.
Pubblicazione: (2026)
di: Tadimeti, Divya, et al.
Pubblicazione: (2026)
CMD: a framework for Context-aware Model self-Detoxification
di: Tang, Zecheng, et al.
Pubblicazione: (2023)
di: Tang, Zecheng, et al.
Pubblicazione: (2023)
Semi-Supervised Learning for Bilingual Lexicon Induction
di: Garnier, Paul, et al.
Pubblicazione: (2024)
di: Garnier, Paul, et al.
Pubblicazione: (2024)
A Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AI
di: Malinga, Melusi, et al.
Pubblicazione: (2024)
di: Malinga, Melusi, et al.
Pubblicazione: (2024)
Autonomous Data Selection with Zero-shot Generative Classifiers for Mathematical Texts
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
Transparent Neighborhood Approximation for Text Classifier Explanation
di: Cai, Yi, et al.
Pubblicazione: (2024)
di: Cai, Yi, et al.
Pubblicazione: (2024)
DSCD: Large Language Model Detoxification with Self-Constrained Decoding
di: Dong, Ming, et al.
Pubblicazione: (2025)
di: Dong, Ming, et al.
Pubblicazione: (2025)
Breaking Bad Tokens: Detoxification of LLMs Using Sparse Autoencoders
di: Goyal, Agam, et al.
Pubblicazione: (2025)
di: Goyal, Agam, et al.
Pubblicazione: (2025)
MultiParaDetox: Extending Text Detoxification with Parallel Data to New Languages
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers
di: Neshaei, Seyed Parsa, et al.
Pubblicazione: (2024)
di: Neshaei, Seyed Parsa, et al.
Pubblicazione: (2024)
A Discriminative Latent-Variable Model for Bilingual Lexicon Induction
di: Ruder, Sebastian, et al.
Pubblicazione: (2018)
di: Ruder, Sebastian, et al.
Pubblicazione: (2018)
Analyzing Bias in False Refusal Behavior of Large Language Models for Hate Speech Detoxification
di: Im, Kyuri, et al.
Pubblicazione: (2026)
di: Im, Kyuri, et al.
Pubblicazione: (2026)
DISCERN: Decoding Systematic Errors in Natural Language for Text Classifiers
di: Menon, Rakesh R., et al.
Pubblicazione: (2024)
di: Menon, Rakesh R., et al.
Pubblicazione: (2024)
LLM in the Loop: Creating the ParaDeHate Dataset for Hate Speech Detoxification
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
di: Cho, Nicole, et al.
Pubblicazione: (2025)
di: Cho, Nicole, et al.
Pubblicazione: (2025)
Fine-Grained Detoxification via Instance-Level Prefixes for Large Language Models
di: Yi, Xin, et al.
Pubblicazione: (2024)
di: Yi, Xin, et al.
Pubblicazione: (2024)
Is There No Such Thing as a Bad Question? H4R: HalluciBot For Ratiocination, Rewriting, Ranking, and Routing
di: Watson, William, et al.
Pubblicazione: (2024)
di: Watson, William, et al.
Pubblicazione: (2024)
Explaining Text Classifiers with Counterfactual Representations
di: Lemberger, Pirmin, et al.
Pubblicazione: (2024)
di: Lemberger, Pirmin, et al.
Pubblicazione: (2024)
Text Sentiment Analysis and Classification Based on Bidirectional Gated Recurrent Units (GRUs) Model
di: Xu, Wei, et al.
Pubblicazione: (2024)
di: Xu, Wei, et al.
Pubblicazione: (2024)
Adaptive Detoxification: Safeguarding General Capabilities of LLMs through Toxicity-Aware Knowledge Editing
di: Lu, Yifan, et al.
Pubblicazione: (2025)
di: Lu, Yifan, et al.
Pubblicazione: (2025)
Breaking Bad Molecules: Are MLLMs Ready for Structure-Level Molecular Detoxification?
di: Lin, Fei, et al.
Pubblicazione: (2025)
di: Lin, Fei, et al.
Pubblicazione: (2025)
Partially Rewriting a Transformer in Natural Language
di: Paulo, Gonçalo, et al.
Pubblicazione: (2025)
di: Paulo, Gonçalo, et al.
Pubblicazione: (2025)
Dr Genre: Reinforcement Learning from Decoupled LLM Feedback for Generic Text Rewriting
di: Li, Yufei, et al.
Pubblicazione: (2025)
di: Li, Yufei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Text Detoxification: Data Efficiency, Semantic Preservation and Model Generalization
di: Yu, Jing, et al.
Pubblicazione: (2025) -
DetoxLLM: A Framework for Detoxification with Explanations
di: Khondaker, Md Tawkat Islam, et al.
Pubblicazione: (2024) -
Multilingual and Explainable Text Detoxification with Parallel Corpora
di: Dementieva, Daryna, et al.
Pubblicazione: (2024) -
Test-Time Detoxification without Training or Learning Anything
di: Saglam, Baturay, et al.
Pubblicazione: (2026) -
DAPI: Domain Adaptive Toxicity Probe Vector Intervention for Fine-Grained Detoxification
di: Hyeonsu, Cho, et al.
Pubblicazione: (2025)