Culture Matters in Toxic Language Detection in Persian
Fuente:
arXiv
Guardado en:
| Autores principales: | Bokaei, Zahra, Magdy, Walid, Webber, Bonnie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Leveraging Hierarchical Prototypes as the Verbalizer for Implicit Discourse Relation Recognition
por: Long, Wanqiu, et al.
Publicado: (2024)
por: Long, Wanqiu, et al.
Publicado: (2024)
Multi-Label Classification for Implicit Discourse Relation Recognition
por: Long, Wanqiu, et al.
Publicado: (2024)
por: Long, Wanqiu, et al.
Publicado: (2024)
Revisiting Common Assumptions about Arabic Dialects in NLP
por: Keleg, Amr, et al.
Publicado: (2025)
por: Keleg, Amr, et al.
Publicado: (2025)
Estimating the Level of Dialectness Predicts Interannotator Agreement in Multi-dialect Arabic Datasets
por: Keleg, Amr, et al.
Publicado: (2024)
por: Keleg, Amr, et al.
Publicado: (2024)
Superlatives in Context: Modeling the Implicit Semantics of Superlatives
por: Pyatkin, Valentina, et al.
Publicado: (2024)
por: Pyatkin, Valentina, et al.
Publicado: (2024)
IslamicMMLU: A Benchmark for Evaluating LLMs on Islamic Knowledge
por: Abdelaal, Ali, et al.
Publicado: (2026)
por: Abdelaal, Ali, et al.
Publicado: (2026)
MELAC: Massive Evaluation of Large Language Models with Alignment of Culture in Persian Language
por: Farsi, Farhan, et al.
Publicado: (2025)
por: Farsi, Farhan, et al.
Publicado: (2025)
PerHalluEval: Persian Hallucination Evaluation Benchmark for Large Language Models
por: Hosseini, Mohammad, et al.
Publicado: (2025)
por: Hosseini, Mohammad, et al.
Publicado: (2025)
ELAB: Extensive LLM Alignment Benchmark in Persian Language
por: Pourbahman, Zahra, et al.
Publicado: (2025)
por: Pourbahman, Zahra, et al.
Publicado: (2025)
PersianRAG: A Retrieval-Augmented Generation System for Persian Language
por: Hosseini, Hossein, et al.
Publicado: (2024)
por: Hosseini, Hossein, et al.
Publicado: (2024)
PerMedCQA: Benchmarking Large Language Models on Medical Consumer Question Answering in Persian Language
por: Jamali, Naghmeh, et al.
Publicado: (2025)
por: Jamali, Naghmeh, et al.
Publicado: (2025)
Large Language Models for Toxic Language Detection in Low-Resource Balkan Languages
por: Muminovic, Amel, et al.
Publicado: (2025)
por: Muminovic, Amel, et al.
Publicado: (2025)
EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models
por: Mirbagheri, Mohammad Reza, et al.
Publicado: (2025)
por: Mirbagheri, Mohammad Reza, et al.
Publicado: (2025)
TookaBERT: A Step Forward for Persian NLU
por: SadraeiJavaheri, MohammadAli, et al.
Publicado: (2024)
por: SadraeiJavaheri, MohammadAli, et al.
Publicado: (2024)
ETT: Expanding the Long Context Understanding Capability of LLMs at Test-Time
por: Zahirnia, Kiarash, et al.
Publicado: (2025)
por: Zahirnia, Kiarash, et al.
Publicado: (2025)
In-game Toxic Language Detection: Shared Task and Attention Residuals
por: Jia, Yuanzhe, et al.
Publicado: (2022)
por: Jia, Yuanzhe, et al.
Publicado: (2022)
Toxicity Detection for Free
por: Hu, Zhanhao, et al.
Publicado: (2024)
por: Hu, Zhanhao, et al.
Publicado: (2024)
Unlocking the Potential of ChatGPT: A Comprehensive Exploration of its Applications, Advantages, Limitations, and Future Directions in Natural Language Processing
por: Hariri, Walid
Publicado: (2023)
por: Hariri, Walid
Publicado: (2023)
TARAZ: Persian Short-Answer Question Benchmark for Cultural Evaluation of Language Models
por: Iranmanesh, Reihaneh, et al.
Publicado: (2026)
por: Iranmanesh, Reihaneh, et al.
Publicado: (2026)
Incorporating Error Level Noise Embedding for Improving LLM-Assisted Robustness in Persian Speech Recognition
por: Rahmani, Zahra, et al.
Publicado: (2025)
por: Rahmani, Zahra, et al.
Publicado: (2025)
Leveraging Online Data to Enhance Medical Knowledge in a Small Persian Language Model
por: Ghassabi, Mehrdad, et al.
Publicado: (2025)
por: Ghassabi, Mehrdad, et al.
Publicado: (2025)
Detecting Toxic Language: Ontology and BERT-based Approaches for Bulgarian Text
por: Berbatova, Melania, et al.
Publicado: (2026)
por: Berbatova, Melania, et al.
Publicado: (2026)
FaMTEB: Massive Text Embedding Benchmark in Persian Language
por: Zinvandi, Erfan, et al.
Publicado: (2025)
por: Zinvandi, Erfan, et al.
Publicado: (2025)
Persian MusicGen: A Large-Scale Dataset and Culturally-Aware Generative Model for Persian Music
por: Sameti, Mohammad Hossein, et al.
Publicado: (2026)
por: Sameti, Mohammad Hossein, et al.
Publicado: (2026)
From Disagreement to Understanding: The Case for Ambiguity Detection in NLI
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
Detecting Bias and Enhancing Diagnostic Accuracy in Large Language Models for Healthcare
por: Zahraei, Pardis Sadat, et al.
Publicado: (2024)
por: Zahraei, Pardis Sadat, et al.
Publicado: (2024)
All Languages Matter: Evaluating LMMs on Culturally Diverse 100 Languages
por: Vayani, Ashmal, et al.
Publicado: (2024)
por: Vayani, Ashmal, et al.
Publicado: (2024)
Dialectal Toxicity Detection: Evaluating LLM-as-a-Judge Consistency Across Language Varieties
por: Faisal, Fahim, et al.
Publicado: (2024)
por: Faisal, Fahim, et al.
Publicado: (2024)
Beyond Toxic: Toxicity Detection Datasets are Not Enough for Brand Safety
por: Korotkova, Elizaveta, et al.
Publicado: (2023)
por: Korotkova, Elizaveta, et al.
Publicado: (2023)
BLADE: Better Language Answers through Dialogue and Explanations
por: Jayaweera, Chathuri, et al.
Publicado: (2026)
por: Jayaweera, Chathuri, et al.
Publicado: (2026)
PolygloToxicityPrompts: Multilingual Evaluation of Neural Toxic Degeneration in Large Language Models
por: Jain, Devansh, et al.
Publicado: (2024)
por: Jain, Devansh, et al.
Publicado: (2024)
MasalBench: A Benchmark for Contextual and Cross-Cultural Understanding of Persian Proverbs in LLMs
por: Kalhor, Ghazal, et al.
Publicado: (2026)
por: Kalhor, Ghazal, et al.
Publicado: (2026)
On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection
por: Elsafoury, Fatma, et al.
Publicado: (2023)
por: Elsafoury, Fatma, et al.
Publicado: (2023)
Reinterpreting 'the Company a Word Keeps': Towards Explainable and Ontologically Grounded Language Models
por: Saba, Walid S.
Publicado: (2024)
por: Saba, Walid S.
Publicado: (2024)
"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth
por: Li, Yaqiong, et al.
Publicado: (2025)
por: Li, Yaqiong, et al.
Publicado: (2025)
PersianMind: A Cross-Lingual Persian-English Large Language Model
por: Rostami, Pedram, et al.
Publicado: (2024)
por: Rostami, Pedram, et al.
Publicado: (2024)
Khayyam Challenge (PersianMMLU): Is Your LLM Truly Wise to The Persian Language?
por: Ghahroodi, Omid, et al.
Publicado: (2024)
por: Ghahroodi, Omid, et al.
Publicado: (2024)
Cultural Compass: Predicting Transfer Learning Success in Offensive Language Detection with Cultural Features
por: Zhou, Li, et al.
Publicado: (2023)
por: Zhou, Li, et al.
Publicado: (2023)
Unmasking the Factual-Conceptual Gap in Persian Language Models
por: Sakhaeirad, Alireza, et al.
Publicado: (2026)
por: Sakhaeirad, Alireza, et al.
Publicado: (2026)
Deep Learning-based Sentiment Analysis in Persian Language
por: Heydari, Mohammad, et al.
Publicado: (2024)
por: Heydari, Mohammad, et al.
Publicado: (2024)
Ejemplares similares
-
Leveraging Hierarchical Prototypes as the Verbalizer for Implicit Discourse Relation Recognition
por: Long, Wanqiu, et al.
Publicado: (2024) -
Multi-Label Classification for Implicit Discourse Relation Recognition
por: Long, Wanqiu, et al.
Publicado: (2024) -
Revisiting Common Assumptions about Arabic Dialects in NLP
por: Keleg, Amr, et al.
Publicado: (2025) -
Estimating the Level of Dialectness Predicts Interannotator Agreement in Multi-dialect Arabic Datasets
por: Keleg, Amr, et al.
Publicado: (2024) -
Superlatives in Context: Modeling the Implicit Semantics of Superlatives
por: Pyatkin, Valentina, et al.
Publicado: (2024)