Investigating the Impact of Semi-Supervised Methods with Data Augmentation on Offensive Language Detection in Romanian Language
Fuente:
arXiv
Salvato in:
| Autori principali: | Nicola, Elena-Beatrice, Cercel, Dumitru-Clementin, Pop, Florin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhancing Romanian Offensive Language Detection through Knowledge Distillation, Multi-Task Learning, and Data Augmentation
di: Matei, Vlad-Cristian, et al.
Pubblicazione: (2024)
di: Matei, Vlad-Cristian, et al.
Pubblicazione: (2024)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
SaRoHead: Detecting Satire in a Multi-Domain Romanian News Headline Dataset
di: Vîrlan, Mihnea-Alexandru, et al.
Pubblicazione: (2025)
di: Vîrlan, Mihnea-Alexandru, et al.
Pubblicazione: (2025)
Multimodal Learning with Augmentation Techniques for Natural Disaster Assessment
di: Urse, Adrian-Dinu, et al.
Pubblicazione: (2025)
di: Urse, Adrian-Dinu, et al.
Pubblicazione: (2025)
Investigating Large Language Models for Complex Word Identification in Multilingual and Multidomain Setups
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2024)
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2024)
Parameter Efficient Multimodal Instruction Tuning for Romanian Vision Language Models
di: Dima, George-Andrei, et al.
Pubblicazione: (2025)
di: Dima, George-Andrei, et al.
Pubblicazione: (2025)
RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian
di: Timpuriu, Mircea, et al.
Pubblicazione: (2026)
di: Timpuriu, Mircea, et al.
Pubblicazione: (2026)
RoD-TAL: A Benchmark for Answering Questions in Romanian Driving License Exams
di: Man, Andrei Vlad, et al.
Pubblicazione: (2025)
di: Man, Andrei Vlad, et al.
Pubblicazione: (2025)
GRAF: Graph Retrieval Augmented by Facts for Romanian Legal Multi-Choice Question Answering
di: Crăciun, Cristian-George, et al.
Pubblicazione: (2024)
di: Crăciun, Cristian-George, et al.
Pubblicazione: (2024)
A Cross-Lingual Meta-Learning Method Based on Domain Adaptation for Speech Emotion Recognition
di: Ion, David-Gabriel, et al.
Pubblicazione: (2024)
di: Ion, David-Gabriel, et al.
Pubblicazione: (2024)
RoQLlama: A Lightweight Romanian Adapted Language Model
di: Dima, George-Andrei, et al.
Pubblicazione: (2024)
di: Dima, George-Andrei, et al.
Pubblicazione: (2024)
MoRoVoc: A Large Dataset for Geographical Variation Identification of the Spoken Romanian Language
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2025)
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2025)
MuSaRoNews: A Multidomain, Multimodal Satire Dataset from Romanian News Articles
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
RoLargeSum: A Large Dialect-Aware Romanian News Dataset for Summary, Headline, and Keyword Generation
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2024)
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2024)
Evaluating Data Augmentation Techniques for Coffee Leaf Disease Classification
di: Gheorghiu, Adrian, et al.
Pubblicazione: (2024)
di: Gheorghiu, Adrian, et al.
Pubblicazione: (2024)
RoIt-XMASA: Multi-Domain Multilingual Sentiment Analysis Dataset for Romanian and Italian
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2026)
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2026)
UniBERT: Adversarial Training for Language-Universal Representations
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2025)
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2025)
IGAff: Benchmarking Adversarial Iterative and Genetic Affine Algorithms on Deep Neural Networks
di: Echim, Sebastian-Vasile, et al.
Pubblicazione: (2025)
di: Echim, Sebastian-Vasile, et al.
Pubblicazione: (2025)
Explainability-Driven Leaf Disease Classification Using Adversarial Training and Knowledge Distillation
di: Echim, Sebastian-Vasile, et al.
Pubblicazione: (2023)
di: Echim, Sebastian-Vasile, et al.
Pubblicazione: (2023)
RoCoISLR: A Romanian Corpus for Isolated Sign Language Recognition
di: Rîpanu, Cătălin-Alexandru, et al.
Pubblicazione: (2025)
di: Rîpanu, Cătălin-Alexandru, et al.
Pubblicazione: (2025)
HistNERo: Historical Named Entity Recognition for the Romanian Language
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2024)
di: Avram, Andrei-Marius, et al.
Pubblicazione: (2024)
LLMic: Romanian Foundation Language Model
di: Bădoiu, Vlad-Andrei, et al.
Pubblicazione: (2025)
di: Bădoiu, Vlad-Andrei, et al.
Pubblicazione: (2025)
FuLG: 150B Romanian Corpus for Language Model Pretraining
di: Bădoiu, Vlad-Andrei, et al.
Pubblicazione: (2024)
di: Bădoiu, Vlad-Andrei, et al.
Pubblicazione: (2024)
RoBiologyDataChoiceQA: A Romanian Dataset for improving Biology understanding of Large Language Models
di: Ghinea, Dragos-Dumitru, et al.
Pubblicazione: (2025)
di: Ghinea, Dragos-Dumitru, et al.
Pubblicazione: (2025)
Scaling Federated Learning Solutions with Kubernetes for Synthesizing Histopathology Images
di: Preda, Andrei-Alexandru, et al.
Pubblicazione: (2025)
di: Preda, Andrei-Alexandru, et al.
Pubblicazione: (2025)
Air Pollution Forecasting in Bucharest
di: Şerban, Dragoş-Andrei, et al.
Pubblicazione: (2025)
di: Şerban, Dragoş-Andrei, et al.
Pubblicazione: (2025)
Detection and Analysis of Offensive Online Content in Hausa Language
di: Adam, Fatima Muhammad, et al.
Pubblicazione: (2023)
di: Adam, Fatima Muhammad, et al.
Pubblicazione: (2023)
OffensiveLang: A Community Based Implicit Offensive Language Dataset
di: Das, Amit, et al.
Pubblicazione: (2024)
di: Das, Amit, et al.
Pubblicazione: (2024)
Subasa - Adapting Language Models for Low-resourced Offensive Language Detection in Sinhala
di: Haturusinghe, Shanilka, et al.
Pubblicazione: (2025)
di: Haturusinghe, Shanilka, et al.
Pubblicazione: (2025)
Guardians of Discourse: Evaluating LLMs on Multilingual Offensive Language Detection
di: He, Jianfei, et al.
Pubblicazione: (2024)
di: He, Jianfei, et al.
Pubblicazione: (2024)
Offensive Language Detection on Social Media Using XLNet
di: Alothman, Reem, et al.
Pubblicazione: (2025)
di: Alothman, Reem, et al.
Pubblicazione: (2025)
Chinese Offensive Language Detection:Current Status and Future Directions
di: Xiao, Yunze, et al.
Pubblicazione: (2024)
di: Xiao, Yunze, et al.
Pubblicazione: (2024)
VocalTweets: Investigating Social Media Offensive Language Among Nigerian Musicians
di: Oluyele, Sunday, et al.
Pubblicazione: (2024)
di: Oluyele, Sunday, et al.
Pubblicazione: (2024)
Lost in Pronunciation: Detecting Chinese Offensive Language Disguised by Phonetic Cloaking Replacement
di: Guo, Haotan, et al.
Pubblicazione: (2025)
di: Guo, Haotan, et al.
Pubblicazione: (2025)
Towards Generalized Offensive Language Identification
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
Conditional Semi-Supervised Data Augmentation for Spam Message Detection with Low Resource Data
di: Nuha, Ulin, et al.
Pubblicazione: (2024)
di: Nuha, Ulin, et al.
Pubblicazione: (2024)
Semi-Supervised Spoken Language Glossification
di: Yao, Huijie, et al.
Pubblicazione: (2024)
di: Yao, Huijie, et al.
Pubblicazione: (2024)
Systematic Offensive Stereotyping (SOS) Bias in Language Models
di: Elsafoury, Fatma
Pubblicazione: (2023)
di: Elsafoury, Fatma
Pubblicazione: (2023)
Cultural Compass: Predicting Transfer Learning Success in Offensive Language Detection with Cultural Features
di: Zhou, Li, et al.
Pubblicazione: (2023)
di: Zhou, Li, et al.
Pubblicazione: (2023)
ToxiCloakCN: Evaluating Robustness of Offensive Language Detection in Chinese with Cloaking Perturbations
di: Xiao, Yunze, et al.
Pubblicazione: (2024)
di: Xiao, Yunze, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Enhancing Romanian Offensive Language Detection through Knowledge Distillation, Multi-Task Learning, and Data Augmentation
di: Matei, Vlad-Cristian, et al.
Pubblicazione: (2024) -
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025) -
SaRoHead: Detecting Satire in a Multi-Domain Romanian News Headline Dataset
di: Vîrlan, Mihnea-Alexandru, et al.
Pubblicazione: (2025) -
Multimodal Learning with Augmentation Techniques for Natural Disaster Assessment
di: Urse, Adrian-Dinu, et al.
Pubblicazione: (2025) -
Investigating Large Language Models for Complex Word Identification in Multilingual and Multidomain Setups
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2024)