L3Cube-MahaSocialNER: A Social Media based Marathi NER Dataset and BERT models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chaudhari, Harsh, Patil, Anuja, Lavekar, Dhanashree, Khairnar, Pranav, Joshi, Raviraj |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
L3Cube-MahaSTS: A Marathi Sentence Similarity Dataset and Models
von: Mirashi, Aishwarya, et al.
Veröffentlicht: (2025)
von: Mirashi, Aishwarya, et al.
Veröffentlicht: (2025)
L3Cube-MahaNews: News-based Short Text and Long Document Classification Datasets in Marathi
von: Mittal, Saloni, et al.
Veröffentlicht: (2024)
von: Mittal, Saloni, et al.
Veröffentlicht: (2024)
MahaParaphrase: A Marathi Paraphrase Detection Corpus and BERT-based Models
von: Jadhav, Suramya, et al.
Veröffentlicht: (2025)
von: Jadhav, Suramya, et al.
Veröffentlicht: (2025)
L3Cube-MahaEmotions: A Marathi Emotion Recognition Dataset with Synthetic Annotations using CoTR prompting and Large Language Models
von: Kowtal, Nidhi, et al.
Veröffentlicht: (2025)
von: Kowtal, Nidhi, et al.
Veröffentlicht: (2025)
L3Cube-MahaSum: A Comprehensive Dataset and BART Models for Abstractive Text Summarization in Marathi
von: Deshmukh, Pranita, et al.
Veröffentlicht: (2024)
von: Deshmukh, Pranita, et al.
Veröffentlicht: (2024)
MahaSQuAD: Bridging Linguistic Divides in Marathi Question-Answering
von: Ghatage, Ruturaj, et al.
Veröffentlicht: (2024)
von: Ghatage, Ruturaj, et al.
Veröffentlicht: (2024)
CMNER: A Chinese Multimodal NER Dataset based on Social Media
von: Ji, Yuanze, et al.
Veröffentlicht: (2024)
von: Ji, Yuanze, et al.
Veröffentlicht: (2024)
Topic Modeling in Marathi
von: Shinde, Sanket, et al.
Veröffentlicht: (2025)
von: Shinde, Sanket, et al.
Veröffentlicht: (2025)
Curating Stopwords in Marathi: A TF-IDF Approach for Improved Text Analysis and Information Retrieval
von: Chavan, Rohan, et al.
Veröffentlicht: (2024)
von: Chavan, Rohan, et al.
Veröffentlicht: (2024)
TriNER: A Series of Named Entity Recognition Models For Hindi, Bengali & Marathi
von: Dhamaskar, Mohammed Amaan, et al.
Veröffentlicht: (2025)
von: Dhamaskar, Mohammed Amaan, et al.
Veröffentlicht: (2025)
Noise reduction in BERT NER models for clinical entity extraction
von: Jiwani, Kuldeep, et al.
Veröffentlicht: (2026)
von: Jiwani, Kuldeep, et al.
Veröffentlicht: (2026)
The GELATO Dataset for Legislative NER
von: Flynn, Matthew, et al.
Veröffentlicht: (2026)
von: Flynn, Matthew, et al.
Veröffentlicht: (2026)
LegalTurk Optimized BERT for Multi-Label Text Classification and NER
von: Zeidi, Farnaz, et al.
Veröffentlicht: (2024)
von: Zeidi, Farnaz, et al.
Veröffentlicht: (2024)
Human-Annotated NER Dataset for the Kyrgyz Language
von: Turatali, Timur, et al.
Veröffentlicht: (2025)
von: Turatali, Timur, et al.
Veröffentlicht: (2025)
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages
von: Mirashi, Aishwarya, et al.
Veröffentlicht: (2024)
von: Mirashi, Aishwarya, et al.
Veröffentlicht: (2024)
Long Range Named Entity Recognition for Marathi Documents
von: Deshmukh, Pranita, et al.
Veröffentlicht: (2024)
von: Deshmukh, Pranita, et al.
Veröffentlicht: (2024)
Label Unification for Cross-Dataset Generalization in Cybersecurity NER
von: Jalocha, Maciej, et al.
Veröffentlicht: (2025)
von: Jalocha, Maciej, et al.
Veröffentlicht: (2025)
WikiNER-fr-gold: A Gold-Standard NER Corpus
von: Cao, Danrun, et al.
Veröffentlicht: (2024)
von: Cao, Danrun, et al.
Veröffentlicht: (2024)
Multilingual Clinical NER for Diseases and Medications Recognition in Cardiology Texts using BERT Embeddings
von: Danu, Manuela Daniela, et al.
Veröffentlicht: (2025)
von: Danu, Manuela Daniela, et al.
Veröffentlicht: (2025)
The Million-Label NER: Breaking Scale Barriers with GLiNER bi-encoder
von: Stepanov, Ihor, et al.
Veröffentlicht: (2026)
von: Stepanov, Ihor, et al.
Veröffentlicht: (2026)
L3Cube-IndicQuest: A Benchmark Question Answering Dataset for Evaluating Knowledge of LLMs in Indic Context
von: Rohera, Pritika, et al.
Veröffentlicht: (2024)
von: Rohera, Pritika, et al.
Veröffentlicht: (2024)
OpenMed NER: Open-Source, Domain-Adapted State-of-the-Art Transformers for Biomedical NER Across 12 Public Datasets
von: Panahi, Maziyar
Veröffentlicht: (2025)
von: Panahi, Maziyar
Veröffentlicht: (2025)
L3Cube-IndicHeadline-ID: A Dataset for Headline Identification and Semantic Evaluation in Low-Resource Indian Languages
von: Tanksale, Nishant, et al.
Veröffentlicht: (2025)
von: Tanksale, Nishant, et al.
Veröffentlicht: (2025)
VerifiNER: Verification-augmented NER via Knowledge-grounded Reasoning with Large Language Models
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
2M-NER: Contrastive Learning for Multilingual and Multimodal NER with Language and Modal Fusion
von: Wang, Dongsheng, et al.
Veröffentlicht: (2024)
von: Wang, Dongsheng, et al.
Veröffentlicht: (2024)
Leveraging Parameter Efficient Training Methods for Low Resource Text Classification: A Case Study in Marathi
von: Deshmukh, Pranita, et al.
Veröffentlicht: (2024)
von: Deshmukh, Pranita, et al.
Veröffentlicht: (2024)
Augmenting NER Datasets with LLMs: Towards Automated and Refined Annotation
von: Naraki, Yuji, et al.
Veröffentlicht: (2024)
von: Naraki, Yuji, et al.
Veröffentlicht: (2024)
Towards Building Efficient Sentence BERT Models using Layer Pruning
von: Shelke, Anushka, et al.
Veröffentlicht: (2024)
von: Shelke, Anushka, et al.
Veröffentlicht: (2024)
MariNER: A Dataset for Historical Brazilian Portuguese Named Entity Recognition
von: Sarcinelli, João Lucas Luz Lima, et al.
Veröffentlicht: (2025)
von: Sarcinelli, João Lucas Luz Lima, et al.
Veröffentlicht: (2025)
Do LLMs Surpass Encoders for Biomedical NER?
von: Obeidat, Motasem S, et al.
Veröffentlicht: (2025)
von: Obeidat, Motasem S, et al.
Veröffentlicht: (2025)
Comparative Analysis of Extrinsic Factors for NER in French
von: Yang, Grace, et al.
Veröffentlicht: (2024)
von: Yang, Grace, et al.
Veröffentlicht: (2024)
Novel Benchmark for NER in the Wastewater and Stormwater Domain
von: Cardillo, Franco Alberto, et al.
Veröffentlicht: (2025)
von: Cardillo, Franco Alberto, et al.
Veröffentlicht: (2025)
On-the-fly Definition Augmentation of LLMs for Biomedical NER
von: Munnangi, Monica, et al.
Veröffentlicht: (2024)
von: Munnangi, Monica, et al.
Veröffentlicht: (2024)
NER- RoBERTa: Fine-Tuning RoBERTa for Named Entity Recognition (NER) within low-resource languages
von: Abdullah, Abdulhady Abas, et al.
Veröffentlicht: (2024)
von: Abdullah, Abdulhady Abas, et al.
Veröffentlicht: (2024)
FiNER-ORD: Financial Named Entity Recognition Open Research Dataset
von: Shah, Agam, et al.
Veröffentlicht: (2023)
von: Shah, Agam, et al.
Veröffentlicht: (2023)
ANCHOLIK-NER: A Benchmark Dataset for Bangla Regional Named Entity Recognition
von: Paul, Bidyarthi, et al.
Veröffentlicht: (2025)
von: Paul, Bidyarthi, et al.
Veröffentlicht: (2025)
YoNER: A New Yorùbá Multi-domain Named Entity Recognition Dataset
von: Falola, Peace Busola, et al.
Veröffentlicht: (2026)
von: Falola, Peace Busola, et al.
Veröffentlicht: (2026)
PrionNER: A Named Entity Recognition Dataset for Prion Disease Biomedical Literature
von: Dao, An, et al.
Veröffentlicht: (2026)
von: Dao, An, et al.
Veröffentlicht: (2026)
StanfordNER Gender-Classifier
von: Mareike Schumacher, et al.
Veröffentlicht: (2021)
von: Mareike Schumacher, et al.
Veröffentlicht: (2021)
QueryNER: Segmentation of E-commerce Queries
von: Palen-Michel, Chester, et al.
Veröffentlicht: (2024)
von: Palen-Michel, Chester, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
L3Cube-MahaSTS: A Marathi Sentence Similarity Dataset and Models
von: Mirashi, Aishwarya, et al.
Veröffentlicht: (2025) -
L3Cube-MahaNews: News-based Short Text and Long Document Classification Datasets in Marathi
von: Mittal, Saloni, et al.
Veröffentlicht: (2024) -
MahaParaphrase: A Marathi Paraphrase Detection Corpus and BERT-based Models
von: Jadhav, Suramya, et al.
Veröffentlicht: (2025) -
L3Cube-MahaEmotions: A Marathi Emotion Recognition Dataset with Synthetic Annotations using CoTR prompting and Large Language Models
von: Kowtal, Nidhi, et al.
Veröffentlicht: (2025) -
L3Cube-MahaSum: A Comprehensive Dataset and BART Models for Abstractive Text Summarization in Marathi
von: Deshmukh, Pranita, et al.
Veröffentlicht: (2024)