Benchmarking BERT-based Models for Sentence-level Topic Classification in Nepali Language
Fuente:
arXiv
Saved in:
| Main Authors: | Karki, Nischal, Subedi, Bipesh, Poudyal, Prakash, Ghimire, Rupak Raj, Bal, Bal Krishna |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NepTam: A Nepali-Tamang Parallel Corpus and Baseline Machine Translation Experiments
by: Ghimire, Rupak Raj, et al.
Published: (2026)
by: Ghimire, Rupak Raj, et al.
Published: (2026)
A Comprehensive Study of the Current State-of-the-Art in Nepali Automatic Speech Recognition Systems
by: Ghimire, Rupak Raj, et al.
Published: (2024)
by: Ghimire, Rupak Raj, et al.
Published: (2024)
Development of Pre-Trained Transformer-based Models for the Nepali Language
by: Thapa, Prajwal, et al.
Published: (2024)
by: Thapa, Prajwal, et al.
Published: (2024)
Consolidating and Developing Benchmarking Datasets for the Nepali Natural Language Understanding Tasks
by: Nyachhyon, Jinu, et al.
Published: (2024)
by: Nyachhyon, Jinu, et al.
Published: (2024)
Nwāchā Munā: A Devanagari Speech Corpus and Proximal Transfer Benchmark for Nepal Bhasha ASR
by: Sharma, Rishikesh Kumar, et al.
Published: (2026)
by: Sharma, Rishikesh Kumar, et al.
Published: (2026)
Nepali Passport Question Answering: A Low-Resource Dataset for Public Service Applications
by: Begha, Funghang Limbu, et al.
Published: (2026)
by: Begha, Funghang Limbu, et al.
Published: (2026)
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining
by: Bal, Melis Ilayda, et al.
Published: (2025)
by: Bal, Melis Ilayda, et al.
Published: (2025)
Breaking BERT: Gradient Attack on Twitter Sentiment Analysis for Targeted Misclassification
by: Subedi, Akil Raj, et al.
Published: (2025)
by: Subedi, Akil Raj, et al.
Published: (2025)
Topic Modeling with Fine-tuning LLMs and Bag of Sentences
by: Schneider, Johannes
Published: (2024)
by: Schneider, Johannes
Published: (2024)
Towards Building Efficient Sentence BERT Models using Layer Pruning
by: Shelke, Anushka, et al.
Published: (2024)
by: Shelke, Anushka, et al.
Published: (2024)
Empowering Interdisciplinary Research with BERT-Based Models: An Approach Through SciBERT-CNN with Topic Modeling
by: Likhareva, Darya, et al.
Published: (2024)
by: Likhareva, Darya, et al.
Published: (2024)
Efficient and Flexible Topic Modeling using Pretrained Embeddings and Bag of Sentences
by: Schneider, Johannes
Published: (2023)
by: Schneider, Johannes
Published: (2023)
A BERT-based Hierarchical Classification Model with Applications in Chinese Commodity Classification
by: Liu, Kun, et al.
Published: (2025)
by: Liu, Kun, et al.
Published: (2025)
Improving Sampling Methods for Fine-tuning SentenceBERT in Text Streams
by: Garcia, Cristiano Mesquita, et al.
Published: (2024)
by: Garcia, Cristiano Mesquita, et al.
Published: (2024)
Leveraging Large Language Models and Topic Modeling for Toxicity Classification
by: Oskouie, Haniyeh Ehsani, et al.
Published: (2024)
by: Oskouie, Haniyeh Ehsani, et al.
Published: (2024)
Sentence Curve Language Models
by: Heo, DongNyeong, et al.
Published: (2026)
by: Heo, DongNyeong, et al.
Published: (2026)
Can Perplexity Predict Fine-tuning Performance? An Investigation of Tokenization Effects on Sequential Language Models for Nepali
by: Luitel, Nishant, et al.
Published: (2024)
by: Luitel, Nishant, et al.
Published: (2024)
2-Tier SimCSE: Elevating BERT for Robust Sentence Embeddings
by: Wang, Yumeng, et al.
Published: (2025)
by: Wang, Yumeng, et al.
Published: (2025)
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models
by: Zhang, Ying, et al.
Published: (2024)
by: Zhang, Ying, et al.
Published: (2024)
ColBERT: Using BERT Sentence Embedding in Parallel Neural Networks for Computational Humor
by: Annamoradnejad, Issa, et al.
Published: (2020)
by: Annamoradnejad, Issa, et al.
Published: (2020)
Computational Sentence-level Metrics Predicting Human Sentence Comprehension
by: Sun, Kun, et al.
Published: (2024)
by: Sun, Kun, et al.
Published: (2024)
MaxPoolBERT: Enhancing BERT Classification via Layer- and Token-Wise Aggregation
by: Behrendt, Maike, et al.
Published: (2025)
by: Behrendt, Maike, et al.
Published: (2025)
HyperBERT: Mixing Hypergraph-Aware Layers with Language Models for Node Classification on Text-Attributed Hypergraphs
by: Bazaga, Adrián, et al.
Published: (2024)
by: Bazaga, Adrián, et al.
Published: (2024)
LakotaBERT: A Transformer-based Model for Low Resource Lakota Language
by: Parankusham, Kanishka, et al.
Published: (2025)
by: Parankusham, Kanishka, et al.
Published: (2025)
BERT-ASC: Auxiliary-Sentence Construction for Implicit Aspect Learning in Sentiment Analysis
by: Ahmed, Murtadha, et al.
Published: (2022)
by: Ahmed, Murtadha, et al.
Published: (2022)
Topics as Entity Clusters: Entity-based Topics from Large Language Models and Graph Neural Networks
by: Loureiro, Manuel V., et al.
Published: (2023)
by: Loureiro, Manuel V., et al.
Published: (2023)
Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models
by: Eleraqi, Amr, et al.
Published: (2026)
by: Eleraqi, Amr, et al.
Published: (2026)
GottBERT: a pure German Language Model
by: Scheible, Raphael, et al.
Published: (2020)
by: Scheible, Raphael, et al.
Published: (2020)
Small Vision-Language Models: A Survey on Compact Architectures and Techniques
by: Patnaik, Nitesh, et al.
Published: (2025)
by: Patnaik, Nitesh, et al.
Published: (2025)
Domain-adaptative Continual Learning for Low-resource Tasks: Evaluation on Nepali
by: Duwal, Sharad, et al.
Published: (2024)
by: Duwal, Sharad, et al.
Published: (2024)
Improving Socratic Question Generation using Data Augmentation and Preference Optimization
by: Kumar, Nischal Ashok, et al.
Published: (2024)
by: Kumar, Nischal Ashok, et al.
Published: (2024)
DE$^3$-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks
by: He, Jianing, et al.
Published: (2024)
by: He, Jianing, et al.
Published: (2024)
Topic mining based on fine-tuning Sentence-BERT and LDA
by: Li, Jianheng, et al.
Published: (2025)
by: Li, Jianheng, et al.
Published: (2025)
Improving the Efficiency of Long Document Classification using Sentence Ranking Approach
by: Kokate, Prathamesh, et al.
Published: (2025)
by: Kokate, Prathamesh, et al.
Published: (2025)
Hierarchical Classification of Transversal Skills in Job Ads Based on Sentence Embeddings
by: Leon, Florin, et al.
Published: (2024)
by: Leon, Florin, et al.
Published: (2024)
On Importance of Layer Pruning for Smaller BERT Models and Low Resource Languages
by: Shirke, Mayur, et al.
Published: (2025)
by: Shirke, Mayur, et al.
Published: (2025)
Enhancing Foundation Models in Transaction Understanding with LLM-based Sentence Embeddings
by: Fan, Xiran, et al.
Published: (2025)
by: Fan, Xiran, et al.
Published: (2025)
Ensemble BERT for Medication Event Classification on Electronic Health Records (EHRs)
by: Sarker, Shouvon, et al.
Published: (2025)
by: Sarker, Shouvon, et al.
Published: (2025)
LimTopic: LLM-based Topic Modeling and Text Summarization for Analyzing Scientific Articles limitations
by: Azhar, Ibrahim Al, et al.
Published: (2025)
by: Azhar, Ibrahim Al, et al.
Published: (2025)
Research on the Application of Deep Learning-based BERT Model in Sentiment Analysis
by: Wu, Yichao, et al.
Published: (2024)
by: Wu, Yichao, et al.
Published: (2024)
Similar Items
-
NepTam: A Nepali-Tamang Parallel Corpus and Baseline Machine Translation Experiments
by: Ghimire, Rupak Raj, et al.
Published: (2026) -
A Comprehensive Study of the Current State-of-the-Art in Nepali Automatic Speech Recognition Systems
by: Ghimire, Rupak Raj, et al.
Published: (2024) -
Development of Pre-Trained Transformer-based Models for the Nepali Language
by: Thapa, Prajwal, et al.
Published: (2024) -
Consolidating and Developing Benchmarking Datasets for the Nepali Natural Language Understanding Tasks
by: Nyachhyon, Jinu, et al.
Published: (2024) -
Nwāchā Munā: A Devanagari Speech Corpus and Proximal Transfer Benchmark for Nepal Bhasha ASR
by: Sharma, Rishikesh Kumar, et al.
Published: (2026)