Saved in:
| Main Author: | Khan, Muhammad Shahmeer |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.00444 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RefactorCoderQA: Benchmarking LLMs for Multi-Domain Coding Question Solutions in Cloud and Edge Deployment
by: Rahman, Shadikur, et al.
Published: (2025)
by: Rahman, Shadikur, et al.
Published: (2025)
From Transformers to LLMs: A Systematic Survey of Efficiency Considerations in NLP
by: Ansar, Wazib, et al.
Published: (2024)
by: Ansar, Wazib, et al.
Published: (2024)
Measuring the Robustness of NLP Models to Domain Shifts
by: Calderon, Nitay, et al.
Published: (2023)
by: Calderon, Nitay, et al.
Published: (2023)
SciNLP: A Domain-Specific Benchmark for Full-Text Scientific Entity and Relation Extraction in NLP
by: Duan, Decheng, et al.
Published: (2025)
by: Duan, Decheng, et al.
Published: (2025)
NLP-ADBench: NLP Anomaly Detection Benchmark
by: Li, Yuangang, et al.
Published: (2024)
by: Li, Yuangang, et al.
Published: (2024)
Privacy Evaluation Benchmarks for NLP Models
by: Huang, Wei, et al.
Published: (2024)
by: Huang, Wei, et al.
Published: (2024)
NLP for Social Good: A Survey and Outlook of Challenges, Opportunities, and Responsible Deployment
by: Karamolegkou, Antonia, et al.
Published: (2025)
by: Karamolegkou, Antonia, et al.
Published: (2025)
mhGPT: A Lightweight Generative Pre-Trained Transformer for Mental Health Text Analysis
by: Kim, Dae-young, et al.
Published: (2024)
by: Kim, Dae-young, et al.
Published: (2024)
PIIvot: A Lightweight NLP Anonymization Framework for Question-Anchored Tutoring Dialogues
by: Zent, Matthew, et al.
Published: (2025)
by: Zent, Matthew, et al.
Published: (2025)
EvalxNLP: A Framework for Benchmarking Post-Hoc Explainability Methods on NLP Models
by: Dhaini, Mahdi, et al.
Published: (2025)
by: Dhaini, Mahdi, et al.
Published: (2025)
Towards Inclusive NLP: Assessing Compressed Multilingual Transformers across Diverse Language Benchmarks
by: Alshehhi, Maitha, et al.
Published: (2025)
by: Alshehhi, Maitha, et al.
Published: (2025)
Robust Explanations for User Trust in Enterprise NLP Systems
by: Zhang, Guilin, et al.
Published: (2026)
by: Zhang, Guilin, et al.
Published: (2026)
SEMFED: Semantic-Aware Resource-Efficient Federated Learning for Heterogeneous NLP Tasks
by: Hussain, Sajid, et al.
Published: (2025)
by: Hussain, Sajid, et al.
Published: (2025)
Divide et Impera: Multi-Transformer Architectures for Complex NLP-Tasks
by: Helland, Solveig, et al.
Published: (2023)
by: Helland, Solveig, et al.
Published: (2023)
Towards Incremental Transformers: An Empirical Analysis of Transformer Models for Incremental NLU
by: Kahardipraja, Patrick, et al.
Published: (2021)
by: Kahardipraja, Patrick, et al.
Published: (2021)
AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
by: Nakash, Itay, et al.
Published: (2025)
by: Nakash, Itay, et al.
Published: (2025)
A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP
by: Ono, Shinnosuke, et al.
Published: (2025)
by: Ono, Shinnosuke, et al.
Published: (2025)
LLM and Agent-Driven Data Analysis: A Systematic Approach for Enterprise Applications and System-level Deployment
by: Wang, Xi, et al.
Published: (2025)
by: Wang, Xi, et al.
Published: (2025)
ECBD: Evidence-Centered Benchmark Design for NLP
by: Liu, Yu Lu, et al.
Published: (2024)
by: Liu, Yu Lu, et al.
Published: (2024)
BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models
by: Aung, Thura, et al.
Published: (2026)
by: Aung, Thura, et al.
Published: (2026)
Benchmarking NLP-supported Language Sample Analysis for Swiss Children's Speech
by: Ryser, Anja, et al.
Published: (2025)
by: Ryser, Anja, et al.
Published: (2025)
Enterprise Benchmarks for Large Language Model Evaluation
by: Zhang, Bing, et al.
Published: (2024)
by: Zhang, Bing, et al.
Published: (2024)
Detecting PTSD in Clinical Interviews: A Comparative Analysis of NLP Methods and Large Language Models
by: Chen, Feng, et al.
Published: (2025)
by: Chen, Feng, et al.
Published: (2025)
DRBench: A Realistic Benchmark for Enterprise Deep Research
by: Abaskohi, Amirhossein, et al.
Published: (2025)
by: Abaskohi, Amirhossein, et al.
Published: (2025)
Exploring NLP Benchmarks in an Extremely Low-Resource Setting
by: Nuha, Ulin, et al.
Published: (2025)
by: Nuha, Ulin, et al.
Published: (2025)
When does MAML Work the Best? An Empirical Study on Model-Agnostic Meta-Learning in NLP Applications
by: Liu, Zequn, et al.
Published: (2020)
by: Liu, Zequn, et al.
Published: (2020)
Comparative Analysis of Efficient Adapter-Based Fine-Tuning of State-of-the-Art Transformer Models
by: Siddiqui, Saad Mashkoor, et al.
Published: (2025)
by: Siddiqui, Saad Mashkoor, et al.
Published: (2025)
Transformer Enhanced Relation Classification: A Comparative Analysis of Contextuality, Data Efficiency and Sequence Complexity
by: Jing, Bowen, et al.
Published: (2025)
by: Jing, Bowen, et al.
Published: (2025)
Benchmarking Sociolinguistic Diversity in Swahili NLP: A Taxonomy-Guided Approach
by: Oketch, Kezia, et al.
Published: (2025)
by: Oketch, Kezia, et al.
Published: (2025)
Large Language Model for Multi-Domain Translation: Benchmarking and Domain CoT Fine-tuning
by: Hu, Tianxiang, et al.
Published: (2024)
by: Hu, Tianxiang, et al.
Published: (2024)
Indian Legal NLP Benchmarks : A Survey
by: Kalamkar, Prathamesh, et al.
Published: (2021)
by: Kalamkar, Prathamesh, et al.
Published: (2021)
BriefMe: A Legal NLP Benchmark for Assisting with Legal Briefs
by: Woo, Jesse, et al.
Published: (2025)
by: Woo, Jesse, et al.
Published: (2025)
Evaluating Large Language Models (LLMs) in Financial NLP: A Comparative Study on Financial Report Analysis
by: Mohsin, Md Talha
Published: (2025)
by: Mohsin, Md Talha
Published: (2025)
LogogramNLP: Comparing Visual and Textual Representations of Ancient Logographic Writing Systems for NLP
by: Chen, Danlu, et al.
Published: (2024)
by: Chen, Danlu, et al.
Published: (2024)
Benchmarking Large Language Models on Multiple Tasks in Bioinformatics NLP with Prompting
by: Jiang, Jiyue, et al.
Published: (2025)
by: Jiang, Jiyue, et al.
Published: (2025)
Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP
by: Visser, Ruan, et al.
Published: (2026)
by: Visser, Ruan, et al.
Published: (2026)
Towards Sustainable NLP: Insights from Benchmarking Inference Energy in Large Language Models
by: Poddar, Soham, et al.
Published: (2025)
by: Poddar, Soham, et al.
Published: (2025)
An Empirical Analysis of Fine-Tuning Large Language Models on Bioinformatics Literature: PRSGPT and BioStarsGPT
by: Muneeb, Muhammad, et al.
Published: (2025)
by: Muneeb, Muhammad, et al.
Published: (2025)
LightPAL: Lightweight Passage Retrieval for Open Domain Multi-Document Summarization
by: Enomoto, Masafumi, et al.
Published: (2024)
by: Enomoto, Masafumi, et al.
Published: (2024)
Current State in Privacy-Preserving Text Preprocessing for Domain-Agnostic NLP
by: Sinha, Abhirup, et al.
Published: (2025)
by: Sinha, Abhirup, et al.
Published: (2025)
Similar Items
-
RefactorCoderQA: Benchmarking LLMs for Multi-Domain Coding Question Solutions in Cloud and Edge Deployment
by: Rahman, Shadikur, et al.
Published: (2025) -
From Transformers to LLMs: A Systematic Survey of Efficiency Considerations in NLP
by: Ansar, Wazib, et al.
Published: (2024) -
Measuring the Robustness of NLP Models to Domain Shifts
by: Calderon, Nitay, et al.
Published: (2023) -
SciNLP: A Domain-Specific Benchmark for Full-Text Scientific Entity and Relation Extraction in NLP
by: Duan, Decheng, et al.
Published: (2025) -
NLP-ADBench: NLP Anomaly Detection Benchmark
by: Li, Yuangang, et al.
Published: (2024)