Enhancing Assamese NLP Capabilities: Introducing a Centralized Dataset Repository
Fuente:
arXiv
Saved in:
| Main Authors: | Tamang, S., Bora, D. J. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Tokenizer Performance of Large Language Models Across Official Indian Languages
by: Tamang, S., et al.
Published: (2024)
by: Tamang, S., et al.
Published: (2024)
Performance Evaluation of Tokenizers in Large Language Models for the Assamese Language
by: Tamang, Sagar, et al.
Published: (2024)
by: Tamang, Sagar, et al.
Published: (2024)
An Innovative Solution: AI-Based Digital Screen-Integrated Tables for Educational Settings
by: Tamang, S., et al.
Published: (2024)
by: Tamang, S., et al.
Published: (2024)
Machine Learning-based NLP for Emotion Classification on a Cholera X Dataset
by: Jideani, Paul, et al.
Published: (2024)
by: Jideani, Paul, et al.
Published: (2024)
AC-Lite : A Lightweight Image Captioning Model for Low-Resource Assamese Language
by: Choudhury, Pankaj, et al.
Published: (2025)
by: Choudhury, Pankaj, et al.
Published: (2025)
Building a Multivariate Time Series Benchmarking Datasets Inspired by Natural Language Processing (NLP)
by: Mustafa, Mohammad Asif Ibna, et al.
Published: (2024)
by: Mustafa, Mohammad Asif Ibna, et al.
Published: (2024)
Enhancing Reasoning Capabilities in SLMs with Reward Guided Dataset Distillation
by: Padarha, Shreyansh
Published: (2025)
by: Padarha, Shreyansh
Published: (2025)
COMI-LINGUA: Expert Annotated Large-Scale Dataset for Multitask NLP in Hindi-English Code-Mixing
by: Sheth, Rajvee, et al.
Published: (2025)
by: Sheth, Rajvee, et al.
Published: (2025)
Introducing A Bangla Sentence - Gloss Pair Dataset for Bangla Sign Language Translation and Research
by: Saha, Neelavro, et al.
Published: (2025)
by: Saha, Neelavro, et al.
Published: (2025)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
by: Powers, Maximus, et al.
Published: (2024)
by: Powers, Maximus, et al.
Published: (2024)
State of NLP in Kenya: A Survey
by: Amol, Cynthia Jayne, et al.
Published: (2024)
by: Amol, Cynthia Jayne, et al.
Published: (2024)
Enforcement Agents: Enhancing Accountability and Resilience in Multi-Agent AI Frameworks
by: Tamang, Sagar, et al.
Published: (2025)
by: Tamang, Sagar, et al.
Published: (2025)
NLP Security and Ethics, in the Wild
by: Lent, Heather, et al.
Published: (2025)
by: Lent, Heather, et al.
Published: (2025)
SCARV: Structure-Constrained Aggregation for Stable Sample Ranking in Redundant NLP Datasets
by: Zheng, Xu, et al.
Published: (2026)
by: Zheng, Xu, et al.
Published: (2026)
Beyond Catalogue Counts: the Dataset Visibility Asymmetry in Low-Resource Multilingual NLP
by: Tan, Zhiyin, et al.
Published: (2026)
by: Tan, Zhiyin, et al.
Published: (2026)
ToolBridge: An Open-Source Dataset to Equip LLMs with External Tool Capabilities
by: Jin, Zhenchao, et al.
Published: (2024)
by: Jin, Zhenchao, et al.
Published: (2024)
Practising responsibility: Ethics in NLP as a hands-on course
by: Nissim, Malvina, et al.
Published: (2025)
by: Nissim, Malvina, et al.
Published: (2025)
Deanthropomorphising NLP: Can a Language Model Be Conscious?
by: Shardlow, Matthew, et al.
Published: (2022)
by: Shardlow, Matthew, et al.
Published: (2022)
An Audit on the Perspectives and Challenges of Hallucinations in NLP
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
UW-BioNLP at ChemoTimelines 2025: Thinking, Fine-Tuning, and Dictionary-Enhanced LLM Systems for Chemotherapy Timeline Extraction
by: Zhang, Tianmai M., et al.
Published: (2025)
by: Zhang, Tianmai M., et al.
Published: (2025)
Optimizing Sentence Embedding with Pseudo-Labeling and Model Ensembles: A Hierarchical Framework for Enhanced NLP Tasks
by: Liu, Ziwei, et al.
Published: (2025)
by: Liu, Ziwei, et al.
Published: (2025)
RLAP: A Reinforcement Learning Enhanced Adaptive Planning Framework for Multi-step NLP Task Solving
by: Ding, Zepeng, et al.
Published: (2025)
by: Ding, Zepeng, et al.
Published: (2025)
Evaluation Metrics for Text Data Augmentation in NLP
by: Amadeus, Marcellus, et al.
Published: (2024)
by: Amadeus, Marcellus, et al.
Published: (2024)
Faithfulness and the Notion of Adversarial Sensitivity in NLP Explanations
by: Manna, Supriya, et al.
Published: (2024)
by: Manna, Supriya, et al.
Published: (2024)
Select, Label, Evaluate: Active Testing in NLP
by: Purificato, Antonio, et al.
Published: (2026)
by: Purificato, Antonio, et al.
Published: (2026)
Undesirable Biases in NLP: Addressing Challenges of Measurement
by: van der Wal, Oskar, et al.
Published: (2022)
by: van der Wal, Oskar, et al.
Published: (2022)
Speaking of Language: Reflections on Metalanguage Research in NLP
by: Schneider, Nathan, et al.
Published: (2026)
by: Schneider, Nathan, et al.
Published: (2026)
Self-Review Framework for Enhancing Instruction Following Capability of LLM
by: Park, Sihyun
Published: (2025)
by: Park, Sihyun
Published: (2025)
Introducing OmniGEC: A Silver Multilingual Dataset for Grammatical Error Correction
by: Kovalchuk, Roman, et al.
Published: (2025)
by: Kovalchuk, Roman, et al.
Published: (2025)
Indian Legal NLP Benchmarks : A Survey
by: Kalamkar, Prathamesh, et al.
Published: (2021)
by: Kalamkar, Prathamesh, et al.
Published: (2021)
Towards Supporting Legal Argumentation with NLP: Is More Data Really All You Need?
by: Santosh, T. Y. S. S, et al.
Published: (2024)
by: Santosh, T. Y. S. S, et al.
Published: (2024)
RepoGraph: Enhancing AI Software Engineering with Repository-level Code Graph
by: Ouyang, Siru, et al.
Published: (2024)
by: Ouyang, Siru, et al.
Published: (2024)
Cooperative Strategic Planning Enhances Reasoning Capabilities in Large Language Models
by: Wang, Danqing, et al.
Published: (2024)
by: Wang, Danqing, et al.
Published: (2024)
Intertwining CP and NLP: The Generation of Unreasonably Constrained Sentences
by: Bonlarron, Alexandre, et al.
Published: (2024)
by: Bonlarron, Alexandre, et al.
Published: (2024)
Large Language Models Meet NLP: A Survey
by: Qin, Libo, et al.
Published: (2024)
by: Qin, Libo, et al.
Published: (2024)
Advancing NLP Security by Leveraging LLMs as Adversarial Engines
by: Srinivasan, Sudarshan, et al.
Published: (2024)
by: Srinivasan, Sudarshan, et al.
Published: (2024)
Facilitating Opinion Diversity through Hybrid NLP Approaches
by: van der Meer, Michiel
Published: (2024)
by: van der Meer, Michiel
Published: (2024)
Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities
by: Liu, Zhichen, et al.
Published: (2026)
by: Liu, Zhichen, et al.
Published: (2026)
Towards Open-Ended Discovery for Low-Resource NLP
by: Dossou, Bonaventure F. P., et al.
Published: (2025)
by: Dossou, Bonaventure F. P., et al.
Published: (2025)
GR-NLP-TOOLKIT: An Open-Source NLP Toolkit for Modern Greek
by: Loukas, Lefteris, et al.
Published: (2024)
by: Loukas, Lefteris, et al.
Published: (2024)
Similar Items
-
Evaluating Tokenizer Performance of Large Language Models Across Official Indian Languages
by: Tamang, S., et al.
Published: (2024) -
Performance Evaluation of Tokenizers in Large Language Models for the Assamese Language
by: Tamang, Sagar, et al.
Published: (2024) -
An Innovative Solution: AI-Based Digital Screen-Integrated Tables for Educational Settings
by: Tamang, S., et al.
Published: (2024) -
Machine Learning-based NLP for Emotion Classification on a Cholera X Dataset
by: Jideani, Paul, et al.
Published: (2024) -
AC-Lite : A Lightweight Image Captioning Model for Low-Resource Assamese Language
by: Choudhury, Pankaj, et al.
Published: (2025)