A Benchmark Corpus and Neural Approach for Sanskrit Derivative Nouns Analysis

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Singh, Arun Kumar, Dave, Sushant, P., Prathosh A., Lall, Brejesh, Mehta, Shresth
Format: Preprint
Published: 2020
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910587166916608
author Singh, Arun Kumar
Dave, Sushant
P., Prathosh A.
Lall, Brejesh
Mehta, Shresth
author_facet Singh, Arun Kumar
Dave, Sushant
P., Prathosh A.
Lall, Brejesh
Mehta, Shresth
contents This paper presents first benchmark corpus of Sanskrit Pratyaya (suffix) and inflectional words (padas) formed due to suffixes along with neural network based approaches to process the formation and splitting of inflectional words. Inflectional words spans the primary and secondary derivative nouns as the scope of current work. Pratyayas are an important dimension of morphological analysis of Sanskrit texts. There have been Sanskrit Computational Linguistics tools for processing and analyzing Sanskrit texts. Unfortunately there has not been any work to standardize & validate these tools specifically for derivative nouns analysis. In this work, we prepared a Sanskrit suffix benchmark corpus called Pratyaya-Kosh to evaluate the performance of tools. We also present our own neural approach for derivative nouns analysis while evaluating the same on most prominent Sanskrit Morphological Analysis tools. This benchmark will be freely dedicated and available to researchers worldwide and we hope it will motivate all to improve morphological analysis in Sanskrit Language.
format Preprint
id arxiv_https___arxiv_org_abs_2010_12937
institution arXiv
publishDate 2020
record_format arxiv
spellingShingle A Benchmark Corpus and Neural Approach for Sanskrit Derivative Nouns Analysis
Singh, Arun Kumar
Dave, Sushant
P., Prathosh A.
Lall, Brejesh
Mehta, Shresth
Computation and Language
This paper presents first benchmark corpus of Sanskrit Pratyaya (suffix) and inflectional words (padas) formed due to suffixes along with neural network based approaches to process the formation and splitting of inflectional words. Inflectional words spans the primary and secondary derivative nouns as the scope of current work. Pratyayas are an important dimension of morphological analysis of Sanskrit texts. There have been Sanskrit Computational Linguistics tools for processing and analyzing Sanskrit texts. Unfortunately there has not been any work to standardize & validate these tools specifically for derivative nouns analysis. In this work, we prepared a Sanskrit suffix benchmark corpus called Pratyaya-Kosh to evaluate the performance of tools. We also present our own neural approach for derivative nouns analysis while evaluating the same on most prominent Sanskrit Morphological Analysis tools. This benchmark will be freely dedicated and available to researchers worldwide and we hope it will motivate all to improve morphological analysis in Sanskrit Language.
title A Benchmark Corpus and Neural Approach for Sanskrit Derivative Nouns Analysis
topic Computation and Language
url https://arxiv.org/abs/2010.12937