| _version_ | 1866901248248119296 |
|---|---|
| author | Hoang Thi My Le |
| author_facet | Hoang Thi My Le |
| contents | Word segmentation is a processing process aimed at determining the boundaries of words in a sentence. Words can also be single words, compound words, etc. In natural language processing, in order to determine the grammatical structure of a sentence, the word class of a word in a sentence, the requirement is to determine the words in the sentence. The word segmentation problem is always the first problem to solve the problem of automatic translation or problems in natural language processing. Besides the problems of studying Vietnamese word segmentation, the problem of Ede word segmentation has not yet been published and shared for the purpose of studying Ede language processing. To solve this problem and based on the current situation of processing ethnic minority languages in Vietnam in general and Ede language in particular, this article proposes a solution to segment Ede words using the maximum matching method based on the Ede vocabulary database, in order to contribute to solving the problem of segmentation in Ede language processing.. |
| format | Recurso digital |
| id | zenodo_https___doi_org_10_5281_zenodo_18104339 |
| institution | Zenodo |
| language | |
| publishDate | 2025 |
| publisher | Zenodo |
| record_format | zenodo |
| spellingShingle | Proposing a Solution to Separate Ede Words Based on the Ede Vocabulary Hoang Thi My Le Bilingual vocabulary database; ethnic minority; Ede vocabulary separation; Ede language processing; maximum matching method. Word segmentation is a processing process aimed at determining the boundaries of words in a sentence. Words can also be single words, compound words, etc. In natural language processing, in order to determine the grammatical structure of a sentence, the word class of a word in a sentence, the requirement is to determine the words in the sentence. The word segmentation problem is always the first problem to solve the problem of automatic translation or problems in natural language processing. Besides the problems of studying Vietnamese word segmentation, the problem of Ede word segmentation has not yet been published and shared for the purpose of studying Ede language processing. To solve this problem and based on the current situation of processing ethnic minority languages in Vietnam in general and Ede language in particular, this article proposes a solution to segment Ede words using the maximum matching method based on the Ede vocabulary database, in order to contribute to solving the problem of segmentation in Ede language processing.. |
| title | Proposing a Solution to Separate Ede Words Based on the Ede Vocabulary |
| topic | Bilingual vocabulary database; ethnic minority; Ede vocabulary separation; Ede language processing; maximum matching method. |
| url | https://doi.org/10.5281/zenodo.18104339 |