Slovo: Russian Sign Language Dataset

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Kapitanov, Alexander, Kvanchiani, Karina, Nagaev, Alexander, Petrova, Elizaveta
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910639391244288
author Kapitanov, Alexander
Kvanchiani, Karina
Nagaev, Alexander
Petrova, Elizaveta
author_facet Kapitanov, Alexander
Kvanchiani, Karina
Nagaev, Alexander
Petrova, Elizaveta
contents One of the main challenges of the sign language recognition task is the difficulty of collecting a suitable dataset due to the gap between hard-of-hearing and hearing societies. In addition, the sign language in each country differs significantly, which obliges the creation of new data for each of them. This paper presents the Russian Sign Language (RSL) video dataset Slovo, produced using crowdsourcing platforms. The dataset contains 20,000 FullHD recordings, divided into 1,000 classes of isolated RSL gestures received by 194 signers. We also provide the entire dataset creation pipeline, from data collection to video annotation, with the following demo application. Several neural networks are trained and evaluated on the Slovo to demonstrate its teaching ability. Proposed data and pre-trained models are publicly available.
format Preprint
id arxiv_https___arxiv_org_abs_2305_14527
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Slovo: Russian Sign Language Dataset
Kapitanov, Alexander
Kvanchiani, Karina
Nagaev, Alexander
Petrova, Elizaveta
Computer Vision and Pattern Recognition
One of the main challenges of the sign language recognition task is the difficulty of collecting a suitable dataset due to the gap between hard-of-hearing and hearing societies. In addition, the sign language in each country differs significantly, which obliges the creation of new data for each of them. This paper presents the Russian Sign Language (RSL) video dataset Slovo, produced using crowdsourcing platforms. The dataset contains 20,000 FullHD recordings, divided into 1,000 classes of isolated RSL gestures received by 194 signers. We also provide the entire dataset creation pipeline, from data collection to video annotation, with the following demo application. Several neural networks are trained and evaluated on the Slovo to demonstrate its teaching ability. Proposed data and pre-trained models are publicly available.
title Slovo: Russian Sign Language Dataset
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2305.14527