Sign Language Recognition using Bidirectional Reservoir Computing

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Singh, Nitin Kumar, Syulistyo, Arie Rachmad, Tanaka, Yuichiro, Tamukoh, Hakaru
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866911334287802368
author Singh, Nitin Kumar
Syulistyo, Arie Rachmad
Tanaka, Yuichiro
Tamukoh, Hakaru
author_facet Singh, Nitin Kumar
Syulistyo, Arie Rachmad
Tanaka, Yuichiro
Tamukoh, Hakaru
contents Sign language recognition (SLR) facilitates communication between deaf and hearing individuals. Deep learning is widely used to develop SLR-based systems; however, it is computationally intensive and requires substantial computational resources, making it unsuitable for resource-constrained devices. To address this, we propose an efficient sign language recognition system using MediaPipe and an echo state network (ESN)-based bidirectional reservoir computing (BRC) architecture. MediaPipe extracts hand joint coordinates, which serve as inputs to the ESN-based BRC architecture. The BRC processes these features in both forward and backward directions, efficiently capturing temporal dependencies. The resulting states of BRC are concatenated to form a robust representation for classification. We evaluated our method on the Word-Level American Sign Language (WLASL) video dataset, achieving a competitive accuracy of 57.71% and a significantly lower training time of only 9 seconds, in contrast to the 55 minutes and $38$ seconds required by the deep learning-based Bi-GRU approach. Consequently, the BRC-based SLR system is well-suited for edge devices.
format Preprint
id arxiv_https___arxiv_org_abs_2512_00777
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Sign Language Recognition using Bidirectional Reservoir Computing
Singh, Nitin Kumar
Syulistyo, Arie Rachmad
Tanaka, Yuichiro
Tamukoh, Hakaru
Robotics
Computer Vision and Pattern Recognition
Sign language recognition (SLR) facilitates communication between deaf and hearing individuals. Deep learning is widely used to develop SLR-based systems; however, it is computationally intensive and requires substantial computational resources, making it unsuitable for resource-constrained devices. To address this, we propose an efficient sign language recognition system using MediaPipe and an echo state network (ESN)-based bidirectional reservoir computing (BRC) architecture. MediaPipe extracts hand joint coordinates, which serve as inputs to the ESN-based BRC architecture. The BRC processes these features in both forward and backward directions, efficiently capturing temporal dependencies. The resulting states of BRC are concatenated to form a robust representation for classification. We evaluated our method on the Word-Level American Sign Language (WLASL) video dataset, achieving a competitive accuracy of 57.71% and a significantly lower training time of only 9 seconds, in contrast to the 55 minutes and $38$ seconds required by the deep learning-based Bi-GRU approach. Consequently, the BRC-based SLR system is well-suited for edge devices.
title Sign Language Recognition using Bidirectional Reservoir Computing
topic Robotics
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2512.00777