Staff View: :: Library Catalog

Saved in:

Bibliographic Details
Main Authors:	Rastgoo, Razieh, Kiani, Kourosh, Escalera, Sergio
Format:	Preprint
Published:	2024
Subjects:	Computer Vision and Pattern Recognition
Online Access:	https://arxiv.org/abs/2402.14720
Tags:	Add Tag No Tags, Be the first to tag this record!

_version_	1866910340893114368
author	Rastgoo, Razieh Kiani, Kourosh Escalera, Sergio
author_facet	Rastgoo, Razieh Kiani, Kourosh Escalera, Sergio
contents	Sign Language Recognition (SLR) has garnered significant attention from researchers in recent years, particularly the intricate domain of Continuous Sign Language Recognition (CSLR), which presents heightened complexity compared to Isolated Sign Language Recognition (ISLR). One of the prominent challenges in CSLR pertains to accurately detecting the boundaries of isolated signs within a continuous video stream. Additionally, the reliance on handcrafted features in existing models poses a challenge to achieving optimal accuracy. To surmount these challenges, we propose a novel approach utilizing a Transformer-based model. Unlike traditional models, our approach focuses on enhancing accuracy while eliminating the need for handcrafted features. The Transformer model is employed for both ISLR and CSLR. The training process involves using isolated sign videos, where hand keypoint features extracted from the input video are enriched using the Transformer model. Subsequently, these enriched features are forwarded to the final classification layer. The trained model, coupled with a post-processing method, is then applied to detect isolated sign boundaries within continuous sign videos. The evaluation of our model is conducted on two distinct datasets, including both continuous signs and their corresponding isolated signs, demonstrates promising results.
format	Preprint
id	arxiv_https___arxiv_org_abs_2402_14720
institution	arXiv
publishDate	2024
record_format	arxiv
spellingShingle	A Transformer Model for Boundary Detection in Continuous Sign Language Rastgoo, Razieh Kiani, Kourosh Escalera, Sergio Computer Vision and Pattern Recognition Sign Language Recognition (SLR) has garnered significant attention from researchers in recent years, particularly the intricate domain of Continuous Sign Language Recognition (CSLR), which presents heightened complexity compared to Isolated Sign Language Recognition (ISLR). One of the prominent challenges in CSLR pertains to accurately detecting the boundaries of isolated signs within a continuous video stream. Additionally, the reliance on handcrafted features in existing models poses a challenge to achieving optimal accuracy. To surmount these challenges, we propose a novel approach utilizing a Transformer-based model. Unlike traditional models, our approach focuses on enhancing accuracy while eliminating the need for handcrafted features. The Transformer model is employed for both ISLR and CSLR. The training process involves using isolated sign videos, where hand keypoint features extracted from the input video are enriched using the Transformer model. Subsequently, these enriched features are forwarded to the final classification layer. The trained model, coupled with a post-processing method, is then applied to detect isolated sign boundaries within continuous sign videos. The evaluation of our model is conducted on two distinct datasets, including both continuous signs and their corresponding isolated signs, demonstrates promising results.
title	A Transformer Model for Boundary Detection in Continuous Sign Language
topic	Computer Vision and Pattern Recognition
url	https://arxiv.org/abs/2402.14720

Similar Items