Word-level Sign Language Recognition with Multi-stream Neural Networks Focusing on Local Regions and Skeletal Information
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Maruyama, Mizuki, Singh, Shrey, Inoue, Katsufumi, Roy, Partha Pratim, Iwamura, Masakazu, Yoshioka, Michifumi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing EEG Signal-Based Emotion Recognition with Synthetic Data: Diffusion Model Approach
von: Siddhad, Gourav, et al.
Veröffentlicht: (2024)
von: Siddhad, Gourav, et al.
Veröffentlicht: (2024)
Empirical Investigation of the Impact of Phase Information on Fault Diagnosis of Rotating Machinery
von: Nagahama, Hiroyoshi, et al.
Veröffentlicht: (2025)
von: Nagahama, Hiroyoshi, et al.
Veröffentlicht: (2025)
Awake at the Wheel: Enhancing Automotive Safety through EEG-Based Fatigue Detection
von: Siddhad, Gourav, et al.
Veröffentlicht: (2024)
von: Siddhad, Gourav, et al.
Veröffentlicht: (2024)
TARQ: Tail-Aware Reconstruction Quantization for Rare-Word Robust Automatic Speech Recognition
von: Wang, Xinyu, et al.
Veröffentlicht: (2026)
von: Wang, Xinyu, et al.
Veröffentlicht: (2026)
Realistic Virtual Flood Experience System Using 360° Videos and 3D City Models Constructed from Building Footprints
von: Banno, Tatsuro, et al.
Veröffentlicht: (2026)
von: Banno, Tatsuro, et al.
Veröffentlicht: (2026)
Building and Evaluating a Realistic Virtual World for Large Scale Urban Exploration from 360° Videos
von: Takenawa, Mizuki, et al.
Veröffentlicht: (2025)
von: Takenawa, Mizuki, et al.
Veröffentlicht: (2025)
360CityGML: Realistic and Interactive Urban Visualization System Integrating CityGML Model and 360° Videos
von: Banno, Tatsuro, et al.
Veröffentlicht: (2025)
von: Banno, Tatsuro, et al.
Veröffentlicht: (2025)
Hierarchical Sub-action Tree for Continuous Sign Language Recognition
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
Manipulated Regions Localization For Partially Deepfake Audio: A Survey
von: He, Jiayi, et al.
Veröffentlicht: (2025)
von: He, Jiayi, et al.
Veröffentlicht: (2025)
Emotional Cues Extraction and Fusion for Multi-modal Emotion Prediction and Recognition in Conversation
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
Enabling American Sign Language Communication Under Low Data Rates
von: Santhalingam, Panneer Selvam, et al.
Veröffentlicht: (2025)
von: Santhalingam, Panneer Selvam, et al.
Veröffentlicht: (2025)
FPGA‐Based Deep Neural Network Implementation for Handwritten Digit Recognition
von: Matej Štajnbrikner, et al.
Veröffentlicht: (2025)
von: Matej Štajnbrikner, et al.
Veröffentlicht: (2025)
MAR3: Multi-Agent Recognition, Reasoning, and Reflection for Reference Audio-Visual Segmentation
von: Zhao, Yuan, et al.
Veröffentlicht: (2026)
von: Zhao, Yuan, et al.
Veröffentlicht: (2026)
Continuous Sign Language Recognition System using Deep Learning with MediaPipe Holistic
von: Srivastava, Sharvani, et al.
Veröffentlicht: (2024)
von: Srivastava, Sharvani, et al.
Veröffentlicht: (2024)
Towards Real-World Stickers Use: A New Dataset for Multi-Tag Sticker Recognition
von: Wang, Bingbing, et al.
Veröffentlicht: (2024)
von: Wang, Bingbing, et al.
Veröffentlicht: (2024)
Cross-domain Multi-step Thinking: Zero-shot Fine-grained Traffic Sign Recognition in the Wild
von: Gan, Yaozong, et al.
Veröffentlicht: (2024)
von: Gan, Yaozong, et al.
Veröffentlicht: (2024)
Text2Sign Diffusion: A Generative Approach for Gloss-Free Sign Language Production
von: Feng, Liqian, et al.
Veröffentlicht: (2025)
von: Feng, Liqian, et al.
Veröffentlicht: (2025)
DBDH: A Dual-Branch Dual-Head Neural Network for Invisible Embedded Regions Localization
von: Zhao, Chengxin, et al.
Veröffentlicht: (2024)
von: Zhao, Chengxin, et al.
Veröffentlicht: (2024)
A Comprehensive Review of Sign Language Recognition: Different Types, Modalities, and Datasets
von: Madhiarasan, M., et al.
Veröffentlicht: (2022)
von: Madhiarasan, M., et al.
Veröffentlicht: (2022)
CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
Dynamic resolution switching for live streaming
von: Xiong, Xin, et al.
Veröffentlicht: (2026)
von: Xiong, Xin, et al.
Veröffentlicht: (2026)
Efficient Prompt Tuning for Hierarchical Ingredient Recognition
von: Gui, Yinxuan, et al.
Veröffentlicht: (2025)
von: Gui, Yinxuan, et al.
Veröffentlicht: (2025)
Multimodal Emotion Recognition with Large Language Models
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
Sign-IDD: Iconicity Disentangled Diffusion for Sign Language Production
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
High-level Codes and Fine-grained Weights for Online Multi-modal Hashing Retrieval
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2024)
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2024)
Cross-domain Few-shot In-context Learning for Enhancing Traffic Sign Recognition
von: Gan, Yaozong, et al.
Veröffentlicht: (2024)
von: Gan, Yaozong, et al.
Veröffentlicht: (2024)
Modality-Aware Contrastive and Uncertainty-Regularized Emotion Recognition
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
Angle-Optimized Partial Disentanglement for Multimodal Emotion Recognition in Conversation
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
VCEMO: Multi-Modal Emotion Recognition for Chinese Voiceprints
von: Tang, Jinghua, et al.
Veröffentlicht: (2024)
von: Tang, Jinghua, et al.
Veröffentlicht: (2024)
Hallucination Localization in Video Captioning
von: Nakada, Shota, et al.
Veröffentlicht: (2025)
von: Nakada, Shota, et al.
Veröffentlicht: (2025)
Dialogue Understandability: Why are we streaming movies with subtitles?
von: Martinez, Helard Becerra, et al.
Veröffentlicht: (2024)
von: Martinez, Helard Becerra, et al.
Veröffentlicht: (2024)
Evolutionary Multimodal Reasoning via Hierarchical Semantic Representation for Intent Recognition
von: Zhou, Qianrui, et al.
Veröffentlicht: (2026)
von: Zhou, Qianrui, et al.
Veröffentlicht: (2026)
Orthogonal Disentanglement with Projected Feature Alignment for Multimodal Emotion Recognition in Conversation
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
SpikEmo: Enhancing Emotion Recognition With Spiking Temporal Dynamics in Conversations
von: Yu, Xiaomin, et al.
Veröffentlicht: (2024)
von: Yu, Xiaomin, et al.
Veröffentlicht: (2024)
Real-Time Word-Level Temporal Segmentation in Streaming Speech Recognition
von: Nishida, Naoto, et al.
Veröffentlicht: (2025)
von: Nishida, Naoto, et al.
Veröffentlicht: (2025)
Fine-grained Knowledge Graph-driven Video-Language Learning for Action Recognition
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
Multimodal Fusion via Hypergraph Autoencoder and Contrastive Learning for Emotion Recognition in Conversation
von: Yi, Zijian, et al.
Veröffentlicht: (2024)
von: Yi, Zijian, et al.
Veröffentlicht: (2024)
State-Anchored Complete-View Distillation for Robust Conversational Multimodal Emotion Recognition
von: Pan, Zhaoyan, et al.
Veröffentlicht: (2026)
von: Pan, Zhaoyan, et al.
Veröffentlicht: (2026)
Mitigating Multimodal Inconsistency via Cognitive Dual-Pathway Reasoning for Intent Recognition
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
Ada2I: Enhancing Modality Balance for Multimodal Conversational Emotion Recognition
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Enhancing EEG Signal-Based Emotion Recognition with Synthetic Data: Diffusion Model Approach
von: Siddhad, Gourav, et al.
Veröffentlicht: (2024) -
Empirical Investigation of the Impact of Phase Information on Fault Diagnosis of Rotating Machinery
von: Nagahama, Hiroyoshi, et al.
Veröffentlicht: (2025) -
Awake at the Wheel: Enhancing Automotive Safety through EEG-Based Fatigue Detection
von: Siddhad, Gourav, et al.
Veröffentlicht: (2024) -
TARQ: Tail-Aware Reconstruction Quantization for Rare-Word Robust Automatic Speech Recognition
von: Wang, Xinyu, et al.
Veröffentlicht: (2026) -
Realistic Virtual Flood Experience System Using 360° Videos and 3D City Models Constructed from Building Footprints
von: Banno, Tatsuro, et al.
Veröffentlicht: (2026)