Surface EMG-Based Inter-Session/Inter-Subject Gesture Recognition by Leveraging Lightweight All-ConvNet and Transfer Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Islam, Md. Rabiul, Massicotte, Daniel, Massicotte, Philippe Y., Zhu, Wei-Ping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
COVID-19 Diagnosis from Cough Acoustics using ConvNets and Data Augmentation
von: Mahanta, Saranga Kingkor, et al.
Veröffentlicht: (2021)
von: Mahanta, Saranga Kingkor, et al.
Veröffentlicht: (2021)
Study on Inter and Intra Speaker Variability in Speaker Recognition
von: Okhotnikov, Anton, et al.
Veröffentlicht: (2024)
von: Okhotnikov, Anton, et al.
Veröffentlicht: (2024)
InterBiasing: Boost Unseen Word Recognition through Biasing Intermediate Predictions
von: Nakagome, Yu, et al.
Veröffentlicht: (2024)
von: Nakagome, Yu, et al.
Veröffentlicht: (2024)
Confidence-Based Self-Training for EMG-to-Speech: Leveraging Synthetic EMG for Robust Modeling
von: Chen, Xiaodan, et al.
Veröffentlicht: (2025)
von: Chen, Xiaodan, et al.
Veröffentlicht: (2025)
Deep Filter Estimation from Inter-Frame Correlations for Monaural Speech Dereverberation
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2026)
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2026)
SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
Inter-Speaker Relative Cues for Two-Stage Text-Guided Target Speech Extraction
von: Dai, Wang, et al.
Veröffentlicht: (2026)
von: Dai, Wang, et al.
Veröffentlicht: (2026)
Speech Representation Analysis based on Inter- and Intra-Model Similarities
von: Kheir, Yassine El, et al.
Veröffentlicht: (2024)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2024)
WiRD-Gest: Gesture Recognition In The Real World Using Range-Doppler Wi-Fi Sensing on COTS Hardware
von: Sanson, Jessica, et al.
Veröffentlicht: (2026)
von: Sanson, Jessica, et al.
Veröffentlicht: (2026)
Exploring an Inter-Pausal Unit (IPU) based Approach for Indic End-to-End TTS Systems
von: Prakash, Anusha, et al.
Veröffentlicht: (2024)
von: Prakash, Anusha, et al.
Veröffentlicht: (2024)
ERVQ: Enhanced Residual Vector Quantization with Intra-and-Inter-Codebook Optimization for Neural Audio Codecs
von: Zheng, Rui-Chen, et al.
Veröffentlicht: (2024)
von: Zheng, Rui-Chen, et al.
Veröffentlicht: (2024)
Inter-Speaker Relative Cues for Text-Guided Target Speech Extraction
von: Dai, Wang, et al.
Veröffentlicht: (2025)
von: Dai, Wang, et al.
Veröffentlicht: (2025)
Improving Speech Enhancement by Integrating Inter-Channel and Band Features with Dual-branch Conformer
von: Li, Jizhen, et al.
Veröffentlicht: (2024)
von: Li, Jizhen, et al.
Veröffentlicht: (2024)
Towards EMG-to-Speech with a Necklace Form Factor
von: Wu, Peter, et al.
Veröffentlicht: (2024)
von: Wu, Peter, et al.
Veröffentlicht: (2024)
UL-UNAS: Ultra-Lightweight U-Nets for Real-Time Speech Enhancement via Network Architecture Search
von: Rong, Xiaobin, et al.
Veröffentlicht: (2025)
von: Rong, Xiaobin, et al.
Veröffentlicht: (2025)
Leveraging Content and Acoustic Representations for Speech Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
InterGridNet: An Electric Network Frequency Approach for Audio Source Location Classification Using Convolutional Neural Networks
von: Korgialas, Christos, et al.
Veröffentlicht: (2025)
von: Korgialas, Christos, et al.
Veröffentlicht: (2025)
SIM: Surface-based fMRI Analysis for Inter-Subject Multimodal Decoding from Movie-Watching Experiments
von: Dahan, Simon, et al.
Veröffentlicht: (2025)
von: Dahan, Simon, et al.
Veröffentlicht: (2025)
Articulatory Feature Prediction from Surface EMG during Speech Production
von: Lee, Jihwan, et al.
Veröffentlicht: (2025)
von: Lee, Jihwan, et al.
Veröffentlicht: (2025)
Affect Decoding in Phonated and Silent Speech Production from Surface EMG
von: Pistrosch, Simon, et al.
Veröffentlicht: (2026)
von: Pistrosch, Simon, et al.
Veröffentlicht: (2026)
Lightweight and Robust Multi-Channel End-to-End Speech Recognition with Spherical Harmonic Transform
von: Kong, Xiangzhu, et al.
Veröffentlicht: (2025)
von: Kong, Xiangzhu, et al.
Veröffentlicht: (2025)
Gesture-Aware Zero-Shot Speech Recognition for Patients with Language Disorders
von: Kim, Seungbae, et al.
Veröffentlicht: (2025)
von: Kim, Seungbae, et al.
Veröffentlicht: (2025)
Revisiting Modeling and Evaluation Approaches in Speech Emotion Recognition: Considering Subjectivity of Annotators and Ambiguity of Emotions
von: Chou, Huang-Cheng, et al.
Veröffentlicht: (2025)
von: Chou, Huang-Cheng, et al.
Veröffentlicht: (2025)
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
von: Jia, Zhenqi, et al.
Veröffentlicht: (2024)
von: Jia, Zhenqi, et al.
Veröffentlicht: (2024)
Multi-Level Embedding Conformer Framework for Bengali Automatic Speech Recognition
von: Sakib, Md. Nazmus, et al.
Veröffentlicht: (2025)
von: Sakib, Md. Nazmus, et al.
Veröffentlicht: (2025)
Leveraging ASR Pretrained Conformers for Speaker Verification through Transfer Learning and Knowledge Distillation
von: Cai, Danwei, et al.
Veröffentlicht: (2023)
von: Cai, Danwei, et al.
Veröffentlicht: (2023)
A Novel Transfer Learning Approach for Mental Stability Classification from Voice Signal
von: Islam, Rafiul, et al.
Veröffentlicht: (2026)
von: Islam, Rafiul, et al.
Veröffentlicht: (2026)
TF-CorrNet: Leveraging Spatial Correlation for Continuous Speech Separation
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2025)
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2025)
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
LPGNet: A Lightweight Network with Parallel Attention and Gated Fusion for Multimodal Emotion Recognition
von: He, Zhining, et al.
Veröffentlicht: (2025)
von: He, Zhining, et al.
Veröffentlicht: (2025)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
von: Poncelet, Jakob, et al.
Veröffentlicht: (2025)
von: Poncelet, Jakob, et al.
Veröffentlicht: (2025)
ConvDTW-ACS: Audio Segmentation for Track Type Detection During Car Manufacturing
von: López-Chilet, Álvaro, et al.
Veröffentlicht: (2024)
von: López-Chilet, Álvaro, et al.
Veröffentlicht: (2024)
Beyond Lips: Integrating Gesture and Lip Cues for Robust Audio-visual Speaker Extraction
von: Pan, Zexu, et al.
Veröffentlicht: (2026)
von: Pan, Zexu, et al.
Veröffentlicht: (2026)
AS-ASR: A Lightweight Framework for Aphasia-Specific Automatic Speech Recognition
von: Bao, Chen, et al.
Veröffentlicht: (2025)
von: Bao, Chen, et al.
Veröffentlicht: (2025)
TBDM-Net: Bidirectional Dense Networks with Gender Information for Speech Emotion Recognition
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
WaveNeXt 2: ConvNeXt-Based Fast Neural Vocoders With Residual Denoising and Sub-Modeling for GAN and Diffusion Models
von: Zhou, Wangzixi, et al.
Veröffentlicht: (2026)
von: Zhou, Wangzixi, et al.
Veröffentlicht: (2026)
Leveraging LLM for Stuttering Speech: A Unified Architecture Bridging Recognition and Event Detection
von: Huang, Shangkun, et al.
Veröffentlicht: (2025)
von: Huang, Shangkun, et al.
Veröffentlicht: (2025)
IITKGP-ABSP Submission to LRE22: Language Recognition in Low-Resource Settings
von: Dey, Spandan, et al.
Veröffentlicht: (2025)
von: Dey, Spandan, et al.
Veröffentlicht: (2025)
ISPA: Inter-Species Phonetic Alphabet for Transcribing Animal Sounds
von: Hagiwara, Masato, et al.
Veröffentlicht: (2024)
von: Hagiwara, Masato, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021) -
COVID-19 Diagnosis from Cough Acoustics using ConvNets and Data Augmentation
von: Mahanta, Saranga Kingkor, et al.
Veröffentlicht: (2021) -
Study on Inter and Intra Speaker Variability in Speaker Recognition
von: Okhotnikov, Anton, et al.
Veröffentlicht: (2024) -
InterBiasing: Boost Unseen Word Recognition through Biasing Intermediate Predictions
von: Nakagome, Yu, et al.
Veröffentlicht: (2024) -
Confidence-Based Self-Training for EMG-to-Speech: Leveraging Synthetic EMG for Robust Modeling
von: Chen, Xiaodan, et al.
Veröffentlicht: (2025)