Comparison of sEMG Encoding Accuracy Across Speech Modes Using Articulatory and Phoneme Features
Fuente:
arXiv
Guardado en:
| Autores principales: | Le, Chenqian, Li, Ruisi, Fumagalli, Beatrice, Esmaeili, Yasamin, Chen, Xupeng, Khalilian-Gourtani, Amirhossein, He, Tianyu, Flinker, Adeen, Wang, Yao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Machine Learning-Based Prediction of Speech Arrest During Direct Cortical Stimulation Mapping
por: Emami, Nikasadat, et al.
Publicado: (2025)
por: Emami, Nikasadat, et al.
Publicado: (2025)
GroupCDL: Interpretable Denoising and Compressed Sensing MRI via Learned Group-Sparsity and Circulant Attention
por: Janjusevic, Nikola, et al.
Publicado: (2024)
por: Janjusevic, Nikola, et al.
Publicado: (2024)
Articulatory Feature Prediction from Surface EMG during Speech Production
por: Lee, Jihwan, et al.
Publicado: (2025)
por: Lee, Jihwan, et al.
Publicado: (2025)
Speaker- and Text-Independent Estimation of Articulatory Movements and Phoneme Alignments from Speech
por: Weise, Tobias, et al.
Publicado: (2024)
por: Weise, Tobias, et al.
Publicado: (2024)
Evaluating Multichannel Speech Enhancement Algorithms at the Phoneme Scale Across Genders
por: Monir, Nasser-Eddine, et al.
Publicado: (2025)
por: Monir, Nasser-Eddine, et al.
Publicado: (2025)
EMG-to-Speech with Fewer Channels
por: Hwang, Injune, et al.
Publicado: (2026)
por: Hwang, Injune, et al.
Publicado: (2026)
Interpretable Modeling of Articulatory Temporal Dynamics from real-time MRI for Phoneme Recognition
por: Park, Jay, et al.
Publicado: (2025)
por: Park, Jay, et al.
Publicado: (2025)
Deep Speech Synthesis from Multimodal Articulatory Representations
por: Wu, Peter, et al.
Publicado: (2024)
por: Wu, Peter, et al.
Publicado: (2024)
Speech Emotion Recognition with Phonation Excitation Information and Articulatory Kinematics
por: Zhang, Ziqian, et al.
Publicado: (2025)
por: Zhang, Ziqian, et al.
Publicado: (2025)
Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection
por: Xue, Jun, et al.
Publicado: (2026)
por: Xue, Jun, et al.
Publicado: (2026)
Probing Human Articulatory Constraints in End-to-End TTS with Reverse and Mismatched Speech-Text Directions
por: Khadse, Parth, et al.
Publicado: (2026)
por: Khadse, Parth, et al.
Publicado: (2026)
Phoneme-Level Feature Discrepancies: A Key to Detecting Sophisticated Speech Deepfakes
por: Zhang, Kuiyuan, et al.
Publicado: (2024)
por: Zhang, Kuiyuan, et al.
Publicado: (2024)
Gabor is Enough: Interpretable Deep Denoising with a Gabor Synthesis Dictionary Prior
por: Janjušević, Nikola, et al.
Publicado: (2022)
por: Janjušević, Nikola, et al.
Publicado: (2022)
CDLNet: Noise-Adaptive Convolutional Dictionary Learning Network for Blind Denoising and Demosaicing
por: Janjušević, Nikola, et al.
Publicado: (2021)
por: Janjušević, Nikola, et al.
Publicado: (2021)
Prosody Labeling with Phoneme-BERT and Speech Foundation Models
por: Koriyama, Tomoki
Publicado: (2025)
por: Koriyama, Tomoki
Publicado: (2025)
PhonemeDF: A Synthetic Speech Dataset for Audio Deepfake Detection and Naturalness Evaluation
por: Nallaguntla, Vamshi, et al.
Publicado: (2026)
por: Nallaguntla, Vamshi, et al.
Publicado: (2026)
UTI-LLM: A Personalized Articulatory-Speech Therapy Assistance System Based on Multimodal Large Language Model
por: Yang, Yudong, et al.
Publicado: (2025)
por: Yang, Yudong, et al.
Publicado: (2025)
On the Relationship between Accent Strength and Articulatory Features
por: Huang, Kevin, et al.
Publicado: (2025)
por: Huang, Kevin, et al.
Publicado: (2025)
A Phoneme-Scale Assessment of Multichannel Speech Enhancement Algorithms
por: Monir, Nasser-Eddine, et al.
Publicado: (2024)
por: Monir, Nasser-Eddine, et al.
Publicado: (2024)
Phoneme-Level Analysis for Person-of-Interest Speech Deepfake Detection
por: Salvi, Davide, et al.
Publicado: (2025)
por: Salvi, Davide, et al.
Publicado: (2025)
Speech Rhythm-Based Speaker Embeddings Extraction from Phonemes and Phoneme Duration for Multi-Speaker Speech Synthesis
por: Fujita, Kenichi, et al.
Publicado: (2024)
por: Fujita, Kenichi, et al.
Publicado: (2024)
MRI2Speech: Speech Synthesis from Articulatory Movements Recorded by Real-time MRI
por: Shah, Neil, et al.
Publicado: (2024)
por: Shah, Neil, et al.
Publicado: (2024)
FabasedVC: Enhancing Voice Conversion with Text Modality Fusion and Phoneme-Level SSL Features
por: Wang, Wenyu, et al.
Publicado: (2025)
por: Wang, Wenyu, et al.
Publicado: (2025)
A Scalable Pipeline for Estimating Verb Frame Frequencies Using Large Language Models
por: Morgan, Adam M., et al.
Publicado: (2025)
por: Morgan, Adam M., et al.
Publicado: (2025)
Confidence-Based Self-Training for EMG-to-Speech: Leveraging Synthetic EMG for Robust Modeling
por: Chen, Xiaodan, et al.
Publicado: (2025)
por: Chen, Xiaodan, et al.
Publicado: (2025)
Articulation-Informed ASR: Integrating Articulatory Features into ASR via Auxiliary Speech Inversion and Cross-Attention Fusion
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
ELF: Encoding Speaker-Specific Latent Speech Feature for Speech Synthesis
por: Kong, Jungil, et al.
Publicado: (2023)
por: Kong, Jungil, et al.
Publicado: (2023)
Tracking Articulatory Dynamics in Speech with a Fixed-Weight BiLSTM-CNN Architecture
por: Pillai, Leena G, et al.
Publicado: (2025)
por: Pillai, Leena G, et al.
Publicado: (2025)
Acoustic to Articulatory Inversion of Speech; Data Driven Approaches, Challenges, Applications, and Future Scope
por: Pillai, Leena G, et al.
Publicado: (2025)
por: Pillai, Leena G, et al.
Publicado: (2025)
VAE-based Phoneme Alignment Using Gradient Annealing and SSL Acoustic Features
por: Koriyama, Tomoki
Publicado: (2024)
por: Koriyama, Tomoki
Publicado: (2024)
Multitask Learning for Grapheme-to-Phoneme Conversion of Anglicisms in German Speech Recognition
por: Pritzen, Julia, et al.
Publicado: (2021)
por: Pritzen, Julia, et al.
Publicado: (2021)
Phonikud: Hebrew Grapheme-to-Phoneme Conversion for Real-Time Text-to-Speech
por: Kolani, Yakov, et al.
Publicado: (2025)
por: Kolani, Yakov, et al.
Publicado: (2025)
Articulatory strategy as a source of variation in acoustic vowel dynamics
por: Strycharczuk, Patrycja, et al.
Publicado: (2026)
por: Strycharczuk, Patrycja, et al.
Publicado: (2026)
GTR-Voice: Articulatory Phonetics Informed Controllable Expressive Speech Synthesis
por: Li, Zehua Kcriss, et al.
Publicado: (2024)
por: Li, Zehua Kcriss, et al.
Publicado: (2024)
Empowering Global Voices: A Data-Efficient, Phoneme-Tone Adaptive Approach to High-Fidelity Speech Synthesis
por: Geng, Yizhong, et al.
Publicado: (2025)
por: Geng, Yizhong, et al.
Publicado: (2025)
DyPCL: Dynamic Phoneme-level Contrastive Learning for Dysarthric Speech Recognition
por: Lee, Wonjun, et al.
Publicado: (2025)
por: Lee, Wonjun, et al.
Publicado: (2025)
Self-Supervised Models for Phoneme Recognition: Applications in Children's Speech for Reading Learning
por: Medin, Lucas Block, et al.
Publicado: (2025)
por: Medin, Lucas Block, et al.
Publicado: (2025)
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
por: Ohnaka, Hien, et al.
Publicado: (2025)
por: Ohnaka, Hien, et al.
Publicado: (2025)
Frequency-Weighted Training Losses for Phoneme-Level DNN-based Speech Enhancement
por: Monir, Nasser-Eddine, et al.
Publicado: (2025)
por: Monir, Nasser-Eddine, et al.
Publicado: (2025)
DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units
por: Poli, Maxime, et al.
Publicado: (2026)
por: Poli, Maxime, et al.
Publicado: (2026)
Ejemplares similares
-
Machine Learning-Based Prediction of Speech Arrest During Direct Cortical Stimulation Mapping
por: Emami, Nikasadat, et al.
Publicado: (2025) -
GroupCDL: Interpretable Denoising and Compressed Sensing MRI via Learned Group-Sparsity and Circulant Attention
por: Janjusevic, Nikola, et al.
Publicado: (2024) -
Articulatory Feature Prediction from Surface EMG during Speech Production
por: Lee, Jihwan, et al.
Publicado: (2025) -
Speaker- and Text-Independent Estimation of Articulatory Movements and Phoneme Alignments from Speech
por: Weise, Tobias, et al.
Publicado: (2024) -
Evaluating Multichannel Speech Enhancement Algorithms at the Phoneme Scale Across Genders
por: Monir, Nasser-Eddine, et al.
Publicado: (2025)