Non-invasive electromyographic speech neuroprosthesis: a geometric perspective
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gowda, Harshavardhana T., Miller, Lee M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
emg2speech: Synthesizing speech from electromyography using self-supervised speech models
von: Gowda, Harshavardhana T., et al.
Veröffentlicht: (2025)
von: Gowda, Harshavardhana T., et al.
Veröffentlicht: (2025)
Data-driven grapheme-to-phoneme representations for a lexicon-free text-to-speech
von: Garg, Abhinav, et al.
Veröffentlicht: (2024)
von: Garg, Abhinav, et al.
Veröffentlicht: (2024)
Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2026)
von: Lee, Dongheon, et al.
Veröffentlicht: (2026)
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
von: Saon, George, et al.
Veröffentlicht: (2025)
von: Saon, George, et al.
Veröffentlicht: (2025)
SLM-S2ST: A multimodal language model for direct speech-to-speech translation
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
Assessing speech quality metrics for evaluation of neural audio codecs under clean speech conditions
von: Mack, Wolfgang, et al.
Veröffentlicht: (2025)
von: Mack, Wolfgang, et al.
Veröffentlicht: (2025)
From the perspective of perceptual speech quality: The robustness of frequency bands to noise
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
Predicting speech intelligibility in older adults for speech enhancement using the Gammachirp Envelope Similarity Index, GESI
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2025)
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2025)
Audio-conditioned phonemic and prosodic annotation for building text-to-speech models from unlabeled speech data
von: Shirahata, Yuma, et al.
Veröffentlicht: (2024)
von: Shirahata, Yuma, et al.
Veröffentlicht: (2024)
Distilling a speech and music encoder with task arithmetic
von: Ritter-Gutierrez, Fabian, et al.
Veröffentlicht: (2025)
von: Ritter-Gutierrez, Fabian, et al.
Veröffentlicht: (2025)
Good practices for evaluation of synthesized speech
von: Cooper, Erica, et al.
Veröffentlicht: (2025)
von: Cooper, Erica, et al.
Veröffentlicht: (2025)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
von: Li, Junjie, et al.
Veröffentlicht: (2024)
von: Li, Junjie, et al.
Veröffentlicht: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
mmWave-Whisper: Phone Call Eavesdropping and Transcription Using Millimeter-Wave Radar
von: Basak, Suryoday, et al.
Veröffentlicht: (2024)
von: Basak, Suryoday, et al.
Veröffentlicht: (2024)
Text-To-Speech with Chain-of-Details: modeling temporal dynamics in speech generation
von: Ma, Jianbo, et al.
Veröffentlicht: (2026)
von: Ma, Jianbo, et al.
Veröffentlicht: (2026)
Transcribe, Align and Segment: Creating speech datasets for low-resource languages
von: Sereda, Taras
Veröffentlicht: (2024)
von: Sereda, Taras
Veröffentlicht: (2024)
Towards the Synthesis of Non-speech Vocalizations
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
On the relationship between speech and hearing
von: Umesh, Srinivasan, et al.
Veröffentlicht: (2024)
von: Umesh, Srinivasan, et al.
Veröffentlicht: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
TokenSE: a Mamba-based discrete token speech enhancement framework for cochlear implants
von: Chiang, Hsin-Tien, et al.
Veröffentlicht: (2026)
von: Chiang, Hsin-Tien, et al.
Veröffentlicht: (2026)
Adversarial speech for voice privacy protection from Personalized Speech generation
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
Joint decoding method for controllable contextual speech recognition based on Speech LLM
von: Fang, Yangui, et al.
Veröffentlicht: (2025)
von: Fang, Yangui, et al.
Veröffentlicht: (2025)
MELA-TTS: Joint transformer-diffusion model with representation alignment for speech synthesis
von: An, Keyu, et al.
Veröffentlicht: (2025)
von: An, Keyu, et al.
Veröffentlicht: (2025)
Evaluating pretrained speech embedding systems for dysarthria detection across heterogenous datasets
von: Wihlborg, Lovisa, et al.
Veröffentlicht: (2025)
von: Wihlborg, Lovisa, et al.
Veröffentlicht: (2025)
Teaching the Teachers: Boosting unsupervised domain adaptation in speech recognition by ensemble update
von: Ahmad, Rehan, et al.
Veröffentlicht: (2026)
von: Ahmad, Rehan, et al.
Veröffentlicht: (2026)
Boosting Diffusion Model for Spectrogram Up-sampling in Text-to-speech: An Empirical Study
von: Zhang, Chong, et al.
Veröffentlicht: (2024)
von: Zhang, Chong, et al.
Veröffentlicht: (2024)
DBMIF: a deep balanced multimodal iterative fusion framework for air- and bone-conduction speech enhancement
von: Wu, Yilei, et al.
Veröffentlicht: (2026)
von: Wu, Yilei, et al.
Veröffentlicht: (2026)
TTS-CtrlNet: Time varying emotion aligned text-to-speech generation with ControlNet
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
Synthesizing speech with selected perceptual voice qualities - A case study with creaky voice
von: Rautenberg, Frederik, et al.
Veröffentlicht: (2025)
von: Rautenberg, Frederik, et al.
Veröffentlicht: (2025)
Spatially constrained vs. unconstrained filtering in neural spatiospectral filters for multichannel speech enhancement
von: Briegleb, Annika, et al.
Veröffentlicht: (2024)
von: Briegleb, Annika, et al.
Veröffentlicht: (2024)
BabAR: from phoneme recognition to developmental measures of young children's speech production
von: Lavechin, Marvin, et al.
Veröffentlicht: (2026)
von: Lavechin, Marvin, et al.
Veröffentlicht: (2026)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
Prominence-aware automatic speech recognition for conversational speech
von: Linke, Julian, et al.
Veröffentlicht: (2025)
von: Linke, Julian, et al.
Veröffentlicht: (2025)
Comparison of linear and nonlinear methods for decoding selective attention to speech from ear-EEG recordings
von: Thornton, Mike, et al.
Veröffentlicht: (2024)
von: Thornton, Mike, et al.
Veröffentlicht: (2024)
Efficient training strategies for natural sounding speech synthesis and speaker adaptation based on FastPitch
von: Răgman, Teodora, et al.
Veröffentlicht: (2024)
von: Răgman, Teodora, et al.
Veröffentlicht: (2024)
On the social bias of speech self-supervised models
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2024)
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2024)
MUSHRA-1S: A scalable and sensitive test approach for evaluating top-tier speech processing systems
von: Lechler, Laura, et al.
Veröffentlicht: (2025)
von: Lechler, Laura, et al.
Veröffentlicht: (2025)
Real-time speech enhancement in noise for throat microphone using neural audio codec as foundation model
von: Hauret, Julien, et al.
Veröffentlicht: (2025)
von: Hauret, Julien, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
emg2speech: Synthesizing speech from electromyography using self-supervised speech models
von: Gowda, Harshavardhana T., et al.
Veröffentlicht: (2025) -
Data-driven grapheme-to-phoneme representations for a lexicon-free text-to-speech
von: Garg, Abhinav, et al.
Veröffentlicht: (2024) -
Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2026) -
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
von: Huang, Ziling, et al.
Veröffentlicht: (2025) -
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
von: Saon, George, et al.
Veröffentlicht: (2025)