On the relationship between speech and hearing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Umesh, Srinivasan, Cohen, Leon, Nelson, Douglas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disentangling peripheral hearing loss from central and cognitive effects on speech intelligibility in older adults
von: Irino, Toshio, et al.
Veröffentlicht: (2025)
von: Irino, Toshio, et al.
Veröffentlicht: (2025)
Real-time multichannel deep speech enhancement in hearing aids: Comparing monaural and binaural processing in complex acoustic scenarios
von: Westhausen, Nils L., et al.
Veröffentlicht: (2024)
von: Westhausen, Nils L., et al.
Veröffentlicht: (2024)
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
Some clues to build a sound analysis relevant to hearing
von: Millot, Laurent
Veröffentlicht: (2024)
von: Millot, Laurent
Veröffentlicht: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
Towards Developing State-of-the-Art TTS Synthesisers for 13 Indian Languages with Signal Processing aided Alignments
von: Prakash, Anusha, et al.
Veröffentlicht: (2022)
von: Prakash, Anusha, et al.
Veröffentlicht: (2022)
Signal processing algorithm effective for sound quality of hearing loss simulators
von: Irino, Toshio, et al.
Veröffentlicht: (2024)
von: Irino, Toshio, et al.
Veröffentlicht: (2024)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
Exploring the anatomy of articulation rate in spontaneous English speech: relationships between utterance length effects and social factors
von: Tanner, James, et al.
Veröffentlicht: (2024)
von: Tanner, James, et al.
Veröffentlicht: (2024)
Do neonates hear what we measure? Assessing neonatal ward soundscapes at the neonates ears
von: Lam, Bhan, et al.
Veröffentlicht: (2025)
von: Lam, Bhan, et al.
Veröffentlicht: (2025)
Enhancing spatial hearing with cochlear implants: exploring the role of AI, multimodal interaction and perceptual training
von: Picinali, Lorenzo, et al.
Veröffentlicht: (2026)
von: Picinali, Lorenzo, et al.
Veröffentlicht: (2026)
Controllable joint noise reduction and hearing loss compensation using a differentiable auditory model
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2025)
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2025)
A dataset and model for auditory scene recognition for hearing devices: AHEAD-DS and OpenYAMNet
von: Zhong, Henry, et al.
Veröffentlicht: (2025)
von: Zhong, Henry, et al.
Veröffentlicht: (2025)
Probing mental health information in speech foundation models
von: de Gennes, Marc, et al.
Veröffentlicht: (2024)
von: de Gennes, Marc, et al.
Veröffentlicht: (2024)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
von: Li, Junjie, et al.
Veröffentlicht: (2024)
von: Li, Junjie, et al.
Veröffentlicht: (2024)
WhisperFlow: speech foundation models in real time
von: Wang, Rongxiang, et al.
Veröffentlicht: (2024)
von: Wang, Rongxiang, et al.
Veröffentlicht: (2024)
Distilling a speech and music encoder with task arithmetic
von: Ritter-Gutierrez, Fabian, et al.
Veröffentlicht: (2025)
von: Ritter-Gutierrez, Fabian, et al.
Veröffentlicht: (2025)
Unsupervised speech enhancement with spectral kurtosis and double deep priors
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
Omni-directional attention mechanism based on Mamba for speech separation
von: Xue, Ke, et al.
Veröffentlicht: (2026)
von: Xue, Ke, et al.
Veröffentlicht: (2026)
BFA: Real-time Multilingual Text-to-speech Forced Alignment
von: Rehman, Abdul, et al.
Veröffentlicht: (2025)
von: Rehman, Abdul, et al.
Veröffentlicht: (2025)
SPGM: Prioritizing Local Features for enhanced speech separation performance
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
Graph-based multi-Feature fusion method for speech emotion recognition
von: Liu, Xueyu, et al.
Veröffentlicht: (2024)
von: Liu, Xueyu, et al.
Veröffentlicht: (2024)
A lightweight and robust method for blind wideband-to-fullband extension of speech
von: Büthe, Jan, et al.
Veröffentlicht: (2024)
von: Büthe, Jan, et al.
Veröffentlicht: (2024)
Monaural speech enhancement on drone via Adapter based transfer learning
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
FreeCodec: A disentangled neural speech codec with fewer tokens
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
Prosodic Parameter Manipulation in TTS generated speech for Controlled Speech Generation
von: Chary, Podakanti Satyajith
Veröffentlicht: (2024)
von: Chary, Podakanti Satyajith
Veröffentlicht: (2024)
Adversarial speech for voice privacy protection from Personalized Speech generation
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
GDiffuSE: Diffusion-based speech enhancement with noise model guidance
von: Yanir, Efrayim, et al.
Veröffentlicht: (2025)
von: Yanir, Efrayim, et al.
Veröffentlicht: (2025)
Expressive paragraph text-to-speech synthesis with multi-step variational autoencoder
von: Li, Xuyuan, et al.
Veröffentlicht: (2023)
von: Li, Xuyuan, et al.
Veröffentlicht: (2023)
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
von: Pascu, Octavian, et al.
Veröffentlicht: (2023)
von: Pascu, Octavian, et al.
Veröffentlicht: (2023)
Enhancement by postfiltering for speech and audio coding in ad-hoc sensor networks
von: Das, Sneha, et al.
Veröffentlicht: (2020)
von: Das, Sneha, et al.
Veröffentlicht: (2020)
Building speech corpus with diverse voice characteristics for its prompt-based representation
von: Watanabe, Aya, et al.
Veröffentlicht: (2024)
von: Watanabe, Aya, et al.
Veröffentlicht: (2024)
Using RLHF to align speech enhancement approaches to mean-opinion quality scores
von: Kumar, Anurag, et al.
Veröffentlicht: (2024)
von: Kumar, Anurag, et al.
Veröffentlicht: (2024)
Multichannel blind speech source separation with a disjoint constraint source model
von: Wang, Jianyu, et al.
Veröffentlicht: (2024)
von: Wang, Jianyu, et al.
Veröffentlicht: (2024)
Towards a generalized monaural and binaural auditory model for psychoacoustics and speech intelligibility
von: Biberger, Thomas, et al.
Veröffentlicht: (2021)
von: Biberger, Thomas, et al.
Veröffentlicht: (2021)
Modeling strategies for speech enhancement in the latent space of a neural audio codec
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
DQR-TTS: Semi-supervised Text-to-speech Synthesis with Dynamic Quantized Representation
von: Wang, Jianzong, et al.
Veröffentlicht: (2023)
von: Wang, Jianzong, et al.
Veröffentlicht: (2023)
Language model integration based on memory control for sequence to sequence speech recognition
von: Cho, Jaejin, et al.
Veröffentlicht: (2018)
von: Cho, Jaejin, et al.
Veröffentlicht: (2018)
Ähnliche Einträge
-
Disentangling peripheral hearing loss from central and cognitive effects on speech intelligibility in older adults
von: Irino, Toshio, et al.
Veröffentlicht: (2025) -
Real-time multichannel deep speech enhancement in hearing aids: Comparing monaural and binaural processing in complex acoustic scenarios
von: Westhausen, Nils L., et al.
Veröffentlicht: (2024) -
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025) -
Some clues to build a sound analysis relevant to hearing
von: Millot, Laurent
Veröffentlicht: (2024) -
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)