Real-time multichannel deep speech enhancement in hearing aids: Comparing monaural and binaural processing in complex acoustic scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Westhausen, Nils L., Kayser, Hendrik, Jansen, Theresa, Meyer, Bernd T. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards a generalized monaural and binaural auditory model for psychoacoustics and speech intelligibility
von: Biberger, Thomas, et al.
Veröffentlicht: (2021)
von: Biberger, Thomas, et al.
Veröffentlicht: (2021)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
On the relationship between speech and hearing
von: Umesh, Srinivasan, et al.
Veröffentlicht: (2024)
von: Umesh, Srinivasan, et al.
Veröffentlicht: (2024)
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
von: Kim, Yunsik, et al.
Veröffentlicht: (2025)
von: Kim, Yunsik, et al.
Veröffentlicht: (2025)
End-to-end multi-channel speaker extraction and binaural speech synthesis
von: Chi, Cheng, et al.
Veröffentlicht: (2024)
von: Chi, Cheng, et al.
Veröffentlicht: (2024)
A tunable binaural audio telepresence system capable of balancing immersive and enhanced modes
von: Hsu, Yicheng, et al.
Veröffentlicht: (2024)
von: Hsu, Yicheng, et al.
Veröffentlicht: (2024)
Unsupervised speech enhancement with spectral kurtosis and double deep priors
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
A two-step approach for speech enhancement in low-SNR scenarios using cyclostationary beamforming and DNNs
von: Bologni, Giovanni, et al.
Veröffentlicht: (2026)
von: Bologni, Giovanni, et al.
Veröffentlicht: (2026)
Interaural time difference loss for binaural target sound extraction
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
Disentangling peripheral hearing loss from central and cognitive effects on speech intelligibility in older adults
von: Irino, Toshio, et al.
Veröffentlicht: (2025)
von: Irino, Toshio, et al.
Veröffentlicht: (2025)
The role of direct sound spherical harmonics representation in externalization using binaural reproduction
von: Miller, Eran, et al.
Veröffentlicht: (2024)
von: Miller, Eran, et al.
Veröffentlicht: (2024)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
Signal processing algorithm effective for sound quality of hearing loss simulators
von: Irino, Toshio, et al.
Veröffentlicht: (2024)
von: Irino, Toshio, et al.
Veröffentlicht: (2024)
SPGM: Prioritizing Local Features for enhanced speech separation performance
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
SonicSim: A customizable simulation platform for speech processing in moving sound source scenarios
von: Li, Kai, et al.
Veröffentlicht: (2024)
von: Li, Kai, et al.
Veröffentlicht: (2024)
Multizone sound field reproduction with direction-of-arrival-distribution-based regularization and its application to binaural-centered mode-matching
von: Matsuda, Ryo, et al.
Veröffentlicht: (2025)
von: Matsuda, Ryo, et al.
Veröffentlicht: (2025)
Robust DOA estimation using deep acoustic imaging
von: Roman, Adrian S., et al.
Veröffentlicht: (2024)
von: Roman, Adrian S., et al.
Veröffentlicht: (2024)
GDiffuSE: Diffusion-based speech enhancement with noise model guidance
von: Yanir, Efrayim, et al.
Veröffentlicht: (2025)
von: Yanir, Efrayim, et al.
Veröffentlicht: (2025)
Monaural speech enhancement on drone via Adapter based transfer learning
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
RaD-Net 2: A causal two-stage repairing and denoising speech enhancement network with knowledge distillation and complex axial self-attention
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
BFA: Real-time Multilingual Text-to-speech Forced Alignment
von: Rehman, Abdul, et al.
Veröffentlicht: (2025)
von: Rehman, Abdul, et al.
Veröffentlicht: (2025)
Modeling strategies for speech enhancement in the latent space of a neural audio codec
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
Using RLHF to align speech enhancement approaches to mean-opinion quality scores
von: Kumar, Anurag, et al.
Veröffentlicht: (2024)
von: Kumar, Anurag, et al.
Veröffentlicht: (2024)
Single-channel speech enhancement by using psychoacoustical model inspired fusion framework
von: Samui, Suman
Veröffentlicht: (2022)
von: Samui, Suman
Veröffentlicht: (2022)
The CHiME-7 UDASE task: Unsupervised domain adaptation for conversational speech enhancement
von: Leglaive, Simon, et al.
Veröffentlicht: (2023)
von: Leglaive, Simon, et al.
Veröffentlicht: (2023)
Deep low-latency joint speech transmission and enhancement over a gaussian channel
von: Bokaei, Mohammad, et al.
Veröffentlicht: (2024)
von: Bokaei, Mohammad, et al.
Veröffentlicht: (2024)
Configurable EBEN: Extreme Bandwidth Extension Network to enhance body-conducted speech capture
von: Hauret, Julien, et al.
Veröffentlicht: (2023)
von: Hauret, Julien, et al.
Veröffentlicht: (2023)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
von: Triantafyllopoulos, Andreas, et al.
Veröffentlicht: (2025)
von: Triantafyllopoulos, Andreas, et al.
Veröffentlicht: (2025)
Human-mimetic binaural ear design and sound source direction estimation for task realization of musculoskeletal humanoids
von: Omura, Yusuke, et al.
Veröffentlicht: (2024)
von: Omura, Yusuke, et al.
Veröffentlicht: (2024)
An adaptive filter bank based neural network approach for time delay estimation and speech enhancement
von: Ma, Lu
Veröffentlicht: (2025)
von: Ma, Lu
Veröffentlicht: (2025)
An automatic mixing speech enhancement system for multi-track audio
von: Liu, Xiaojing, et al.
Veröffentlicht: (2024)
von: Liu, Xiaojing, et al.
Veröffentlicht: (2024)
KS-Net: Multi-band joint speech restoration and enhancement network for 2024 ICASSP SSI Challenge
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
Automatic acoustic detection of birds through deep learning: the first Bird Audio Detection challenge
von: Stowell, Dan, et al.
Veröffentlicht: (2018)
von: Stowell, Dan, et al.
Veröffentlicht: (2018)
Some clues to build a sound analysis relevant to hearing
von: Millot, Laurent
Veröffentlicht: (2024)
von: Millot, Laurent
Veröffentlicht: (2024)
A unified multichannel far-field speech recognition system: combining neural beamforming with attention based end-to-end model
von: Zhao, Dongdi, et al.
Veröffentlicht: (2024)
von: Zhao, Dongdi, et al.
Veröffentlicht: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
Automated speech audiometry: Can it work using open-source pre-trained Kaldi-NL automatic speech recognition?
von: Araiza-Illan, Gloria, et al.
Veröffentlicht: (2023)
von: Araiza-Illan, Gloria, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Towards a generalized monaural and binaural auditory model for psychoacoustics and speech intelligibility
von: Biberger, Thomas, et al.
Veröffentlicht: (2021) -
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021) -
On the relationship between speech and hearing
von: Umesh, Srinivasan, et al.
Veröffentlicht: (2024) -
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
von: Kim, Yunsik, et al.
Veröffentlicht: (2025) -
End-to-end multi-channel speaker extraction and binaural speech synthesis
von: Chi, Cheng, et al.
Veröffentlicht: (2024)