Noise-Robust Hearing Aid Voice Control
Fuente:
arXiv
Guardado en:
| Autores principales: | López-Espejo, Iván, Roselló, Eros, Edraki, Amin, Harte, Naomi, Jensen, Jesper |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Tiny Noise-Robust Voice Activity Detector for Voice Assistants
por: Asl, Hamed Jafarzadeh, et al.
Publicado: (2025)
por: Asl, Hamed Jafarzadeh, et al.
Publicado: (2025)
Visual Cues Support Robust Turn-taking Prediction in Noise
por: Russell, Sam O'Connor, et al.
Publicado: (2025)
por: Russell, Sam O'Connor, et al.
Publicado: (2025)
On Speech Pre-emphasis as a Simple and Inexpensive Method to Boost Speech Enhancement
por: López-Espejo, Iván, et al.
Publicado: (2024)
por: López-Espejo, Iván, et al.
Publicado: (2024)
Interpreting the Role of Visemes in Audio-Visual Speech Recognition
por: Papadopoulos, Aristeidis, et al.
Publicado: (2025)
por: Papadopoulos, Aristeidis, et al.
Publicado: (2025)
ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge
por: Martín-Doñas, Juan M., et al.
Publicado: (2024)
por: Martín-Doñas, Juan M., et al.
Publicado: (2024)
Uncovering the Visual Contribution in Audio-Visual Speech Recognition
por: Lin, Zhaofeng, et al.
Publicado: (2024)
por: Lin, Zhaofeng, et al.
Publicado: (2024)
PolySinger: Singing-Voice to Singing-Voice Translation from English to Japanese
por: Antonisen, Silas, et al.
Publicado: (2024)
por: Antonisen, Silas, et al.
Publicado: (2024)
VisG AV-HuBERT: Viseme-Guided AV-HuBERT
por: Papadopoulos, Aristeidis, et al.
Publicado: (2026)
por: Papadopoulos, Aristeidis, et al.
Publicado: (2026)
Audio-Visual Feature Synchronization for Robust Speech Enhancement in Hearing Aids
por: Saleem, Nasir, et al.
Publicado: (2025)
por: Saleem, Nasir, et al.
Publicado: (2025)
Advances in Intelligent Hearing Aids: Deep Learning Approaches to Selective Noise Cancellation
por: Khan, Haris, et al.
Publicado: (2025)
por: Khan, Haris, et al.
Publicado: (2025)
Sound Zone Control Robust To Sound Speed Change
por: Bhattacharjee, Sankha Subhra, et al.
Publicado: (2024)
por: Bhattacharjee, Sankha Subhra, et al.
Publicado: (2024)
Hearing-Loss Compensation Using Deep Neural Networks: A Framework and Results From a Listening Test
por: Leer, Peter, et al.
Publicado: (2024)
por: Leer, Peter, et al.
Publicado: (2024)
Online Single-Channel Audio-Based Sound Speed Estimation for Robust Multi-Channel Audio Control
por: Fuglsig, Andreas Jonas, et al.
Publicado: (2026)
por: Fuglsig, Andreas Jonas, et al.
Publicado: (2026)
Robust Fixed-Filter Sound Zone Control with Audio-Based Position Tracking
por: Bhattacharjee, Sankha Subhra, et al.
Publicado: (2024)
por: Bhattacharjee, Sankha Subhra, et al.
Publicado: (2024)
Binaural Localization Model for Speech in Noise
por: Tokala, Vikas, et al.
Publicado: (2025)
por: Tokala, Vikas, et al.
Publicado: (2025)
A Multi-stage Low-latency Enhancement System for Hearing Aids
por: Ouyang, Chengwei, et al.
Publicado: (2025)
por: Ouyang, Chengwei, et al.
Publicado: (2025)
Noise-Robust Target-Speaker Voice Activity Detection Through Self-Supervised Pretraining
por: Bovbjerg, Holger Severin, et al.
Publicado: (2025)
por: Bovbjerg, Holger Severin, et al.
Publicado: (2025)
Using Speech Foundational Models in Loss Functions for Hearing Aid Speech Enhancement
por: Sutherland, Robert, et al.
Publicado: (2024)
por: Sutherland, Robert, et al.
Publicado: (2024)
Non-Intrusive Intelligibility Prediction for Hearing Aids: Recent Advances, Trends, and Challenges
por: Zezario, Ryandhimas E.
Publicado: (2025)
por: Zezario, Ryandhimas E.
Publicado: (2025)
Efficient Personalization of Amplification in Hearing Aids via Multi-band Bayesian Machine Learning
por: Ni, Aoxin, et al.
Publicado: (2024)
por: Ni, Aoxin, et al.
Publicado: (2024)
Towards Environmental Preference Based Speech Enhancement For Individualised Multi-Modal Hearing Aids
por: Kirton-Wingate, Jasper, et al.
Publicado: (2024)
por: Kirton-Wingate, Jasper, et al.
Publicado: (2024)
HyWA: Hypernetwork Weight Adapting Personalized Voice Activity Detection
por: Nejad, Mahsa Ghazvini, et al.
Publicado: (2025)
por: Nejad, Mahsa Ghazvini, et al.
Publicado: (2025)
Feature Importance across Domains for Improving Non-Intrusive Speech Intelligibility Prediction in Hearing Aids
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
Affine Modulation-based Audiogram Fusion Network for Joint Noise Reduction and Hearing Loss Compensation
por: Ni, Ye, et al.
Publicado: (2025)
por: Ni, Ye, et al.
Publicado: (2025)
Noise-Robust Voice Conversion by Conditional Denoising Training Using Latent Variables of Recording Quality and Environment
por: Igarashi, Takuto, et al.
Publicado: (2024)
por: Igarashi, Takuto, et al.
Publicado: (2024)
Multi-Speaker DOA Estimation in Binaural Hearing Aids using Deep Learning and Speaker Count Fusion
por: Jazaeri, Farnaz, et al.
Publicado: (2025)
por: Jazaeri, Farnaz, et al.
Publicado: (2025)
DiffVQE: Hybrid Diffusion Voice Quality Enhancement Under Acoustic Echo and Noise
por: Girao, Haljan Lugo, et al.
Publicado: (2026)
por: Girao, Haljan Lugo, et al.
Publicado: (2026)
Robust Speech Activity Detection in the Presence of Singing Voice
por: Grundhuber, Philipp, et al.
Publicado: (2025)
por: Grundhuber, Philipp, et al.
Publicado: (2025)
A Study on Zero-Shot Non-Intrusive Speech Intelligibility for Hearing Aids Using Large Language Models
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
Frame-Aligned Fusion of Canary and WavLM for Non-Intrusive Intelligibility Prediction of Hearing-Aid-Processed Speech
por: Nakazawa, Kazushi
Publicado: (2026)
por: Nakazawa, Kazushi
Publicado: (2026)
How to train your ears: Auditory-model emulation for large-dynamic-range inputs and mild-to-severe hearing losses
por: Leer, Peter, et al.
Publicado: (2024)
por: Leer, Peter, et al.
Publicado: (2024)
Domain-Agnostic Incremental Learning for Sound Classification. A DCASE 2026 Challenge task
por: Casciotti, Riccardo, et al.
Publicado: (2026)
por: Casciotti, Riccardo, et al.
Publicado: (2026)
LIWhiz: A Non-Intrusive Lyric Intelligibility Prediction System for the Cadenza Challenge
por: Shekar, Ram C. M. C., et al.
Publicado: (2025)
por: Shekar, Ram C. M. C., et al.
Publicado: (2025)
The ICASSP SP Cadenza Challenge: Music Demixing/Remixing for Hearing Aids
por: Dabike, Gerardo Roa, et al.
Publicado: (2023)
por: Dabike, Gerardo Roa, et al.
Publicado: (2023)
Deep low-latency joint speech transmission and enhancement over a gaussian channel
por: Bokaei, Mohammad, et al.
Publicado: (2024)
por: Bokaei, Mohammad, et al.
Publicado: (2024)
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation
por: Peng, Puyuan, et al.
Publicado: (2025)
por: Peng, Puyuan, et al.
Publicado: (2025)
Multi-Microphone Noise Data Augmentation for DNN-based Own Voice Reconstruction for Hearables in Noisy Environments
por: Ohlenbusch, Mattes, et al.
Publicado: (2023)
por: Ohlenbusch, Mattes, et al.
Publicado: (2023)
Evaluating Speech Enhancement Systems Through Listening Effort
por: Gelderblom, Femke B., et al.
Publicado: (2024)
por: Gelderblom, Femke B., et al.
Publicado: (2024)
Binaural Speech Enhancement Using Deep Complex Convolutional Transformer Networks
por: Tokala, Vikas, et al.
Publicado: (2024)
por: Tokala, Vikas, et al.
Publicado: (2024)
Head-steered channel selection method for hearing aid applications using remote microphones
por: Sathyapriyan, Vasudha, et al.
Publicado: (2025)
por: Sathyapriyan, Vasudha, et al.
Publicado: (2025)
Ejemplares similares
-
Tiny Noise-Robust Voice Activity Detector for Voice Assistants
por: Asl, Hamed Jafarzadeh, et al.
Publicado: (2025) -
Visual Cues Support Robust Turn-taking Prediction in Noise
por: Russell, Sam O'Connor, et al.
Publicado: (2025) -
On Speech Pre-emphasis as a Simple and Inexpensive Method to Boost Speech Enhancement
por: López-Espejo, Iván, et al.
Publicado: (2024) -
Interpreting the Role of Visemes in Audio-Visual Speech Recognition
por: Papadopoulos, Aristeidis, et al.
Publicado: (2025) -
ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge
por: Martín-Doñas, Juan M., et al.
Publicado: (2024)