Some clues to build a sound analysis relevant to hearing
Fuente:
arXiv
Saved in:
| Main Author: | Millot, Laurent |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Signal processing algorithm effective for sound quality of hearing loss simulators
by: Irino, Toshio, et al.
Published: (2024)
by: Irino, Toshio, et al.
Published: (2024)
Listening broadband physical model for microphones: a first step
by: Millot, Laurent, et al.
Published: (2024)
by: Millot, Laurent, et al.
Published: (2024)
Using perceptive subbands analysis to perform audio scenes cartography
by: Millot, Laurent, et al.
Published: (2024)
by: Millot, Laurent, et al.
Published: (2024)
On the relationship between speech and hearing
by: Umesh, Srinivasan, et al.
Published: (2024)
by: Umesh, Srinivasan, et al.
Published: (2024)
What do MLLMs hear? Examining reasoning with text and sound components in Multimodal Large Language Models
by: Çoban, Enis Berk, et al.
Published: (2024)
by: Çoban, Enis Berk, et al.
Published: (2024)
Controllable joint noise reduction and hearing loss compensation using a differentiable auditory model
by: Gonzalez, Philippe, et al.
Published: (2025)
by: Gonzalez, Philippe, et al.
Published: (2025)
Do neonates hear what we measure? Assessing neonatal ward soundscapes at the neonates ears
by: Lam, Bhan, et al.
Published: (2025)
by: Lam, Bhan, et al.
Published: (2025)
Enhancing spatial hearing with cochlear implants: exploring the role of AI, multimodal interaction and perceptual training
by: Picinali, Lorenzo, et al.
Published: (2026)
by: Picinali, Lorenzo, et al.
Published: (2026)
Disentangling peripheral hearing loss from central and cognitive effects on speech intelligibility in older adults
by: Irino, Toshio, et al.
Published: (2025)
by: Irino, Toshio, et al.
Published: (2025)
A dataset and model for auditory scene recognition for hearing devices: AHEAD-DS and OpenYAMNet
by: Zhong, Henry, et al.
Published: (2025)
by: Zhong, Henry, et al.
Published: (2025)
Differentiable physics for sound field reconstruction
by: Verburg, Samuel A., et al.
Published: (2025)
by: Verburg, Samuel A., et al.
Published: (2025)
Frequency-aware convolution for sound event detection
by: Song, Tao, et al.
Published: (2024)
by: Song, Tao, et al.
Published: (2024)
Real-time multichannel deep speech enhancement in hearing aids: Comparing monaural and binaural processing in complex acoustic scenarios
by: Westhausen, Nils L., et al.
Published: (2024)
by: Westhausen, Nils L., et al.
Published: (2024)
The Neural-SRP method for positional sound source localization
by: Grinstein, Eric, et al.
Published: (2024)
by: Grinstein, Eric, et al.
Published: (2024)
Multispecies bird sound recognition using a fully convolutional neural network
by: García-Ordás, María Teresa, et al.
Published: (2024)
by: García-Ordás, María Teresa, et al.
Published: (2024)
RELATE: Subjective evaluation dataset for automatic evaluation of relevance between text and audio
by: Kanamori, Yusuke, et al.
Published: (2025)
by: Kanamori, Yusuke, et al.
Published: (2025)
Interaural time difference loss for binaural target sound extraction
by: Hernandez-Olivan, Carlos, et al.
Published: (2024)
by: Hernandez-Olivan, Carlos, et al.
Published: (2024)
Onset and offset weighted loss function for sound event detection
by: Song, Tao
Published: (2024)
by: Song, Tao
Published: (2024)
Fine-tune the pretrained ATST model for sound event detection
by: Shao, Nian, et al.
Published: (2023)
by: Shao, Nian, et al.
Published: (2023)
Binaural sound source localization using a hybrid time and frequency domain model
by: Geva, Gil, et al.
Published: (2024)
by: Geva, Gil, et al.
Published: (2024)
Representational learning for an anomalous sound detection system with source separation model
by: Shin, Seunghyeon, et al.
Published: (2024)
by: Shin, Seunghyeon, et al.
Published: (2024)
Full-frequency dynamic convolution: a physical frequency-dependent convolution for sound event detection
by: Yue, Haobo, et al.
Published: (2024)
by: Yue, Haobo, et al.
Published: (2024)
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
by: Ronchini, Francesca, et al.
Published: (2023)
by: Ronchini, Francesca, et al.
Published: (2023)
The role of direct sound spherical harmonics representation in externalization using binaural reproduction
by: Miller, Eran, et al.
Published: (2024)
by: Miller, Eran, et al.
Published: (2024)
InsectSet459: an open dataset of insect sounds for bioacoustic machine learning
by: Faiß, Marius, et al.
Published: (2025)
by: Faiß, Marius, et al.
Published: (2025)
Simi-SFX: A similarity-based conditioning method for controllable sound effect synthesis
by: Liu, Yunyi, et al.
Published: (2024)
by: Liu, Yunyi, et al.
Published: (2024)
Efficient learning-based sound propagation for virtual and real-world audio processing applications
by: Ratnarajah, Anton Jeran
Published: (2024)
by: Ratnarajah, Anton Jeran
Published: (2024)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
by: Deng, Qingkun, et al.
Published: (2024)
by: Deng, Qingkun, et al.
Published: (2024)
Adaptive high-precision sound source localization at low frequencies based on convolutional neural network
by: Ma, Wenbo, et al.
Published: (2024)
by: Ma, Wenbo, et al.
Published: (2024)
Text2Move: Text-to-moving sound generation via trajectory prediction and temporal alignment
by: Liu, Yunyi, et al.
Published: (2025)
by: Liu, Yunyi, et al.
Published: (2025)
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
by: Gao, Wenmiao, et al.
Published: (2025)
by: Gao, Wenmiao, et al.
Published: (2025)
Resnet-conformer network with shared weights and attention mechanism for sound event localization, detection, and distance estimation
by: Vo, Quoc Thinh, et al.
Published: (2025)
by: Vo, Quoc Thinh, et al.
Published: (2025)
Multizone sound field reproduction with direction-of-arrival-distribution-based regularization and its application to binaural-centered mode-matching
by: Matsuda, Ryo, et al.
Published: (2025)
by: Matsuda, Ryo, et al.
Published: (2025)
Serial-OE: Anomalous sound detection based on serial method with outlier exposure capable of using small amounts of anomalous data for training
by: Kuroyanagi, Ibuki, et al.
Published: (2025)
by: Kuroyanagi, Ibuki, et al.
Published: (2025)
Speech foundation models on intelligibility prediction for hearing-impaired listeners
by: Cuervo, Santiago, et al.
Published: (2024)
by: Cuervo, Santiago, et al.
Published: (2024)
Time-domain sound field estimation using kernel ridge regression
by: Brunnström, Jesper, et al.
Published: (2025)
by: Brunnström, Jesper, et al.
Published: (2025)
The first Cadenza challenges: using machine learning competitions to improve music for listeners with a hearing loss
by: Dabike, Gerardo Roa, et al.
Published: (2024)
by: Dabike, Gerardo Roa, et al.
Published: (2024)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
by: Garcia-Barrios, Guillermo, et al.
Published: (2024)
by: Garcia-Barrios, Guillermo, et al.
Published: (2024)
Synthetic training set generation using text-to-audio models for environmental sound classification
by: Ronchini, Francesca, et al.
Published: (2024)
by: Ronchini, Francesca, et al.
Published: (2024)
Human-mimetic binaural ear design and sound source direction estimation for task realization of musculoskeletal humanoids
by: Omura, Yusuke, et al.
Published: (2024)
by: Omura, Yusuke, et al.
Published: (2024)
Similar Items
-
Signal processing algorithm effective for sound quality of hearing loss simulators
by: Irino, Toshio, et al.
Published: (2024) -
Listening broadband physical model for microphones: a first step
by: Millot, Laurent, et al.
Published: (2024) -
Using perceptive subbands analysis to perform audio scenes cartography
by: Millot, Laurent, et al.
Published: (2024) -
On the relationship between speech and hearing
by: Umesh, Srinivasan, et al.
Published: (2024) -
What do MLLMs hear? Examining reasoning with text and sound components in Multimodal Large Language Models
by: Çoban, Enis Berk, et al.
Published: (2024)