A Machine Learning Approach for Denoising and Upsampling HRTFs
Fuente:
arXiv
Guardado en:
| Autores principales: | Hu, Xuyi, Li, Jian, Picinali, Lorenzo, Hogg, Aidan O. T. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HRTFformer: A Spatially-Aware Transformer for Individual HRTF Upsampling in Immersive Audio Rendering
por: Hu, Xuyi, et al.
Publicado: (2025)
por: Hu, Xuyi, et al.
Publicado: (2025)
HRTF upsampling with a generative adversarial network using a gnomonic equiangular projection
por: Hogg, Aidan O. T., et al.
Publicado: (2023)
por: Hogg, Aidan O. T., et al.
Publicado: (2023)
Learning to Upsample and Upmix Audio in the Latent Domain
por: Bralios, Dimitrios, et al.
Publicado: (2025)
por: Bralios, Dimitrios, et al.
Publicado: (2025)
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
por: Miotello, Federico, et al.
Publicado: (2024)
por: Miotello, Federico, et al.
Publicado: (2024)
ADNAC: Audio Denoiser using Neural Audio Codec
por: Jimon, Daniel, et al.
Publicado: (2025)
por: Jimon, Daniel, et al.
Publicado: (2025)
Cross-Validated Cross-Channel Self-Attention and Denoising for Automatic Modulation Classification
por: Suman, Prakash, et al.
Publicado: (2026)
por: Suman, Prakash, et al.
Publicado: (2026)
Music Genre Classification Using Machine Learning Techniques
por: Mishra, Alokit, et al.
Publicado: (2025)
por: Mishra, Alokit, et al.
Publicado: (2025)
Clustering of Indonesian and Western Gamelan Orchestras through Machine Learning of Performance Parameters
por: Linke, Simon, et al.
Publicado: (2024)
por: Linke, Simon, et al.
Publicado: (2024)
Automatic Contextual Audio Denoising
por: Luong, Diep, et al.
Publicado: (2026)
por: Luong, Diep, et al.
Publicado: (2026)
Machine Learning Approaches to Vocal Register Classification in Contemporary Male Pop Music
por: Kim, Alexander, et al.
Publicado: (2025)
por: Kim, Alexander, et al.
Publicado: (2025)
CyIN: Cyclic Informative Latent Space for Bridging Complete and Incomplete Multimodal Learning
por: Lin, Ronghao, et al.
Publicado: (2026)
por: Lin, Ronghao, et al.
Publicado: (2026)
Privacy-Enhancing Infant Cry Classification with Federated Transformers and Denoising Regularization
por: Owino, Geofrey, et al.
Publicado: (2025)
por: Owino, Geofrey, et al.
Publicado: (2025)
Time Series Diffusion Method: A Denoising Diffusion Probabilistic Model for Vibration Signal Generation
por: Yi, Haiming, et al.
Publicado: (2023)
por: Yi, Haiming, et al.
Publicado: (2023)
Denoising by neural network for muzzle blast detection
por: Pujol, Hadrien, et al.
Publicado: (2025)
por: Pujol, Hadrien, et al.
Publicado: (2025)
When Denoising Hinders: Revisiting Zero-Shot ASR with SAM-Audio and Whisper
por: Islam, Akif, et al.
Publicado: (2026)
por: Islam, Akif, et al.
Publicado: (2026)
Voice-Driven Mortality Prediction in Hospitalized Heart Failure Patients: A Machine Learning Approach Enhanced with Diagnostic Biomarkers
por: Ahmadli, Nihat, et al.
Publicado: (2024)
por: Ahmadli, Nihat, et al.
Publicado: (2024)
Transformer Based Machine Fault Detection From Audio Input
por: Holla, Kiran Voderhobli
Publicado: (2026)
por: Holla, Kiran Voderhobli
Publicado: (2026)
Are Deep Speech Denoising Models Robust to Adversarial Noise?
por: Schwarzer, Will, et al.
Publicado: (2025)
por: Schwarzer, Will, et al.
Publicado: (2025)
A Framework for Evaluating Faithfulness in Explainable AI for Machine Anomalous Sound Detection Using Frequency-Band Perturbation
por: Buck, Alexander, et al.
Publicado: (2026)
por: Buck, Alexander, et al.
Publicado: (2026)
PACE: Pretrained Audio Continual Learning
por: Li, Chang, et al.
Publicado: (2026)
por: Li, Chang, et al.
Publicado: (2026)
End-to-End Efficiency in Keyword Spotting: A System-Level Approach for Embedded Microcontrollers
por: Bartoli, Pietro, et al.
Publicado: (2025)
por: Bartoli, Pietro, et al.
Publicado: (2025)
Uncertainty Quantification in Machine Learning for Joint Speaker Diarization and Identification
por: McKnight, Simon W., et al.
Publicado: (2023)
por: McKnight, Simon W., et al.
Publicado: (2023)
Sound and Music Biases in Deep Music Transcription Models: A Systematic Analysis
por: Marták, Lukáš Samuel, et al.
Publicado: (2025)
por: Marták, Lukáš Samuel, et al.
Publicado: (2025)
Improving BERT for Symbolic Music Understanding Using Token Denoising and Pianoroll Prediction
por: Wang, Jun-You, et al.
Publicado: (2025)
por: Wang, Jun-You, et al.
Publicado: (2025)
Knowledge Distillation for Speech Denoising by Latent Representation Alignment with Cosine Distance
por: Luong, Diep, et al.
Publicado: (2025)
por: Luong, Diep, et al.
Publicado: (2025)
A Novel Score-CAM based Denoiser for Spectrographic Signature Extraction without Ground Truth
por: Elias, Noel
Publicado: (2024)
por: Elias, Noel
Publicado: (2024)
Tri-MTL: A Triple Multitask Learning Approach for Respiratory Disease Diagnosis
por: Kim, June-Woo, et al.
Publicado: (2025)
por: Kim, June-Woo, et al.
Publicado: (2025)
Predicting Global HRTFs From Scanned Head Geometry Using Deep Learning and Compact Representations
por: Wang, Yuxiang, et al.
Publicado: (2022)
por: Wang, Yuxiang, et al.
Publicado: (2022)
A Recurrent Neural Network Approach to the Answering Machine Detection Problem
por: Altwlkany, Kemal, et al.
Publicado: (2024)
por: Altwlkany, Kemal, et al.
Publicado: (2024)
Unsupervised CP-UNet Framework for Denoising DAS Data with Decay Noise
por: Huang, Tianye, et al.
Publicado: (2025)
por: Huang, Tianye, et al.
Publicado: (2025)
A Multimodal Framework for Dementia Detection via Linguistic and Acoustic Representation Learning
por: Ilias, Loukas, et al.
Publicado: (2026)
por: Ilias, Loukas, et al.
Publicado: (2026)
LibriVAD: A Scalable Open Dataset with Deep Learning Benchmarks for Voice Activity Detection
por: Stylianou, Ioannis, et al.
Publicado: (2025)
por: Stylianou, Ioannis, et al.
Publicado: (2025)
SCRAPL: Scattering Transform with Random Paths for Machine Learning
por: Mitcheltree, Christopher, et al.
Publicado: (2026)
por: Mitcheltree, Christopher, et al.
Publicado: (2026)
Multi-Representation Attention Framework for Underwater Bioacoustic Denoising and Recognition
por: Razig, Amine, et al.
Publicado: (2025)
por: Razig, Amine, et al.
Publicado: (2025)
SpikCommander: A High-performance Spiking Transformer with Multi-view Learning for Efficient Speech Command Recognition
por: Wang, Jiaqi, et al.
Publicado: (2025)
por: Wang, Jiaqi, et al.
Publicado: (2025)
Descriptor-Injected Cross-Modal Learning: A Systematic Exploration of Audio-MIDI Alignment via Spectral and Melodic Features
por: Méndez, Mariano Fernández
Publicado: (2026)
por: Méndez, Mariano Fernández
Publicado: (2026)
Adversarial Training of Denoising Diffusion Model Using Dual Discriminators for High-Fidelity Multi-Speaker TTS
por: Ko, Myeongjin, et al.
Publicado: (2023)
por: Ko, Myeongjin, et al.
Publicado: (2023)
Voice Signal Processing for Machine Learning. The Case of Speaker Isolation
por: Ganchev, Radan
Publicado: (2024)
por: Ganchev, Radan
Publicado: (2024)
Prototypical Contrastive Learning For Improved Few-Shot Audio Classification
por: Sgouropoulos, Christos, et al.
Publicado: (2025)
por: Sgouropoulos, Christos, et al.
Publicado: (2025)
Acoustic and Machine Learning Methods for Speech-Based Suicide Risk Assessment: A Systematic Review
por: Marie, Ambre, et al.
Publicado: (2025)
por: Marie, Ambre, et al.
Publicado: (2025)
Ejemplares similares
-
HRTFformer: A Spatially-Aware Transformer for Individual HRTF Upsampling in Immersive Audio Rendering
por: Hu, Xuyi, et al.
Publicado: (2025) -
HRTF upsampling with a generative adversarial network using a gnomonic equiangular projection
por: Hogg, Aidan O. T., et al.
Publicado: (2023) -
Learning to Upsample and Upmix Audio in the Latent Domain
por: Bralios, Dimitrios, et al.
Publicado: (2025) -
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
por: Miotello, Federico, et al.
Publicado: (2024) -
ADNAC: Audio Denoiser using Neural Audio Codec
por: Jimon, Daniel, et al.
Publicado: (2025)