Towards HRTF Personalization using Denoising Diffusion Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Sánchez, Juan Camilo Albarracín, Comanducci, Luca, Pezzoli, Mirco, Antonacci, Fabio |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reconstruction of Sound Field through Diffusion Models
di: Miotello, Federico, et al.
Pubblicazione: (2023)
di: Miotello, Federico, et al.
Pubblicazione: (2023)
Room Transfer Function Reconstruction Using Complex-valued Neural Networks and Irregularly Distributed Microphones
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Interpreting End-to-End Deep Learning Models for Speech Source Localization Using Layer-wise Relevance Propagation
di: Comanducci, Luca, et al.
Pubblicazione: (2024)
di: Comanducci, Luca, et al.
Pubblicazione: (2024)
DiffusionRIR: Room Impulse Response Interpolation using Diffusion Models
di: Della Torre, Sagi, et al.
Pubblicazione: (2025)
di: Della Torre, Sagi, et al.
Pubblicazione: (2025)
Synthetic training set generation using text-to-audio models for environmental sound classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
MambaFoley: Foley Sound Generation using Selective State-Space Models
di: Colombo, Marco Furio, et al.
Pubblicazione: (2024)
di: Colombo, Marco Furio, et al.
Pubblicazione: (2024)
Low-Rank Adaptation of Deep Prior Neural Networks For Room Impulse Response Reconstruction
di: Pezzoli, Mirco, et al.
Pubblicazione: (2025)
di: Pezzoli, Mirco, et al.
Pubblicazione: (2025)
Physics-Informed Transfer Learning for Data-Driven Sound Source Reconstruction in Near-Field Acoustic Holography
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
Acoustic source localization in the spherical harmonics domain exploiting low-rank approximations
di: Cobos, Maximo, et al.
Pubblicazione: (2023)
di: Cobos, Maximo, et al.
Pubblicazione: (2023)
PAGURI: a user experience study of creative interaction with text-to-music models
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Implicit neural representation with physics-informed neural networks for the reconstruction of the early part of room impulse responses
di: Pezzoli, Mirco, et al.
Pubblicazione: (2023)
di: Pezzoli, Mirco, et al.
Pubblicazione: (2023)
VR-PTOLEMAIC: A Virtual Environment for the Perceptual Testing of Spatial Audio Algorithms
di: Ostan, Paolo, et al.
Pubblicazione: (2025)
di: Ostan, Paolo, et al.
Pubblicazione: (2025)
AI-Assisted Music Production: A User Study on Text-to-Music Models
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
Diffused Responsibility: Analyzing the Energy Consumption of Generative Text-to-Audio Diffusion Models
di: Passoni, Riccardo, et al.
Pubblicazione: (2025)
di: Passoni, Riccardo, et al.
Pubblicazione: (2025)
Mitigating data replication in text-to-audio generative diffusion models through anti-memorization guidance
di: Messina, Francisco, et al.
Pubblicazione: (2025)
di: Messina, Francisco, et al.
Pubblicazione: (2025)
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
di: Miotello, Federico, et al.
Pubblicazione: (2024)
di: Miotello, Federico, et al.
Pubblicazione: (2024)
Towards Perception-Informed Latent HRTF Representations
di: Zhang, You, et al.
Pubblicazione: (2025)
di: Zhang, You, et al.
Pubblicazione: (2025)
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2024)
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2024)
FakeMusicCaps: a Dataset for Detection and Attribution of Synthetic Music Generated via Text-to-Music Models
di: Comanducci, Luca, et al.
Pubblicazione: (2024)
di: Comanducci, Luca, et al.
Pubblicazione: (2024)
Phase-Retrieval-Based Physics-Informed Neural Networks For Acoustic Magnitude Field Reconstruction
di: Schrader, Karl, et al.
Pubblicazione: (2026)
di: Schrader, Karl, et al.
Pubblicazione: (2026)
Binaural Target Speaker Extraction using Individualized HRTF
di: Ellinson, Yoav, et al.
Pubblicazione: (2025)
di: Ellinson, Yoav, et al.
Pubblicazione: (2025)
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
di: Miotello, Federico, et al.
Pubblicazione: (2024)
di: Miotello, Federico, et al.
Pubblicazione: (2024)
Synthesis of Soundfields through Irregular Loudspeaker Arrays Based on Convolutional Neural Networks
di: Comanducci, Luca, et al.
Pubblicazione: (2022)
di: Comanducci, Luca, et al.
Pubblicazione: (2022)
Physics-Informed Machine Learning For Sound Field Estimation
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
The Extended SONICOM HRTF Dataset and Spatial Audio Metrics Toolbox
di: Poole, Katarina C., et al.
Pubblicazione: (2025)
di: Poole, Katarina C., et al.
Pubblicazione: (2025)
HRTF-guided Binaural Target Speaker Extraction with Real-World Validation
di: Ellinson, Yoav, et al.
Pubblicazione: (2026)
di: Ellinson, Yoav, et al.
Pubblicazione: (2026)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
Assessing the Potential Impact of Direction-Dependent HRTF Selection on Sound Localization Accuracy
di: Goldring, Sapir, et al.
Pubblicazione: (2024)
di: Goldring, Sapir, et al.
Pubblicazione: (2024)
HRTF Estimation using a Score-based Prior
di: Thuillier, Etienne, et al.
Pubblicazione: (2024)
di: Thuillier, Etienne, et al.
Pubblicazione: (2024)
Physics-Informed Neural Network-Driven Sparse Field Discretization Method for Near-Field Acoustic Holography
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
di: Gayer, Yhonatan, et al.
Pubblicazione: (2025)
di: Gayer, Yhonatan, et al.
Pubblicazione: (2025)
Past, Present, and Future of Spatial Audio and Room Acoustics
di: Koyama, Shoichi, et al.
Pubblicazione: (2025)
di: Koyama, Shoichi, et al.
Pubblicazione: (2025)
Complex Image-Generative Diffusion Transformer for Audio Denoising
di: Li, Junhui, et al.
Pubblicazione: (2024)
di: Li, Junhui, et al.
Pubblicazione: (2024)
LL-SDR: Low-Latency Speech enhancement through Discrete Representations
di: Li, Jingyi, et al.
Pubblicazione: (2026)
di: Li, Jingyi, et al.
Pubblicazione: (2026)
Speech Denoising with Auditory Models
di: Saddler, Mark R., et al.
Pubblicazione: (2020)
di: Saddler, Mark R., et al.
Pubblicazione: (2020)
Dynamic Real-Time Ambisonics Order Adaptation for Immersive Networked Music Performances
di: Ostan, Paolo, et al.
Pubblicazione: (2025)
di: Ostan, Paolo, et al.
Pubblicazione: (2025)
HRTFformer: A Spatially-Aware Transformer for Individual HRTF Upsampling in Immersive Audio Rendering
di: Hu, Xuyi, et al.
Pubblicazione: (2025)
di: Hu, Xuyi, et al.
Pubblicazione: (2025)
Denoising of photogrammetric dummy head ear point clouds for individual Head-Related Transfer Functions computation
di: Di Giusto, Fabio, et al.
Pubblicazione: (2024)
di: Di Giusto, Fabio, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Reconstruction of Sound Field through Diffusion Models
di: Miotello, Federico, et al.
Pubblicazione: (2023) -
Room Transfer Function Reconstruction Using Complex-valued Neural Networks and Irregularly Distributed Microphones
di: Ronchini, Francesca, et al.
Pubblicazione: (2024) -
Interpreting End-to-End Deep Learning Models for Speech Source Localization Using Layer-wise Relevance Propagation
di: Comanducci, Luca, et al.
Pubblicazione: (2024) -
DiffusionRIR: Room Impulse Response Interpolation using Diffusion Models
di: Della Torre, Sagi, et al.
Pubblicazione: (2025) -
Synthetic training set generation using text-to-audio models for environmental sound classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)