Room Transfer Function Reconstruction Using Complex-valued Neural Networks and Irregularly Distributed Microphones
Fuente:
arXiv
Guardado en:
| Autores principales: | Ronchini, Francesca, Comanducci, Luca, Pezzoli, Mirco, Antonacci, Fabio, Sarti, Augusto |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Synthetic training set generation using text-to-audio models for environmental sound classification
por: Ronchini, Francesca, et al.
Publicado: (2024)
por: Ronchini, Francesca, et al.
Publicado: (2024)
Reconstruction of Sound Field through Diffusion Models
por: Miotello, Federico, et al.
Publicado: (2023)
por: Miotello, Federico, et al.
Publicado: (2023)
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
por: Miotello, Federico, et al.
Publicado: (2024)
por: Miotello, Federico, et al.
Publicado: (2024)
Implicit neural representation with physics-informed neural networks for the reconstruction of the early part of room impulse responses
por: Pezzoli, Mirco, et al.
Publicado: (2023)
por: Pezzoli, Mirco, et al.
Publicado: (2023)
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
por: Miotello, Federico, et al.
Publicado: (2024)
por: Miotello, Federico, et al.
Publicado: (2024)
AI-Assisted Music Production: A User Study on Text-to-Music Models
por: Ronchini, Francesca, et al.
Publicado: (2025)
por: Ronchini, Francesca, et al.
Publicado: (2025)
Physics-Informed Transfer Learning for Data-Driven Sound Source Reconstruction in Near-Field Acoustic Holography
por: Luan, Xinmeng, et al.
Publicado: (2025)
por: Luan, Xinmeng, et al.
Publicado: (2025)
Mitigating data replication in text-to-audio generative diffusion models through anti-memorization guidance
por: Messina, Francisco, et al.
Publicado: (2025)
por: Messina, Francisco, et al.
Publicado: (2025)
Interpreting End-to-End Deep Learning Models for Speech Source Localization Using Layer-wise Relevance Propagation
por: Comanducci, Luca, et al.
Publicado: (2024)
por: Comanducci, Luca, et al.
Publicado: (2024)
Physics-Informed Neural Network-Driven Sparse Field Discretization Method for Near-Field Acoustic Holography
por: Luan, Xinmeng, et al.
Publicado: (2025)
por: Luan, Xinmeng, et al.
Publicado: (2025)
Low-Rank Adaptation of Deep Prior Neural Networks For Room Impulse Response Reconstruction
por: Pezzoli, Mirco, et al.
Publicado: (2025)
por: Pezzoli, Mirco, et al.
Publicado: (2025)
PAGURI: a user experience study of creative interaction with text-to-music models
por: Ronchini, Francesca, et al.
Publicado: (2024)
por: Ronchini, Francesca, et al.
Publicado: (2024)
Acoustic source localization in the spherical harmonics domain exploiting low-rank approximations
por: Cobos, Maximo, et al.
Publicado: (2023)
por: Cobos, Maximo, et al.
Publicado: (2023)
Towards HRTF Personalization using Denoising Diffusion Models
por: Sánchez, Juan Camilo Albarracín, et al.
Publicado: (2025)
por: Sánchez, Juan Camilo Albarracín, et al.
Publicado: (2025)
MambaFoley: Foley Sound Generation using Selective State-Space Models
por: Colombo, Marco Furio, et al.
Publicado: (2024)
por: Colombo, Marco Furio, et al.
Publicado: (2024)
Synthesis of Soundfields through Irregular Loudspeaker Arrays Based on Convolutional Neural Networks
por: Comanducci, Luca, et al.
Publicado: (2022)
por: Comanducci, Luca, et al.
Publicado: (2022)
A Zero-Shot Physics-Informed Dictionary Learning Approach for Sound Field Reconstruction
por: Damiano, Stefano, et al.
Publicado: (2024)
por: Damiano, Stefano, et al.
Publicado: (2024)
Mind the Prompt: Prompting Strategies in Audio Generations for Improving Sound Classification
por: Ronchini, Francesca, et al.
Publicado: (2025)
por: Ronchini, Francesca, et al.
Publicado: (2025)
BRUDEX Database: Binaural Room Impulse Responses with Uniformly Distributed External Microphones
por: Fejgin, Daniel, et al.
Publicado: (2023)
por: Fejgin, Daniel, et al.
Publicado: (2023)
DiffusionRIR: Room Impulse Response Interpolation using Diffusion Models
por: Della Torre, Sagi, et al.
Publicado: (2025)
por: Della Torre, Sagi, et al.
Publicado: (2025)
VR-PTOLEMAIC: A Virtual Environment for the Perceptual Testing of Spatial Audio Algorithms
por: Ostan, Paolo, et al.
Publicado: (2025)
por: Ostan, Paolo, et al.
Publicado: (2025)
Diffused Responsibility: Analyzing the Energy Consumption of Generative Text-to-Audio Diffusion Models
por: Passoni, Riccardo, et al.
Publicado: (2025)
por: Passoni, Riccardo, et al.
Publicado: (2025)
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
Physics-Informed Neural Network for Volumetric Sound field Reconstruction of Speech Signals
por: Olivieri, Marco, et al.
Publicado: (2024)
por: Olivieri, Marco, et al.
Publicado: (2024)
Phase-Retrieval-Based Physics-Informed Neural Networks For Acoustic Magnitude Field Reconstruction
por: Schrader, Karl, et al.
Publicado: (2026)
por: Schrader, Karl, et al.
Publicado: (2026)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
por: Brümann, Klaus, et al.
Publicado: (2024)
por: Brümann, Klaus, et al.
Publicado: (2024)
EchoScan: Scanning Complex Room Geometries via Acoustic Echoes
por: Yeon, Inmo, et al.
Publicado: (2023)
por: Yeon, Inmo, et al.
Publicado: (2023)
Exploiting an External Microphone for Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
por: Fejgin, Daniel, et al.
Publicado: (2023)
por: Fejgin, Daniel, et al.
Publicado: (2023)
HiRIS: an Airborne Sonar Sensor with a 1024 Channel Microphone Array for In-Air Acoustic Imaging
por: Laurijssen, Dennis, et al.
Publicado: (2024)
por: Laurijssen, Dennis, et al.
Publicado: (2024)
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
por: Gupta, Shubham, et al.
Publicado: (2024)
por: Gupta, Shubham, et al.
Publicado: (2024)
Low-Complexity Neural Wind Noise Reduction for Audio Recordings
por: Eftekhari, Hesam, et al.
Publicado: (2025)
por: Eftekhari, Hesam, et al.
Publicado: (2025)
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
por: Fejgin, Daniel, et al.
Publicado: (2025)
por: Fejgin, Daniel, et al.
Publicado: (2025)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
por: Yu, Chin-Yun, et al.
Publicado: (2024)
por: Yu, Chin-Yun, et al.
Publicado: (2024)
Acoustivision Pro: An Open-Source Interactive Platform for Room Impulse Response Analysis and Acoustic Characterization
por: Goswami, Mandip
Publicado: (2026)
por: Goswami, Mandip
Publicado: (2026)
Listenable Maps for Zero-Shot Audio Classifiers
por: Paissan, Francesco, et al.
Publicado: (2024)
por: Paissan, Francesco, et al.
Publicado: (2024)
Differentiable Acoustic Radiance Transfer
por: Lee, Sungho, et al.
Publicado: (2025)
por: Lee, Sungho, et al.
Publicado: (2025)
Tool Wear Prediction in CNC Turning Operations using Ultrasonic Microphone Arrays and CNNs
por: Steckel, Jan, et al.
Publicado: (2024)
por: Steckel, Jan, et al.
Publicado: (2024)
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
por: Ronchini, Francesca, et al.
Publicado: (2023)
por: Ronchini, Francesca, et al.
Publicado: (2023)
FakeMusicCaps: a Dataset for Detection and Attribution of Synthetic Music Generated via Text-to-Music Models
por: Comanducci, Luca, et al.
Publicado: (2024)
por: Comanducci, Luca, et al.
Publicado: (2024)
Enhancing Anti-spoofing Countermeasures Robustness through Joint Optimization and Transfer Learning
por: Wang, Yikang, et al.
Publicado: (2024)
por: Wang, Yikang, et al.
Publicado: (2024)
Ejemplares similares
-
Synthetic training set generation using text-to-audio models for environmental sound classification
por: Ronchini, Francesca, et al.
Publicado: (2024) -
Reconstruction of Sound Field through Diffusion Models
por: Miotello, Federico, et al.
Publicado: (2023) -
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
por: Miotello, Federico, et al.
Publicado: (2024) -
Implicit neural representation with physics-informed neural networks for the reconstruction of the early part of room impulse responses
por: Pezzoli, Mirco, et al.
Publicado: (2023) -
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
por: Miotello, Federico, et al.
Publicado: (2024)