Reconstruction of Sound Field through Diffusion Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Miotello, Federico, Comanducci, Luca, Pezzoli, Mirco, Bernardini, Alberto, Antonacci, Fabio, Sarti, Augusto |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Room Transfer Function Reconstruction Using Complex-valued Neural Networks and Irregularly Distributed Microphones
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
di: Miotello, Federico, et al.
Pubblicazione: (2024)
di: Miotello, Federico, et al.
Pubblicazione: (2024)
Implicit neural representation with physics-informed neural networks for the reconstruction of the early part of room impulse responses
di: Pezzoli, Mirco, et al.
Pubblicazione: (2023)
di: Pezzoli, Mirco, et al.
Pubblicazione: (2023)
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
di: Miotello, Federico, et al.
Pubblicazione: (2024)
di: Miotello, Federico, et al.
Pubblicazione: (2024)
A Zero-Shot Physics-Informed Dictionary Learning Approach for Sound Field Reconstruction
di: Damiano, Stefano, et al.
Pubblicazione: (2024)
di: Damiano, Stefano, et al.
Pubblicazione: (2024)
Physics-Informed Transfer Learning for Data-Driven Sound Source Reconstruction in Near-Field Acoustic Holography
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
Low-Rank Adaptation of Deep Prior Neural Networks For Room Impulse Response Reconstruction
di: Pezzoli, Mirco, et al.
Pubblicazione: (2025)
di: Pezzoli, Mirco, et al.
Pubblicazione: (2025)
Towards HRTF Personalization using Denoising Diffusion Models
di: Sánchez, Juan Camilo Albarracín, et al.
Pubblicazione: (2025)
di: Sánchez, Juan Camilo Albarracín, et al.
Pubblicazione: (2025)
Interpreting End-to-End Deep Learning Models for Speech Source Localization Using Layer-wise Relevance Propagation
di: Comanducci, Luca, et al.
Pubblicazione: (2024)
di: Comanducci, Luca, et al.
Pubblicazione: (2024)
Synthetic training set generation using text-to-audio models for environmental sound classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Physics-Informed Neural Network-Driven Sparse Field Discretization Method for Near-Field Acoustic Holography
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
di: Luan, Xinmeng, et al.
Pubblicazione: (2025)
AI-Assisted Music Production: A User Study on Text-to-Music Models
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
Mitigating data replication in text-to-audio generative diffusion models through anti-memorization guidance
di: Messina, Francisco, et al.
Pubblicazione: (2025)
di: Messina, Francisco, et al.
Pubblicazione: (2025)
Acoustic source localization in the spherical harmonics domain exploiting low-rank approximations
di: Cobos, Maximo, et al.
Pubblicazione: (2023)
di: Cobos, Maximo, et al.
Pubblicazione: (2023)
VR-PTOLEMAIC: A Virtual Environment for the Perceptual Testing of Spatial Audio Algorithms
di: Ostan, Paolo, et al.
Pubblicazione: (2025)
di: Ostan, Paolo, et al.
Pubblicazione: (2025)
MambaFoley: Foley Sound Generation using Selective State-Space Models
di: Colombo, Marco Furio, et al.
Pubblicazione: (2024)
di: Colombo, Marco Furio, et al.
Pubblicazione: (2024)
Synthesis of Soundfields through Irregular Loudspeaker Arrays Based on Convolutional Neural Networks
di: Comanducci, Luca, et al.
Pubblicazione: (2022)
di: Comanducci, Luca, et al.
Pubblicazione: (2022)
DiffusionRIR: Room Impulse Response Interpolation using Diffusion Models
di: Della Torre, Sagi, et al.
Pubblicazione: (2025)
di: Della Torre, Sagi, et al.
Pubblicazione: (2025)
PAGURI: a user experience study of creative interaction with text-to-music models
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Diffused Responsibility: Analyzing the Energy Consumption of Generative Text-to-Audio Diffusion Models
di: Passoni, Riccardo, et al.
Pubblicazione: (2025)
di: Passoni, Riccardo, et al.
Pubblicazione: (2025)
On the Extension of Differential Beamforming Theory to Arbitrary Planar Arrays of First-Order Elements
di: Miotello, Federico, et al.
Pubblicazione: (2025)
di: Miotello, Federico, et al.
Pubblicazione: (2025)
Physics-Informed Neural Network for Volumetric Sound field Reconstruction of Speech Signals
di: Olivieri, Marco, et al.
Pubblicazione: (2024)
di: Olivieri, Marco, et al.
Pubblicazione: (2024)
Mind the Prompt: Prompting Strategies in Audio Generations for Improving Sound Classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
Listenable Maps for Zero-Shot Audio Classifiers
di: Paissan, Francesco, et al.
Pubblicazione: (2024)
di: Paissan, Francesco, et al.
Pubblicazione: (2024)
Phase-Retrieval-Based Physics-Informed Neural Networks For Acoustic Magnitude Field Reconstruction
di: Schrader, Karl, et al.
Pubblicazione: (2026)
di: Schrader, Karl, et al.
Pubblicazione: (2026)
Physics-Informed Machine Learning For Sound Field Estimation
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
Resource-Efficient Separation Transformer
di: Della Libera, Luca, et al.
Pubblicazione: (2022)
di: Della Libera, Luca, et al.
Pubblicazione: (2022)
Listenable Maps for Audio Classifiers
di: Paissan, Francesco, et al.
Pubblicazione: (2024)
di: Paissan, Francesco, et al.
Pubblicazione: (2024)
Toward Deep Drum Source Separation
di: Mezza, Alessandro Ilic, et al.
Pubblicazione: (2023)
di: Mezza, Alessandro Ilic, et al.
Pubblicazione: (2023)
Speech Self-Supervised Representations Benchmarking: a Case for Larger Probing Heads
di: Zaiem, Salah, et al.
Pubblicazione: (2023)
di: Zaiem, Salah, et al.
Pubblicazione: (2023)
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
di: Gupta, Shubham, et al.
Pubblicazione: (2024)
di: Gupta, Shubham, et al.
Pubblicazione: (2024)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
di: Torabi, Yasaman, et al.
Pubblicazione: (2025)
di: Torabi, Yasaman, et al.
Pubblicazione: (2025)
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
di: Yao, Shengshi, et al.
Pubblicazione: (2025)
di: Yao, Shengshi, et al.
Pubblicazione: (2025)
Decomposing the Influence of Physical Acoustic Modeling on Neural Personal Sound Zone Rendering: An Ablation Study
di: Jiang, Hao, et al.
Pubblicazione: (2026)
di: Jiang, Hao, et al.
Pubblicazione: (2026)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
di: Diaz-Guerra, David, et al.
Pubblicazione: (2023)
di: Diaz-Guerra, David, et al.
Pubblicazione: (2023)
XAI-Driven Spectral Analysis of Cough Sounds for Respiratory Disease Characterization
di: Amado-Caballero, Patricia, et al.
Pubblicazione: (2025)
di: Amado-Caballero, Patricia, et al.
Pubblicazione: (2025)
GLA-Grad: A Griffin-Lim Extended Waveform Generation Diffusion Model
di: Liu, Haocheng, et al.
Pubblicazione: (2024)
di: Liu, Haocheng, et al.
Pubblicazione: (2024)
GLA-Grad++: An Improved Griffin-Lim Guided Diffusion Model for Speech Synthesis
di: Baoueb, Teysir, et al.
Pubblicazione: (2025)
di: Baoueb, Teysir, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Room Transfer Function Reconstruction Using Complex-valued Neural Networks and Irregularly Distributed Microphones
di: Ronchini, Francesca, et al.
Pubblicazione: (2024) -
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
di: Miotello, Federico, et al.
Pubblicazione: (2024) -
Implicit neural representation with physics-informed neural networks for the reconstruction of the early part of room impulse responses
di: Pezzoli, Mirco, et al.
Pubblicazione: (2023) -
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
di: Miotello, Federico, et al.
Pubblicazione: (2024) -
A Zero-Shot Physics-Informed Dictionary Learning Approach for Sound Field Reconstruction
di: Damiano, Stefano, et al.
Pubblicazione: (2024)