Generative Data Augmentation Challenge: Synthesis of Room Acoustics for Speaker Distance Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Jackie, Götz, Georg, Llopis, Hermes Sampedro, Hafsteinsson, Haukur, Guðjónsson, Steinar, Nielsen, Daniel Gert, Pind, Finnur, Smaragdis, Paris, Manocha, Dinesh, Hershey, John, Kristjansson, Trausti, Kim, Minje |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Room-acoustic simulations as an alternative to measurements for audio-algorithm evaluation
by: Götz, Georg, et al.
Published: (2025)
by: Götz, Georg, et al.
Published: (2025)
Stable Model Reduction for Time‐Domain Room Acoustics: A Structure‐Preserving Formulation for Complex Boundaries
by: Satish Bonthu, et al.
Published: (2026)
by: Satish Bonthu, et al.
Published: (2026)
Generative Data Augmentation Challenge: Zero-Shot Speech Synthesis for Personalized Speech Enhancement
by: Bae, Jae-Sung, et al.
Published: (2025)
by: Bae, Jae-Sung, et al.
Published: (2025)
Reduced basis methods for numerical room acoustic simulations with parametrized boundaries
by: Llopis, Hermes Sampedro, et al.
Published: (2021)
by: Llopis, Hermes Sampedro, et al.
Published: (2021)
Gencho: Room Impulse Response Generation from Reverberant Speech and Text via Diffusion Transformers
by: Lin, Jackie, et al.
Published: (2026)
by: Lin, Jackie, et al.
Published: (2026)
Combolutional Neural Networks
by: Churchwell, Cameron, et al.
Published: (2025)
by: Churchwell, Cameron, et al.
Published: (2025)
User-guided Generative Source Separation
by: Wen, Yutong, et al.
Published: (2025)
by: Wen, Yutong, et al.
Published: (2025)
Deep Room Impulse Response Completion
by: Lin, Jackie, et al.
Published: (2024)
by: Lin, Jackie, et al.
Published: (2024)
Adaptive Slimming for Scalable and Efficient Speech Enhancement
by: Miccini, Riccardo, et al.
Published: (2025)
by: Miccini, Riccardo, et al.
Published: (2025)
Re-Bottleneck: Latent Re-Structuring for Neural Audio Autoencoders
by: Bralios, Dimitrios, et al.
Published: (2025)
by: Bralios, Dimitrios, et al.
Published: (2025)
Learning to Upsample and Upmix Audio in the Latent Domain
by: Bralios, Dimitrios, et al.
Published: (2025)
by: Bralios, Dimitrios, et al.
Published: (2025)
Adaptive Deterministic Flow Matching for Target Speaker Extraction
by: Hsieh, Tsun-An, et al.
Published: (2025)
by: Hsieh, Tsun-An, et al.
Published: (2025)
Continuous-Time Signal Decomposition: An Implicit Neural Generalization of PCA and ICA
by: Azmoodeh, Shayan K., et al.
Published: (2025)
by: Azmoodeh, Shayan K., et al.
Published: (2025)
Scaling Up Adaptive Filter Optimizers
by: Casebeer, Jonah, et al.
Published: (2024)
by: Casebeer, Jonah, et al.
Published: (2024)
Diagnosing Neural Convergence with Topological Alignment Spectra
by: Tavares, Tiago F., et al.
Published: (2024)
by: Tavares, Tiago F., et al.
Published: (2024)
TGIF: Talker Group-Informed Familiarization of Target Speaker Extraction
by: Hsieh, Tsun-An, et al.
Published: (2025)
by: Hsieh, Tsun-An, et al.
Published: (2025)
Hyperbolic Distance-Based Speech Separation
by: Petermann, Darius, et al.
Published: (2024)
by: Petermann, Darius, et al.
Published: (2024)
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
by: Tavares, Tiago, et al.
Published: (2024)
by: Tavares, Tiago, et al.
Published: (2024)
Rethinking Non-Negative Matrix Factorization with Implicit Neural Representations
by: Subramani, Krishna, et al.
Published: (2024)
by: Subramani, Krishna, et al.
Published: (2024)
AV-RIR: Audio-Visual Room Impulse Response Estimation
by: Ratnarajah, Anton, et al.
Published: (2023)
by: Ratnarajah, Anton, et al.
Published: (2023)
A stable decoupled perfectly matched layer for the 3D wave equation using the nodal discontinuous Galerkin method
by: Feriani, Sophia Julia, et al.
Published: (2024)
by: Feriani, Sophia Julia, et al.
Published: (2024)
Physicochemical properties and fish occurrence in 35 Icelandic lakes housing Arctic charr (Salvelinus alpinus)
by: Kristjánsson, Bjarni K, et al.
Published: (2011)
by: Kristjánsson, Bjarni K, et al.
Published: (2011)
(Table 1) Physicochemical properties of 35 Icelandic lakes housing a single morph of Arctic charr (Salvelinus alpinus)
by: Kristjánsson, Bjarni K, et al.
Published: (2011)
by: Kristjánsson, Bjarni K, et al.
Published: (2011)
(Table 2) Water origin, bed rock type and fish occurrence of 35 Icelandic lakes
by: Kristjánsson, Bjarni K, et al.
Published: (2011)
by: Kristjánsson, Bjarni K, et al.
Published: (2011)
Sound Source Separation Using Latent Variational Block-Wise Disentanglement
by: Helwani, Karim, et al.
Published: (2024)
by: Helwani, Karim, et al.
Published: (2024)
PromptSep: Generative Audio Separation via Multimodal Prompting
by: Wen, Yutong, et al.
Published: (2025)
by: Wen, Yutong, et al.
Published: (2025)
Uncovering the Representation Geometry of Minimal Cores in Overcomplete Reasoning Traces
by: Chowdhury, Sanjoy, et al.
Published: (2026)
by: Chowdhury, Sanjoy, et al.
Published: (2026)
PACE: Data-Driven Virtual Agent Interaction in Dense and Cluttered Environments
by: Mullen, James, et al.
Published: (2023)
by: Mullen, James, et al.
Published: (2023)
Listen2Scene: Interactive material-aware binaural sound propagation for reconstructed 3D scenes
by: Ratnarajah, Anton, et al.
Published: (2023)
by: Ratnarajah, Anton, et al.
Published: (2023)
Inst4DGS: Instance-Decomposed 4D Gaussian Splatting with Multi-Video Label Permutation Learning
by: Lee, Yonghan, et al.
Published: (2026)
by: Lee, Yonghan, et al.
Published: (2026)
EM-GANSim: Real-time and Accurate EM Simulation Using Conditional GANs for 3D Indoor Scenes
by: Wang, Ruichen, et al.
Published: (2024)
by: Wang, Ruichen, et al.
Published: (2024)
EH-MAM: Easy-to-Hard Masked Acoustic Modeling for Self-Supervised Speech Representation Learning
by: Seth, Ashish, et al.
Published: (2024)
by: Seth, Ashish, et al.
Published: (2024)
On Randomness in Agentic Evals
by: Bjarnason, Bjarni Haukur, et al.
Published: (2026)
by: Bjarnason, Bjarni Haukur, et al.
Published: (2026)
MemCtrl: Using MLLMs as Active Memory Controllers on Embodied Agents
by: Dorbala, Vishnu Sashank, et al.
Published: (2026)
by: Dorbala, Vishnu Sashank, et al.
Published: (2026)
SLAT-Phys: Fast Material Property Field Prediction from Structured 3D Latents
by: Das, Rocktim Jyoti, et al.
Published: (2026)
by: Das, Rocktim Jyoti, et al.
Published: (2026)
Noise-Robust DSP-Assisted Neural Pitch Estimation with Very Low Complexity
by: Subramani, Krishna, et al.
Published: (2023)
by: Subramani, Krishna, et al.
Published: (2023)
Every projective Oka manifold is elliptic
by: Forstneric, Franc, et al.
Published: (2025)
by: Forstneric, Franc, et al.
Published: (2025)
Regular immersions directed by algebraically elliptic cones
by: Alarcon, Antonio, et al.
Published: (2023)
by: Alarcon, Antonio, et al.
Published: (2023)
A strong parametric h-principle for complete minimal surfaces
by: Alarcon, Antonio, et al.
Published: (2021)
by: Alarcon, Antonio, et al.
Published: (2021)
Dynamics of generic automorphisms of Stein manifolds with the density property
by: Arosio, Leandro, et al.
Published: (2023)
by: Arosio, Leandro, et al.
Published: (2023)
Similar Items
-
Room-acoustic simulations as an alternative to measurements for audio-algorithm evaluation
by: Götz, Georg, et al.
Published: (2025) -
Stable Model Reduction for Time‐Domain Room Acoustics: A Structure‐Preserving Formulation for Complex Boundaries
by: Satish Bonthu, et al.
Published: (2026) -
Generative Data Augmentation Challenge: Zero-Shot Speech Synthesis for Personalized Speech Enhancement
by: Bae, Jae-Sung, et al.
Published: (2025) -
Reduced basis methods for numerical room acoustic simulations with parametrized boundaries
by: Llopis, Hermes Sampedro, et al.
Published: (2021) -
Gencho: Room Impulse Response Generation from Reverberant Speech and Text via Diffusion Transformers
by: Lin, Jackie, et al.
Published: (2026)