Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Chin-Yun, Pauwels, Johan, Fazekas, György |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Differentiable Time-Varying Linear Prediction in the Context of End-to-End Analysis-by-Synthesis
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2022)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2022)
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
Audio synthesizer inversion in symmetric parameter spaces with approximately equivariant flow matching
von: Hayes, Ben, et al.
Veröffentlicht: (2025)
von: Hayes, Ben, et al.
Veröffentlicht: (2025)
Sound Matching an Analogue Levelling Amplifier Using the Newton-Raphson Method
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
Accelerating Automatic Differentiation of Direct Form Digital Filters
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
Zero-Shot Duet Singing Voices Separation with Diffusion Models
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2025)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2025)
Leave-One-EquiVariant: Alleviating invariance-related information loss in contrastive music representations
von: Guinot, Julien, et al.
Veröffentlicht: (2024)
von: Guinot, Julien, et al.
Veröffentlicht: (2024)
Improving Inference-Time Optimisation for Vocal Effects Style Transfer with a Gaussian Prior
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
RIFT: Entropy-Optimised Fractional Wavelet Constellations for Ideal Time-Frequency Estimation
von: Cozens, James M., et al.
Veröffentlicht: (2025)
von: Cozens, James M., et al.
Veröffentlicht: (2025)
Analytical model for the relation between signal bandwidth and spatial resolution in Steered-Response Power Phase Transform (SRP-PHAT) maps
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
Differentiable Acoustic Radiance Transfer
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
Differentiable All-pole Filters for Time-varying Audio Systems
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Phase-Based Signal Representations for Scattering
von: Haider, Daniel, et al.
Veröffentlicht: (2022)
von: Haider, Daniel, et al.
Veröffentlicht: (2022)
DiffVox: A Differentiable Model for Capturing and Analysing Vocal Effects Distributions
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
Enhancing Anti-spoofing Countermeasures Robustness through Joint Optimization and Transfer Learning
von: Wang, Yikang, et al.
Veröffentlicht: (2024)
von: Wang, Yikang, et al.
Veröffentlicht: (2024)
SLAP: Siamese Language-Audio Pretraining Without Negative Samples for Music Understanding
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2024)
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2024)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
Comparison of Frequency-Fusion Mechanisms for Binaural Direction-of-Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
Exploiting an External Microphone for Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
An Investigation of Time-Frequency Representation Discriminators for High-Fidelity Vocoder
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
von: Yao, Shengshi, et al.
Veröffentlicht: (2025)
von: Yao, Shengshi, et al.
Veröffentlicht: (2025)
Time-domain sound field estimation using kernel ridge regression
von: Brunnström, Jesper, et al.
Veröffentlicht: (2025)
von: Brunnström, Jesper, et al.
Veröffentlicht: (2025)
LocaGen: Sub-Sample Time-Delay Learning for Beam Localization
von: Kunwar, Ishaan, et al.
Veröffentlicht: (2025)
von: Kunwar, Ishaan, et al.
Veröffentlicht: (2025)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
Blind Source Separation of Radar Signals in Time Domain Using Deep Learning
von: Hinderer, Sven
Veröffentlicht: (2025)
von: Hinderer, Sven
Veröffentlicht: (2025)
Note-Level Singing Melody Transcription for Time-Aligned Musical Score Generation
von: Kim, Leekyung, et al.
Veröffentlicht: (2025)
von: Kim, Leekyung, et al.
Veröffentlicht: (2025)
Wavelet-Based Time-Frequency Fingerprinting for Feature Extraction of Traditional Irish Music
von: Shore, Noah
Veröffentlicht: (2025)
von: Shore, Noah
Veröffentlicht: (2025)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement
von: Hao, Xiang, et al.
Veröffentlicht: (2020)
von: Hao, Xiang, et al.
Veröffentlicht: (2020)
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
The CARFAC v2 Cochlear Model in Matlab, NumPy, and JAX
von: Lyon, Richard F., et al.
Veröffentlicht: (2024)
von: Lyon, Richard F., et al.
Veröffentlicht: (2024)
A Study on Speech Assessment with Visual Cues
von: Ahmed, Shafique, et al.
Veröffentlicht: (2025)
von: Ahmed, Shafique, et al.
Veröffentlicht: (2025)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Differentiable Time-Varying Linear Prediction in the Context of End-to-End Analysis-by-Synthesis
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024) -
Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2022) -
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023) -
Audio synthesizer inversion in symmetric parameter spaces with approximately equivariant flow matching
von: Hayes, Ben, et al.
Veröffentlicht: (2025) -
Sound Matching an Analogue Levelling Amplifier Using the Newton-Raphson Method
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)