Exploring Frequency-Domain Feature Modeling for HRTF Magnitude Upsampling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xingyu, Bi, Hanwen, Ma, Fei, Zhao, Sipei, Cheng, Eva, Burnett, Ian S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Permutation-Invariant Physics-Informed Neural Network for Region-to-Region Sound Field Reconstruction
von: Chen, Xingyu, et al.
Veröffentlicht: (2026)
von: Chen, Xingyu, et al.
Veröffentlicht: (2026)
Sound Field Reconstruction Using a Compact Acoustics-informed Neural Network
von: Ma, Fei, et al.
Veröffentlicht: (2024)
von: Ma, Fei, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024)
A circular microphone array with virtual microphones based on acoustics-informed neural networks
von: Zhao, Sipei, et al.
Veröffentlicht: (2024)
von: Zhao, Sipei, et al.
Veröffentlicht: (2024)
On HRTF Notch Frequency Prediction Using Anthropometric Features and Neural Networks
von: Arbel, Lior, et al.
Veröffentlicht: (2024)
von: Arbel, Lior, et al.
Veröffentlicht: (2024)
Monaural speech enhancement on drone via Adapter based transfer learning
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
HRTFformer: A Spatially-Aware Transformer for Individual HRTF Upsampling in Immersive Audio Rendering
von: Hu, Xuyi, et al.
Veröffentlicht: (2025)
von: Hu, Xuyi, et al.
Veröffentlicht: (2025)
Towards HRTF Personalization using Denoising Diffusion Models
von: Sánchez, Juan Camilo Albarracín, et al.
Veröffentlicht: (2025)
von: Sánchez, Juan Camilo Albarracín, et al.
Veröffentlicht: (2025)
Towards Perception-Informed Latent HRTF Representations
von: Zhang, You, et al.
Veröffentlicht: (2025)
von: Zhang, You, et al.
Veröffentlicht: (2025)
Enhancing Photogrammetry Reconstruction For HRTF Synthesis Via A Graph Neural Network
von: Pirard, Ludovic, et al.
Veröffentlicht: (2025)
von: Pirard, Ludovic, et al.
Veröffentlicht: (2025)
Impact of HRTF individualisation and head movements in a real/virtual localisation task
von: Martin, Vincent, et al.
Veröffentlicht: (2025)
von: Martin, Vincent, et al.
Veröffentlicht: (2025)
Binaural Target Speaker Extraction using Individualized HRTF
von: Ellinson, Yoav, et al.
Veröffentlicht: (2025)
von: Ellinson, Yoav, et al.
Veröffentlicht: (2025)
The Extended SONICOM HRTF Dataset and Spatial Audio Metrics Toolbox
von: Poole, Katarina C., et al.
Veröffentlicht: (2025)
von: Poole, Katarina C., et al.
Veröffentlicht: (2025)
HRTF-guided Binaural Target Speaker Extraction with Real-World Validation
von: Ellinson, Yoav, et al.
Veröffentlicht: (2026)
von: Ellinson, Yoav, et al.
Veröffentlicht: (2026)
Learning to Upsample and Upmix Audio in the Latent Domain
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
Assessing the Potential Impact of Direction-Dependent HRTF Selection on Sound Localization Accuracy
von: Goldring, Sapir, et al.
Veröffentlicht: (2024)
von: Goldring, Sapir, et al.
Veröffentlicht: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
Magnitude and Phase-based Feature Fusion Using Co-attention Mechanism for Speaker recognition
von: Su, Rongfeng, et al.
Veröffentlicht: (2025)
von: Su, Rongfeng, et al.
Veröffentlicht: (2025)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
HRTF Estimation using a Score-based Prior
von: Thuillier, Etienne, et al.
Veröffentlicht: (2024)
von: Thuillier, Etienne, et al.
Veröffentlicht: (2024)
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
Target Speaker Extraction by Directly Exploiting Contextual Information in the Time-Frequency Domain
von: Yang, Xue, et al.
Veröffentlicht: (2024)
von: Yang, Xue, et al.
Veröffentlicht: (2024)
MP-SENet: A Speech Enhancement Model with Parallel Denoising of Magnitude and Phase Spectra
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2023)
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2023)
HiPPO: Exploring A Novel Hierarchical Pronunciation Assessment Approach for Spoken Languages
von: Yan, Bi-Cheng, et al.
Veröffentlicht: (2025)
von: Yan, Bi-Cheng, et al.
Veröffentlicht: (2025)
Feature Importance across Domains for Improving Non-Intrusive Speech Intelligibility Prediction in Hearing Aids
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2025)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2025)
FLAMO: An Open-Source Library for Frequency-Domain Differentiable Audio Processing
von: Santo, Gloria Dal, et al.
Veröffentlicht: (2024)
von: Santo, Gloria Dal, et al.
Veröffentlicht: (2024)
Point Neuron Learning: A New Physics-Informed Neural Network Architecture
von: Bi, Hanwen, et al.
Veröffentlicht: (2024)
von: Bi, Hanwen, et al.
Veröffentlicht: (2024)
STFTCodec: High-Fidelity Audio Compression through Time-Frequency Domain Representation
von: Feng, Tao, et al.
Veröffentlicht: (2025)
von: Feng, Tao, et al.
Veröffentlicht: (2025)
Self-Supervised Convolutional Audio Models are Flexible Acoustic Feature Learners: A Domain Specificity and Transfer-Learning Study
von: Ogg, Mattson
Veröffentlicht: (2025)
von: Ogg, Mattson
Veröffentlicht: (2025)
ConSep: a Noise- and Reverberation-Robust Speech Separation Framework by Magnitude Conditioning
von: Ho, Kuan-Hsun, et al.
Veröffentlicht: (2024)
von: Ho, Kuan-Hsun, et al.
Veröffentlicht: (2024)
Temporal-Frequency State Space Duality: An Efficient Paradigm for Speech Emotion Recognition
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2024)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2024)
Audio Inpainting in Time-Frequency Domain with Phase-Aware Prior
von: Balušík, Peter, et al.
Veröffentlicht: (2026)
von: Balušík, Peter, et al.
Veröffentlicht: (2026)
Revisiting the Privacy of Low-Frequency Speech Signals: Exploring Resampling Methods, Evaluation Scenarios, and Speaker Characteristics
von: Pohlhausen, Jule, et al.
Veröffentlicht: (2025)
von: Pohlhausen, Jule, et al.
Veröffentlicht: (2025)
Frequency Tracking Features for Data-Efficient Deep Siren Identification
von: Damiano, Stefano, et al.
Veröffentlicht: (2024)
von: Damiano, Stefano, et al.
Veröffentlicht: (2024)
Ambisonics Binaural Rendering via Masked Magnitude Least Squares
von: Berebi, Or, et al.
Veröffentlicht: (2025)
von: Berebi, Or, et al.
Veröffentlicht: (2025)
Learning Magnitude Distribution of Sound Fields via Conditioned Autoencoder
von: Koyama, Shoichi, et al.
Veröffentlicht: (2025)
von: Koyama, Shoichi, et al.
Veröffentlicht: (2025)
Frequency-Domain Sound Field from the Perspective of Band-Limited Functions
von: Iwami, Takahiro, et al.
Veröffentlicht: (2024)
von: Iwami, Takahiro, et al.
Veröffentlicht: (2024)
MuFFIN: Multifaceted Pronunciation Feedback Model with Interactive Hierarchical Neural Modeling
von: Yan, Bi-Cheng, et al.
Veröffentlicht: (2025)
von: Yan, Bi-Cheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Permutation-Invariant Physics-Informed Neural Network for Region-to-Region Sound Field Reconstruction
von: Chen, Xingyu, et al.
Veröffentlicht: (2026) -
Sound Field Reconstruction Using a Compact Acoustics-informed Neural Network
von: Ma, Fei, et al.
Veröffentlicht: (2024) -
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025) -
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024) -
A circular microphone array with virtual microphones based on acoustics-informed neural networks
von: Zhao, Sipei, et al.
Veröffentlicht: (2024)