Graph Neural Field with Spatial-Correlation Augmentation for HRTF Personalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, De, Hu, Junsheng, Jiang, Cuicui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024)
Towards HRTF Personalization using Denoising Diffusion Models
von: Sánchez, Juan Camilo Albarracín, et al.
Veröffentlicht: (2025)
von: Sánchez, Juan Camilo Albarracín, et al.
Veröffentlicht: (2025)
The Extended SONICOM HRTF Dataset and Spatial Audio Metrics Toolbox
von: Poole, Katarina C., et al.
Veröffentlicht: (2025)
von: Poole, Katarina C., et al.
Veröffentlicht: (2025)
HRTFformer: A Spatially-Aware Transformer for Individual HRTF Upsampling in Immersive Audio Rendering
von: Hu, Xuyi, et al.
Veröffentlicht: (2025)
von: Hu, Xuyi, et al.
Veröffentlicht: (2025)
Towards Perception-Informed Latent HRTF Representations
von: Zhang, You, et al.
Veröffentlicht: (2025)
von: Zhang, You, et al.
Veröffentlicht: (2025)
Binaural Target Speaker Extraction using Individualized HRTF
von: Ellinson, Yoav, et al.
Veröffentlicht: (2025)
von: Ellinson, Yoav, et al.
Veröffentlicht: (2025)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
HRTF-guided Binaural Target Speaker Extraction with Real-World Validation
von: Ellinson, Yoav, et al.
Veröffentlicht: (2026)
von: Ellinson, Yoav, et al.
Veröffentlicht: (2026)
HRTF Estimation using a Score-based Prior
von: Thuillier, Etienne, et al.
Veröffentlicht: (2024)
von: Thuillier, Etienne, et al.
Veröffentlicht: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
Assessing the Potential Impact of Direction-Dependent HRTF Selection on Sound Localization Accuracy
von: Goldring, Sapir, et al.
Veröffentlicht: (2024)
von: Goldring, Sapir, et al.
Veröffentlicht: (2024)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
Stereo Audio Rendering for Personal Sound Zones Using a Binaural Spatially Adaptive Neural Network (BSANN)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
persoDA: Personalized Data Augmentation for Personalized ASR
von: Parada, Pablo Peso, et al.
Veröffentlicht: (2025)
von: Parada, Pablo Peso, et al.
Veröffentlicht: (2025)
C2GA: A Class-Controllable Generative Augmentation Framework for Respiratory Sound Classification
von: Ma, Ziqi, et al.
Veröffentlicht: (2026)
von: Ma, Ziqi, et al.
Veröffentlicht: (2026)
SpatialCodec: Neural Spatial Speech Coding
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2023)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2023)
Personalized Neural Speech Codec
von: Jang, Inseon, et al.
Veröffentlicht: (2024)
von: Jang, Inseon, et al.
Veröffentlicht: (2024)
FOA Tokenizer: Low-bitrate Neural Codec for First Order Ambisonics with Spatial Consistency Loss
von: Sudarsanam, Parthasaarathy, et al.
Veröffentlicht: (2025)
von: Sudarsanam, Parthasaarathy, et al.
Veröffentlicht: (2025)
INFER : Learning Implicit Neural Frequency Response Fields for Confined Car Cabin
von: Takawale, Harshvardhan C., et al.
Veröffentlicht: (2025)
von: Takawale, Harshvardhan C., et al.
Veröffentlicht: (2025)
TF-CorrNet: Leveraging Spatial Correlation for Continuous Speech Separation
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2025)
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2025)
Two-Stage Adaptation for Non-Normative Speech Recognition: Revisiting Speaker-Independent Initialization for Personalization
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
Decomposing the Influence of Physical Acoustic Modeling on Neural Personal Sound Zone Rendering: An Ablation Study
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
SALM: Spatial Audio Language Model with Structured Embeddings for Understanding and Editing
von: Hu, Jinbo, et al.
Veröffentlicht: (2025)
von: Hu, Jinbo, et al.
Veröffentlicht: (2025)
Knowledge-Decoupled Functionally Invariant Path with Synthetic Personal Data for Personalized ASR
von: Gu, Yue, et al.
Veröffentlicht: (2025)
von: Gu, Yue, et al.
Veröffentlicht: (2025)
FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates
von: Li, Jiaqi, et al.
Veröffentlicht: (2025)
von: Li, Jiaqi, et al.
Veröffentlicht: (2025)
HRTF upsampling with a generative adversarial network using a gnomonic equiangular projection
von: Hogg, Aidan O. T., et al.
Veröffentlicht: (2023)
von: Hogg, Aidan O. T., et al.
Veröffentlicht: (2023)
LSZone: A Lightweight Spatial Information Modeling Architecture for Real-time In-car Multi-zone Speech Separation
von: Chen, Jun, et al.
Veröffentlicht: (2025)
von: Chen, Jun, et al.
Veröffentlicht: (2025)
MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations
von: Guo, Wenxiang, et al.
Veröffentlicht: (2025)
von: Guo, Wenxiang, et al.
Veröffentlicht: (2025)
Representing Sounds as Neural Amplitude Fields: A Benchmark of Coordinate-MLPs and A Fourier Kolmogorov-Arnold Framework
von: Li, Linfei, et al.
Veröffentlicht: (2026)
von: Li, Linfei, et al.
Veröffentlicht: (2026)
Deep Learning for Personalized Binaural Audio Reproduction
von: Lu, Xikun, et al.
Veröffentlicht: (2025)
von: Lu, Xikun, et al.
Veröffentlicht: (2025)
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
Creating Personalized Synthetic Voices from Articulation Impaired Speech Using Augmented Reconstruction Loss
von: Tian, Yusheng, et al.
Veröffentlicht: (2024)
von: Tian, Yusheng, et al.
Veröffentlicht: (2024)
Spatial-CLAP: Learning Spatially-Aware audio--text Embeddings for Multi-Source Conditions
von: Seki, Kentaro, et al.
Veröffentlicht: (2025)
von: Seki, Kentaro, et al.
Veröffentlicht: (2025)
Data Augmentation Using Neural Acoustic Fields With Retrieval-Augmented Pre-training
von: Ick, Christopher, et al.
Veröffentlicht: (2025)
von: Ick, Christopher, et al.
Veröffentlicht: (2025)
Precise and Simple Audio-to-Score Alignment
von: Peter, Silvan, et al.
Veröffentlicht: (2026)
von: Peter, Silvan, et al.
Veröffentlicht: (2026)
Score-Agnostic Structure Analysis in Large-Scale Performance Datasets
von: Hu, Patricia, et al.
Veröffentlicht: (2026)
von: Hu, Patricia, et al.
Veröffentlicht: (2026)
NSTR: Neural Spectral Transport Representation for Space-Varying Frequency Fields
von: Versace, Plein
Veröffentlicht: (2025)
von: Versace, Plein
Veröffentlicht: (2025)
AnalysisGNN: Unified Music Analysis with Graph Neural Networks
von: Karystinaios, Emmanouil, et al.
Veröffentlicht: (2025)
von: Karystinaios, Emmanouil, et al.
Veröffentlicht: (2025)
CLAIP-Emo: Parameter-Efficient Adaptation of Language-supervised models for In-the-Wild Audiovisual Emotion Recognition
von: Chen, Yin, et al.
Veröffentlicht: (2025)
von: Chen, Yin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025) -
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024) -
Towards HRTF Personalization using Denoising Diffusion Models
von: Sánchez, Juan Camilo Albarracín, et al.
Veröffentlicht: (2025) -
The Extended SONICOM HRTF Dataset and Spatial Audio Metrics Toolbox
von: Poole, Katarina C., et al.
Veröffentlicht: (2025) -
HRTFformer: A Spatially-Aware Transformer for Individual HRTF Upsampling in Immersive Audio Rendering
von: Hu, Xuyi, et al.
Veröffentlicht: (2025)