Gaussian Process Regression of Steering Vectors With Physics-Aware Deep Composite Kernels for Augmented Listening
Fuente:
arXiv
Salvato in:
| Autori principali: | Di Carlo, Diego, Koyama, Shoichi, Arie, Nugraha Aditya, Mathieu, Fontaine, Yoshiaki, Bando, Kazuyoshi, Yoshii |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Neural Steerer: Novel Steering Vector Synthesis with a Causal Neural Field over Frequency and Source Positions
di: Di Carlo, Diego, et al.
Pubblicazione: (2023)
di: Di Carlo, Diego, et al.
Pubblicazione: (2023)
SHAMaNS: Sound Localization with Hybrid Alpha-Stable Spatial Measure and Neural Steerer
di: Di Carlo, Diego, et al.
Pubblicazione: (2025)
di: Di Carlo, Diego, et al.
Pubblicazione: (2025)
SIRUP: A diffusion-based virtual upmixer of steering vectors for highly-directive spatialization with first-order ambisonics
di: Picard, Emilio, et al.
Pubblicazione: (2026)
di: Picard, Emilio, et al.
Pubblicazione: (2026)
Run-Time Adaptation of Neural Beamforming for Robust Speech Dereverberation and Denoising
di: Fujita, Yoto, et al.
Pubblicazione: (2024)
di: Fujita, Yoto, et al.
Pubblicazione: (2024)
Learning Magnitude Distribution of Sound Fields via Conditioned Autoencoder
di: Koyama, Shoichi, et al.
Pubblicazione: (2025)
di: Koyama, Shoichi, et al.
Pubblicazione: (2025)
DOA-Aware Audio-Visual Self-Supervised Learning for Sound Event Localization and Detection
di: Fujita, Yoto, et al.
Pubblicazione: (2024)
di: Fujita, Yoto, et al.
Pubblicazione: (2024)
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
di: Koga, Naoki, et al.
Pubblicazione: (2024)
di: Koga, Naoki, et al.
Pubblicazione: (2024)
Head-Related Transfer Function Individualization Using Anthropometric Features and Spatially Independent Latent Representation
di: Niu, Ryan, et al.
Pubblicazione: (2025)
di: Niu, Ryan, et al.
Pubblicazione: (2025)
Binaural rendering from microphone array signals of arbitrary geometry
di: Iijima, Naoto, et al.
Pubblicazione: (2021)
di: Iijima, Naoto, et al.
Pubblicazione: (2021)
Localizing Acoustic Energy in Sound Field Synthesis by Directionally Weighted Exterior Radiation Suppression
di: Tomita, Yoshihide, et al.
Pubblicazione: (2024)
di: Tomita, Yoshihide, et al.
Pubblicazione: (2024)
Infrastructure-less Localization from Indoor Environmental Sounds Based on Spectral Decomposition and Spatial Likelihood Model
di: Ogiso, Satoki, et al.
Pubblicazione: (2024)
di: Ogiso, Satoki, et al.
Pubblicazione: (2024)
Low-Rank Adaptation of Deep Prior Neural Networks For Room Impulse Response Reconstruction
di: Pezzoli, Mirco, et al.
Pubblicazione: (2025)
di: Pezzoli, Mirco, et al.
Pubblicazione: (2025)
Phase-Retrieval-Based Physics-Informed Neural Networks For Acoustic Magnitude Field Reconstruction
di: Schrader, Karl, et al.
Pubblicazione: (2026)
di: Schrader, Karl, et al.
Pubblicazione: (2026)
Reproducing the Acoustic Velocity Vectors in a Circular Listening Area
di: Wang, Jiarui, et al.
Pubblicazione: (2024)
di: Wang, Jiarui, et al.
Pubblicazione: (2024)
Physics-Informed Machine Learning For Sound Field Estimation
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
Past, Present, and Future of Spatial Audio and Room Acoustics
di: Koyama, Shoichi, et al.
Pubblicazione: (2025)
di: Koyama, Shoichi, et al.
Pubblicazione: (2025)
Contrastive Knowledge Distillation for Embedding Refinement in Personalized Speech Enhancement
di: Serre, Thomas, et al.
Pubblicazione: (2026)
di: Serre, Thomas, et al.
Pubblicazione: (2026)
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
di: Serre, Thomas, et al.
Pubblicazione: (2024)
di: Serre, Thomas, et al.
Pubblicazione: (2024)
Speech dereverberation constrained on room impulse response characteristics
di: Bahrman, Louis, et al.
Pubblicazione: (2024)
di: Bahrman, Louis, et al.
Pubblicazione: (2024)
Time-domain sound field estimation using kernel ridge regression
di: Brunnström, Jesper, et al.
Pubblicazione: (2025)
di: Brunnström, Jesper, et al.
Pubblicazione: (2025)
Streaming Piano Transcription Based on Consistent Onset and Offset Decoding with Sustain Pedal Detection
di: Wei, Weixing, et al.
Pubblicazione: (2025)
di: Wei, Weixing, et al.
Pubblicazione: (2025)
Listen, Think, and Understand
di: Gong, Yuan, et al.
Pubblicazione: (2023)
di: Gong, Yuan, et al.
Pubblicazione: (2023)
A Hybrid Model for Weakly-Supervised Speech Dereverberation
di: Bahrman, Louis, et al.
Pubblicazione: (2025)
di: Bahrman, Louis, et al.
Pubblicazione: (2025)
Modèle physique variationnel pour l'estimation de réponses impulsionnelles de salles
di: Lalay, Louis, et al.
Pubblicazione: (2025)
di: Lalay, Louis, et al.
Pubblicazione: (2025)
Online speaker diarization of meetings guided by speech separation
di: Gruttadauria, Elio, et al.
Pubblicazione: (2024)
di: Gruttadauria, Elio, et al.
Pubblicazione: (2024)
Déréverbération non-supervisée de la parole par modèle hybride
di: Bahrman, Louis, et al.
Pubblicazione: (2025)
di: Bahrman, Louis, et al.
Pubblicazione: (2025)
Listen First, Then Answer: Timestamp-Grounded Speech Reasoning
di: Jeong, Jihoon, et al.
Pubblicazione: (2026)
di: Jeong, Jihoon, et al.
Pubblicazione: (2026)
Evaluating Speech Enhancement Systems Through Listening Effort
di: Gelderblom, Femke B., et al.
Pubblicazione: (2024)
di: Gelderblom, Femke B., et al.
Pubblicazione: (2024)
RF-GML: Reference-Free Generative Machine Listener
di: Biswas, Arijit, et al.
Pubblicazione: (2024)
di: Biswas, Arijit, et al.
Pubblicazione: (2024)
Listen to Extract: Onset-Prompted Target Speaker Extraction
di: Shen, Pengjie, et al.
Pubblicazione: (2025)
di: Shen, Pengjie, et al.
Pubblicazione: (2025)
Listening to Multi-talker Conversations: Modular and End-to-end Perspectives
di: Raj, Desh
Pubblicazione: (2024)
di: Raj, Desh
Pubblicazione: (2024)
Requirements for Mass Adoption of Assistive Listening Technology by the General Public
di: Kaufmann, Thomas B., et al.
Pubblicazione: (2023)
di: Kaufmann, Thomas B., et al.
Pubblicazione: (2023)
DIFFA: Large Language Diffusion Models Can Listen and Understand
di: Zhou, Jiaming, et al.
Pubblicazione: (2025)
di: Zhou, Jiaming, et al.
Pubblicazione: (2025)
Joint Minimum Processing Beamforming and Near-end Listening Enhancement
di: Fuglsig, Andreas J., et al.
Pubblicazione: (2023)
di: Fuglsig, Andreas J., et al.
Pubblicazione: (2023)
Listening broadband physical model for microphones: a first step
di: Millot, Laurent, et al.
Pubblicazione: (2024)
di: Millot, Laurent, et al.
Pubblicazione: (2024)
Active Listener: Continuous Generation of Listener's Head Motion Response in Dyadic Interactions
di: Ghosh, Bishal, et al.
Pubblicazione: (2024)
di: Ghosh, Bishal, et al.
Pubblicazione: (2024)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
di: Chung, Soo-Whan, et al.
Pubblicazione: (2025)
di: Chung, Soo-Whan, et al.
Pubblicazione: (2025)
Reading to Listen at the Cocktail Party: Multi-Modal Speech Separation
di: Rahimi, Akam, et al.
Pubblicazione: (2025)
di: Rahimi, Akam, et al.
Pubblicazione: (2025)
Steer-by-prior Editing of Symbolic Music Loops
di: Jonason, Nicolas, et al.
Pubblicazione: (2024)
di: Jonason, Nicolas, et al.
Pubblicazione: (2024)
U-DREAM: Unsupervised Dereverberation guided by a Reverberation Model
di: Bahrman, Louis, et al.
Pubblicazione: (2025)
di: Bahrman, Louis, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Neural Steerer: Novel Steering Vector Synthesis with a Causal Neural Field over Frequency and Source Positions
di: Di Carlo, Diego, et al.
Pubblicazione: (2023) -
SHAMaNS: Sound Localization with Hybrid Alpha-Stable Spatial Measure and Neural Steerer
di: Di Carlo, Diego, et al.
Pubblicazione: (2025) -
SIRUP: A diffusion-based virtual upmixer of steering vectors for highly-directive spatialization with first-order ambisonics
di: Picard, Emilio, et al.
Pubblicazione: (2026) -
Run-Time Adaptation of Neural Beamforming for Robust Speech Dereverberation and Denoising
di: Fujita, Yoto, et al.
Pubblicazione: (2024) -
Learning Magnitude Distribution of Sound Fields via Conditioned Autoencoder
di: Koyama, Shoichi, et al.
Pubblicazione: (2025)