Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Berghi, Davide, Jackson, Philip J. B. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Leveraging Reverberation and Visual Depth Cues for Sound Event Localization and Detection with Distance Estimation
por: Berghi, Davide, et al.
Publicado: (2024)
por: Berghi, Davide, et al.
Publicado: (2024)
ToS: A Team of Specialists ensemble framework for Stereo Sound Event Localization and Detection with distance estimation in Video
por: Berghi, Davide, et al.
Publicado: (2026)
por: Berghi, Davide, et al.
Publicado: (2026)
Spatial and Semantic Embedding Integration for Stereo Sound Event Localization and Detection in Regular Videos
por: Berghi, Davide, et al.
Publicado: (2025)
por: Berghi, Davide, et al.
Publicado: (2025)
Integrating Spatial and Semantic Embeddings for Stereo Sound Event Localization in Videos
por: Berghi, Davide, et al.
Publicado: (2025)
por: Berghi, Davide, et al.
Publicado: (2025)
Audio-Visual Talker Localization in Video for Spatial Sound Reproduction
por: Berghi, Davide, et al.
Publicado: (2024)
por: Berghi, Davide, et al.
Publicado: (2024)
Dependence on Early and Late Reverberation of Single-Channel Speaker Distance Estimation
por: Neri, Michael, et al.
Publicado: (2026)
por: Neri, Michael, et al.
Publicado: (2026)
A Method for Capturing and Reproducing Directional Reverberation in Six Degrees of Freedom
por: Alary, Benoit, et al.
Publicado: (2021)
por: Alary, Benoit, et al.
Publicado: (2021)
Directional Selective Fixed-Filter Active Noise Control Based on a Convolutional Neural Network in Reverberant Environments
por: Wang, Boxiang, et al.
Publicado: (2026)
por: Wang, Boxiang, et al.
Publicado: (2026)
Sound Event Detection and Localization with Distance Estimation
por: Krause, Daniel Aleksander, et al.
Publicado: (2024)
por: Krause, Daniel Aleksander, et al.
Publicado: (2024)
Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier
por: Dumpis, Martynas, et al.
Publicado: (2026)
por: Dumpis, Martynas, et al.
Publicado: (2026)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
por: Torabi, Yasaman, et al.
Publicado: (2025)
por: Torabi, Yasaman, et al.
Publicado: (2025)
U-DREAM: Unsupervised Dereverberation guided by a Reverberation Model
por: Bahrman, Louis, et al.
Publicado: (2025)
por: Bahrman, Louis, et al.
Publicado: (2025)
String Sound Synthesizer on GPU-accelerated Finite Difference Scheme
por: Lee, Jin Woo, et al.
Publicado: (2023)
por: Lee, Jin Woo, et al.
Publicado: (2023)
A Machine Hearing System for Robust Cough Detection Based on a High-Level Representation of Band-Specific Audio Features
por: Monge-Alvarez, Jesús, et al.
Publicado: (2024)
por: Monge-Alvarez, Jesús, et al.
Publicado: (2024)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
por: Hou, Yuanbo, et al.
Publicado: (2024)
por: Hou, Yuanbo, et al.
Publicado: (2024)
Online Similarity-and-Independence-Aware Beamformer for Low-latency Target Sound Extraction
por: Hiroe, Atsuo
Publicado: (2023)
por: Hiroe, Atsuo
Publicado: (2023)
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
por: Fejgin, Daniel, et al.
Publicado: (2025)
por: Fejgin, Daniel, et al.
Publicado: (2025)
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
por: Yao, Shengshi, et al.
Publicado: (2025)
por: Yao, Shengshi, et al.
Publicado: (2025)
Decomposing the Influence of Physical Acoustic Modeling on Neural Personal Sound Zone Rendering: An Ablation Study
por: Jiang, Hao, et al.
Publicado: (2026)
por: Jiang, Hao, et al.
Publicado: (2026)
RIFT: Entropy-Optimised Fractional Wavelet Constellations for Ideal Time-Frequency Estimation
por: Cozens, James M., et al.
Publicado: (2025)
por: Cozens, James M., et al.
Publicado: (2025)
RIR-Mega-Speech: A Reverberant Speech Corpus with Comprehensive Acoustic Metadata and Reproducible Evaluation
por: Goswami, Mandip
Publicado: (2026)
por: Goswami, Mandip
Publicado: (2026)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
por: Joysingh, S. Johanan, et al.
Publicado: (2024)
por: Joysingh, S. Johanan, et al.
Publicado: (2024)
Acoustical Features as Knee Health Biomarkers: A Critical Analysis
por: Kechris, Christodoulos, et al.
Publicado: (2024)
por: Kechris, Christodoulos, et al.
Publicado: (2024)
Speech-Declipping Transformer with Complex Spectrogram and Learnerble Temporal Features
por: Kwon, Younghoo, et al.
Publicado: (2024)
por: Kwon, Younghoo, et al.
Publicado: (2024)
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
Wavelet-Based Time-Frequency Fingerprinting for Feature Extraction of Traditional Irish Music
por: Shore, Noah
Publicado: (2025)
por: Shore, Noah
Publicado: (2025)
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
LocaGen: Sub-Sample Time-Delay Learning for Beam Localization
por: Kunwar, Ishaan, et al.
Publicado: (2025)
por: Kunwar, Ishaan, et al.
Publicado: (2025)
Region-Specific Audio Tagging for Spatial Sound
por: Zhao, Jinzheng, et al.
Publicado: (2025)
por: Zhao, Jinzheng, et al.
Publicado: (2025)
An Experimental Study on Joint Modeling for Sound Event Localization and Detection with Source Distance Estimation
por: Dong, Yuxuan, et al.
Publicado: (2025)
por: Dong, Yuxuan, et al.
Publicado: (2025)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
por: Brümann, Klaus, et al.
Publicado: (2024)
por: Brümann, Klaus, et al.
Publicado: (2024)
Comparison of Frequency-Fusion Mechanisms for Binaural Direction-of-Arrival Estimation for Multiple Speakers
por: Fejgin, Daniel, et al.
Publicado: (2024)
por: Fejgin, Daniel, et al.
Publicado: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
por: Lee, Gyeong-Tae, et al.
Publicado: (2025)
por: Lee, Gyeong-Tae, et al.
Publicado: (2025)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
por: Abu, Avi, et al.
Publicado: (2024)
por: Abu, Avi, et al.
Publicado: (2024)
Zero- and Few-shot Sound Event Localization and Detection
por: Shimada, Kazuki, et al.
Publicado: (2023)
por: Shimada, Kazuki, et al.
Publicado: (2023)
Exploiting an External Microphone for Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
por: Fejgin, Daniel, et al.
Publicado: (2023)
por: Fejgin, Daniel, et al.
Publicado: (2023)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
por: Yu, Chin-Yun, et al.
Publicado: (2024)
por: Yu, Chin-Yun, et al.
Publicado: (2024)
Towards Improving Speaker Distance Estimation through Generative Impulse Response Augmentation
por: Ratnarajah, Anton, et al.
Publicado: (2026)
por: Ratnarajah, Anton, et al.
Publicado: (2026)
Singing Voice Graph Modeling for SingFake Detection
por: Chen, Xuanjun, et al.
Publicado: (2024)
por: Chen, Xuanjun, et al.
Publicado: (2024)
Ejemplares similares
-
Leveraging Reverberation and Visual Depth Cues for Sound Event Localization and Detection with Distance Estimation
por: Berghi, Davide, et al.
Publicado: (2024) -
ToS: A Team of Specialists ensemble framework for Stereo Sound Event Localization and Detection with distance estimation in Video
por: Berghi, Davide, et al.
Publicado: (2026) -
Spatial and Semantic Embedding Integration for Stereo Sound Event Localization and Detection in Regular Videos
por: Berghi, Davide, et al.
Publicado: (2025) -
Integrating Spatial and Semantic Embeddings for Stereo Sound Event Localization in Videos
por: Berghi, Davide, et al.
Publicado: (2025) -
Audio-Visual Talker Localization in Video for Spatial Sound Reproduction
por: Berghi, Davide, et al.
Publicado: (2024)