MFF-EINV2: Multi-scale Feature Fusion across Spectral-Spatial-Temporal Domains for Sound Event Localization and Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mu, Da, Zhang, Zhicheng, Yue, Haobo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation
von: Mu, Da, et al.
Veröffentlicht: (2024)
von: Mu, Da, et al.
Veröffentlicht: (2024)
MTDA-HSED: Mutual-Assistance Tuning and Dual-Branch Aggregating for Heterogeneous Sound Event Detection
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
Full-frequency dynamic convolution: a physical frequency-dependent convolution for sound event detection
von: Yue, Haobo, et al.
Veröffentlicht: (2024)
von: Yue, Haobo, et al.
Veröffentlicht: (2024)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
Detect Any Sound: Open-Vocabulary Sound Event Detection with Multi-Modal Queries
von: Cai, Pengfei, et al.
Veröffentlicht: (2025)
von: Cai, Pengfei, et al.
Veröffentlicht: (2025)
MUDAS: Mote-scale Unsupervised Domain Adaptation in Multi-label Sound Classification
von: Yun, Jihoon, et al.
Veröffentlicht: (2025)
von: Yun, Jihoon, et al.
Veröffentlicht: (2025)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
Zero- and Few-shot Sound Event Localization and Detection
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
Enhanced Sound Event Localization and Detection in Real 360-degree audio-visual soundscapes
von: Roman, Adrian S., et al.
Veröffentlicht: (2024)
von: Roman, Adrian S., et al.
Veröffentlicht: (2024)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
TLDiffGAN: A Latent Diffusion-GAN Framework with Temporal Information Fusion for Anomalous Sound Detection
von: Ma, Chengyuan, et al.
Veröffentlicht: (2026)
von: Ma, Chengyuan, et al.
Veröffentlicht: (2026)
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Disentangling Hierarchical Features for Anomalous Sound Detection Under Domain Shift
von: Guan, Jian, et al.
Veröffentlicht: (2025)
von: Guan, Jian, et al.
Veröffentlicht: (2025)
Leveraging Language Model Capabilities for Sound Event Detection
von: Wang, Hualei, et al.
Veröffentlicht: (2023)
von: Wang, Hualei, et al.
Veröffentlicht: (2023)
AudioSpa: Spatializing Sound Events with Text
von: Feng, Linfeng, et al.
Veröffentlicht: (2025)
von: Feng, Linfeng, et al.
Veröffentlicht: (2025)
w2v-SELD: A Sound Event Localization and Detection Framework for Self-Supervised Spatial Audio Pre-Training
von: Santos, Orlem Lima dos, et al.
Veröffentlicht: (2023)
von: Santos, Orlem Lima dos, et al.
Veröffentlicht: (2023)
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
von: Bibbó, Gabriel, et al.
Veröffentlicht: (2024)
von: Bibbó, Gabriel, et al.
Veröffentlicht: (2024)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
von: Yu, Hogeon
Veröffentlicht: (2025)
von: Yu, Hogeon
Veröffentlicht: (2025)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
FlexSED: Towards Open-Vocabulary Sound Event Detection
von: Hai, Jiarui, et al.
Veröffentlicht: (2025)
von: Hai, Jiarui, et al.
Veröffentlicht: (2025)
Contrastive Loss Based Frame-wise Feature disentanglement for Polyphonic Sound Event Detection
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
Infrastructure-less Localization from Indoor Environmental Sounds Based on Spectral Decomposition and Spatial Likelihood Model
von: Ogiso, Satoki, et al.
Veröffentlicht: (2024)
von: Ogiso, Satoki, et al.
Veröffentlicht: (2024)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
Deep Generic Representations for Domain-Generalized Anomalous Sound Detection
von: Saengthong, Phurich, et al.
Veröffentlicht: (2024)
von: Saengthong, Phurich, et al.
Veröffentlicht: (2024)
Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
Frequency Dynamic Convolutions for Sound Event Detection
von: Nam, Hyeonuk
Veröffentlicht: (2025)
von: Nam, Hyeonuk
Veröffentlicht: (2025)
Improving Audio Spectrogram Transformers for Sound Event Detection Through Multi-Stage Training
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
Prototype based Masked Audio Model for Self-Supervised Learning of Sound Event Detection
von: Cai, Pengfei, et al.
Veröffentlicht: (2024)
von: Cai, Pengfei, et al.
Veröffentlicht: (2024)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
Towards Understanding of Frequency Dependence on Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Sound Event Detection with Boundary-Aware Optimization and Inference
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
Sound Event Detection and Localization with Distance Estimation
von: Krause, Daniel Aleksander, et al.
Veröffentlicht: (2024)
von: Krause, Daniel Aleksander, et al.
Veröffentlicht: (2024)
Abstract Sound Fusion with Unconditional Inversion Models
von: Liu, Jing, et al.
Veröffentlicht: (2025)
von: Liu, Jing, et al.
Veröffentlicht: (2025)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Effective Pre-Training of Audio Transformers for Sound Event Detection
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation
von: Mu, Da, et al.
Veröffentlicht: (2024) -
MTDA-HSED: Mutual-Assistance Tuning and Dual-Branch Aggregating for Heterogeneous Sound Event Detection
von: Wang, Zehao, et al.
Veröffentlicht: (2024) -
Full-frequency dynamic convolution: a physical frequency-dependent convolution for sound event detection
von: Yue, Haobo, et al.
Veröffentlicht: (2024) -
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
von: Zhang, Xueping, et al.
Veröffentlicht: (2025) -
Detect Any Sound: Open-Vocabulary Sound Event Detection with Multi-Modal Queries
von: Cai, Pengfei, et al.
Veröffentlicht: (2025)