Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Wenmiao, Yin, Han |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
TC-BiMamba: Trans-Chunk bidirectionally within BiMamba for unified streaming and non-streaming ASR
von: She, Qingshun, et al.
Veröffentlicht: (2026)
von: She, Qingshun, et al.
Veröffentlicht: (2026)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
Hierarchical Pooling Structure for Weakly Labeled Sound Event Detection
von: He, Ke-Xin, et al.
Veröffentlicht: (2019)
von: He, Ke-Xin, et al.
Veröffentlicht: (2019)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
BiCrossMamba-ST: Speech Deepfake Detection with Bidirectional Mamba Spectro-Temporal Cross-Attention
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
von: Yu, Hogeon
Veröffentlicht: (2025)
von: Yu, Hogeon
Veröffentlicht: (2025)
SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation
von: Mu, Da, et al.
Veröffentlicht: (2024)
von: Mu, Da, et al.
Veröffentlicht: (2024)
Frequency Dynamic Convolutions for Sound Event Detection
von: Nam, Hyeonuk
Veröffentlicht: (2025)
von: Nam, Hyeonuk
Veröffentlicht: (2025)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
Towards Understanding of Frequency Dependence on Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Zero- and Few-shot Sound Event Localization and Detection
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
Sound Event Detection with Boundary-Aware Optimization and Inference
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
Contrastive Loss Based Frame-wise Feature disentanglement for Polyphonic Sound Event Detection
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
Effective Pre-Training of Audio Transformers for Sound Event Detection
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Trainingless Adaptation of Pretrained Models for Environmental Sound Classification
von: Tonami, Noriyuki, et al.
Veröffentlicht: (2024)
von: Tonami, Noriyuki, et al.
Veröffentlicht: (2024)
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Diversifying and Expanding Frequency-Adaptive Convolution Kernels for Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
VCNAC: A Variable-Channel Neural Audio Codec for Mono, Stereo, and Surround Sound
von: Grötschla, Florian, et al.
Veröffentlicht: (2026)
von: Grötschla, Florian, et al.
Veröffentlicht: (2026)
Sound Event Bounding Boxes
von: Ebbers, Janek, et al.
Veröffentlicht: (2024)
von: Ebbers, Janek, et al.
Veröffentlicht: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae, et al.
Veröffentlicht: (2025)
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
Fine-Grained Engine Fault Sound Event Detection Using Multimodal Signals
von: Fedorishin, Dennis, et al.
Veröffentlicht: (2024)
von: Fedorishin, Dennis, et al.
Veröffentlicht: (2024)
Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
TF-Mamba: A Time-Frequency Network for Sound Source Localization
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
von: Imoto, Keisuke
Veröffentlicht: (2025)
von: Imoto, Keisuke
Veröffentlicht: (2025)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
Automatic Sound Event Detection and Classification of Great Ape Calls Using Neural Networks
von: Jiang, Zifan, et al.
Veröffentlicht: (2023)
von: Jiang, Zifan, et al.
Veröffentlicht: (2023)
Improving Audio Spectrogram Transformers for Sound Event Detection Through Multi-Stage Training
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
von: Koga, Naoki, et al.
Veröffentlicht: (2024)
von: Koga, Naoki, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025) -
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025) -
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025) -
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
von: Yin, Han, et al.
Veröffentlicht: (2024) -
TC-BiMamba: Trans-Chunk bidirectionally within BiMamba for unified streaming and non-streaming ASR
von: She, Qingshun, et al.
Veröffentlicht: (2026)