Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution
Fuente:
arXiv
Salvato in:
| Autori principali: | Nam, Hyeonuk, Park, Yong-Hwa |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
Frequency Dynamic Convolutions for Sound Event Detection
di: Nam, Hyeonuk
Pubblicazione: (2025)
di: Nam, Hyeonuk
Pubblicazione: (2025)
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
Diversifying and Expanding Frequency-Adaptive Convolution Kernels for Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2024)
di: Nam, Hyeonuk, et al.
Pubblicazione: (2024)
Towards Understanding of Frequency Dependence on Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
Self Training and Ensembling Frequency Dependent Networks with Coarse Prediction Pooling and Sound Event Bounding Boxes
di: Nam, Hyeonuk, et al.
Pubblicazione: (2024)
di: Nam, Hyeonuk, et al.
Pubblicazione: (2024)
Auditory Intelligence: Understanding the World Through Sound
di: Nam, Hyeonuk
Pubblicazione: (2025)
di: Nam, Hyeonuk
Pubblicazione: (2025)
Frequency-Domain Sound Field from the Perspective of Band-Limited Functions
di: Iwami, Takahiro, et al.
Pubblicazione: (2024)
di: Iwami, Takahiro, et al.
Pubblicazione: (2024)
Improving Audio Spectrogram Transformers for Sound Event Detection Through Multi-Stage Training
di: Schmid, Florian, et al.
Pubblicazione: (2024)
di: Schmid, Florian, et al.
Pubblicazione: (2024)
Zero- and Few-shot Sound Event Localization and Detection
di: Shimada, Kazuki, et al.
Pubblicazione: (2023)
di: Shimada, Kazuki, et al.
Pubblicazione: (2023)
Sound Event Detection with Boundary-Aware Optimization and Inference
di: Schmid, Florian, et al.
Pubblicazione: (2026)
di: Schmid, Florian, et al.
Pubblicazione: (2026)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
Effective Pre-Training of Audio Transformers for Sound Event Detection
di: Schmid, Florian, et al.
Pubblicazione: (2024)
di: Schmid, Florian, et al.
Pubblicazione: (2024)
Hierarchical Pooling Structure for Weakly Labeled Sound Event Detection
di: He, Ke-Xin, et al.
Pubblicazione: (2019)
di: He, Ke-Xin, et al.
Pubblicazione: (2019)
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
di: Xie, Zeyu, et al.
Pubblicazione: (2023)
di: Xie, Zeyu, et al.
Pubblicazione: (2023)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
di: Xiao, Yang, et al.
Pubblicazione: (2024)
di: Xiao, Yang, et al.
Pubblicazione: (2024)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
di: Yin, Han, et al.
Pubblicazione: (2024)
di: Yin, Han, et al.
Pubblicazione: (2024)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
di: Gao, Wenmiao, et al.
Pubblicazione: (2025)
di: Gao, Wenmiao, et al.
Pubblicazione: (2025)
Sound Event Bounding Boxes
di: Ebbers, Janek, et al.
Pubblicazione: (2024)
di: Ebbers, Janek, et al.
Pubblicazione: (2024)
BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
di: Xin, Detai, et al.
Pubblicazione: (2024)
di: Xin, Detai, et al.
Pubblicazione: (2024)
Fine-Grained Engine Fault Sound Event Detection Using Multimodal Signals
di: Fedorishin, Dennis, et al.
Pubblicazione: (2024)
di: Fedorishin, Dennis, et al.
Pubblicazione: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
di: Yin, Han, et al.
Pubblicazione: (2024)
di: Yin, Han, et al.
Pubblicazione: (2024)
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
di: Roman, Adrian S., et al.
Pubblicazione: (2025)
di: Roman, Adrian S., et al.
Pubblicazione: (2025)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
di: Zhang, Xueping, et al.
Pubblicazione: (2025)
di: Zhang, Xueping, et al.
Pubblicazione: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
di: Yu, Hogeon
Pubblicazione: (2025)
di: Yu, Hogeon
Pubblicazione: (2025)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
di: Hu, Jinbo, et al.
Pubblicazione: (2023)
di: Hu, Jinbo, et al.
Pubblicazione: (2023)
Detect Any Sound: Open-Vocabulary Sound Event Detection with Multi-Modal Queries
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
di: Imoto, Keisuke
Pubblicazione: (2025)
di: Imoto, Keisuke
Pubblicazione: (2025)
Contrastive Loss Based Frame-wise Feature disentanglement for Polyphonic Sound Event Detection
di: Guan, Yadong, et al.
Pubblicazione: (2024)
di: Guan, Yadong, et al.
Pubblicazione: (2024)
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
di: Koga, Naoki, et al.
Pubblicazione: (2024)
di: Koga, Naoki, et al.
Pubblicazione: (2024)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
di: Shen, Yu-Han, et al.
Pubblicazione: (2018)
di: Shen, Yu-Han, et al.
Pubblicazione: (2018)
Automatic Sound Event Detection and Classification of Great Ape Calls Using Neural Networks
di: Jiang, Zifan, et al.
Pubblicazione: (2023)
di: Jiang, Zifan, et al.
Pubblicazione: (2023)
WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
di: Xiao, Yang, et al.
Pubblicazione: (2024)
di: Xiao, Yang, et al.
Pubblicazione: (2024)
AudioSpa: Spatializing Sound Events with Text
di: Feng, Linfeng, et al.
Pubblicazione: (2025)
di: Feng, Linfeng, et al.
Pubblicazione: (2025)
"It is okay to be uncommon": Quantizing Sound Event Detection Networks on Hardware Accelerators with Uncommon Sub-Byte Support
di: Wu, Yushu, et al.
Pubblicazione: (2024)
di: Wu, Yushu, et al.
Pubblicazione: (2024)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
di: Yeow, Jun-Wei, et al.
Pubblicazione: (2025)
di: Yeow, Jun-Wei, et al.
Pubblicazione: (2025)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
Documenti analoghi
-
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025) -
Frequency Dynamic Convolutions for Sound Event Detection
di: Nam, Hyeonuk
Pubblicazione: (2025) -
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025) -
Diversifying and Expanding Frequency-Adaptive Convolution Kernels for Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2024) -
Towards Understanding of Frequency Dependence on Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)