Pseudo Strong Labels from Frame-Level Predictions for Weakly Supervised Sound Event Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yuliang, Defeng, Huang, Togneri, Roberto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Impact of Noisy Labels on Sound Event Detection: Deletion Errors Are More Detrimental Than Insertion Errors
von: Zhang, Yuliang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuliang, et al.
Veröffentlicht: (2024)
Filling MIDI Velocity using U-Net Image Colorizer
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
Hierarchical Pooling Structure for Weakly Labeled Sound Event Detection
von: He, Ke-Xin, et al.
Veröffentlicht: (2019)
von: He, Ke-Xin, et al.
Veröffentlicht: (2019)
From Weak to Strong Sound Event Labels using Adaptive Change-Point Detection and Active Learning
von: Martinsson, John, et al.
Veröffentlicht: (2024)
von: Martinsson, John, et al.
Veröffentlicht: (2024)
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
von: Imoto, Keisuke
Veröffentlicht: (2025)
von: Imoto, Keisuke
Veröffentlicht: (2025)
Contrastive Loss Based Frame-wise Feature disentanglement for Polyphonic Sound Event Detection
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
von: Zhao, PengYuan, et al.
Veröffentlicht: (2024)
von: Zhao, PengYuan, et al.
Veröffentlicht: (2024)
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
von: Koga, Naoki, et al.
Veröffentlicht: (2024)
von: Koga, Naoki, et al.
Veröffentlicht: (2024)
w2v-SELD: A Sound Event Localization and Detection Framework for Self-Supervised Spatial Audio Pre-Training
von: Santos, Orlem Lima dos, et al.
Veröffentlicht: (2023)
von: Santos, Orlem Lima dos, et al.
Veröffentlicht: (2023)
Class-Incremental Learning for Sound Event Localization and Detection
von: Pandey, Ruchi, et al.
Veröffentlicht: (2024)
von: Pandey, Ruchi, et al.
Veröffentlicht: (2024)
Frequency Dynamic Convolutions for Sound Event Detection
von: Nam, Hyeonuk
Veröffentlicht: (2025)
von: Nam, Hyeonuk
Veröffentlicht: (2025)
DG-SED: Domain Generalization for Sound Event Detection with Heterogeneous Training Data
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection
von: Liang, Jinhua, et al.
Veröffentlicht: (2024)
von: Liang, Jinhua, et al.
Veröffentlicht: (2024)
Improving Anomalous Sound Detection through Pseudo-anomalous Set Selection and Pseudo-label Utilization under Unlabeled Conditions
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
SwG-former: A Sliding-Window Graph Convolutional Network for Simultaneous Spatial-Temporal Information Extraction in Sound Event Localization and Detection
von: Huang, Weiming, et al.
Veröffentlicht: (2023)
von: Huang, Weiming, et al.
Veröffentlicht: (2023)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Towards Understanding of Frequency Dependence on Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Zero- and Few-shot Sound Event Localization and Detection
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
Sound Event Detection with Boundary-Aware Optimization and Inference
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Exploring Self-Supervised Audio Models for Generalized Anomalous Sound Detection
von: Han, Bing, et al.
Veröffentlicht: (2025)
von: Han, Bing, et al.
Veröffentlicht: (2025)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
CST-former: Multidimensional Attention-based Transformer for Sound Event Localization and Detection in Real Scenes
von: Shul, Yusun, et al.
Veröffentlicht: (2025)
von: Shul, Yusun, et al.
Veröffentlicht: (2025)
Multi-Iteration Multi-Stage Fine-Tuning of Transformers for Sound Event Detection with Heterogeneous Datasets
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Effective Pre-Training of Audio Transformers for Sound Event Detection
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
Text-Queried Target Sound Event Localization
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2024)
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2024)
MVANet: Multi-Stage Video Attention Network for Sound Event Localization and Detection with Source Distance Estimation
von: Hong, Hengyi, et al.
Veröffentlicht: (2024)
von: Hong, Hengyi, et al.
Veröffentlicht: (2024)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
Weakly Supervised Phonological Features for Pathological Speech Analysis
von: Thienpondt, Jenthe, et al.
Veröffentlicht: (2025)
von: Thienpondt, Jenthe, et al.
Veröffentlicht: (2025)
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
Diversifying and Expanding Frequency-Adaptive Convolution Kernels for Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2024)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
"It is okay to be uncommon": Quantizing Sound Event Detection Networks on Hardware Accelerators with Uncommon Sub-Byte Support
von: Wu, Yushu, et al.
Veröffentlicht: (2024)
von: Wu, Yushu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Impact of Noisy Labels on Sound Event Detection: Deletion Errors Are More Detrimental Than Insertion Errors
von: Zhang, Yuliang, et al.
Veröffentlicht: (2024) -
Filling MIDI Velocity using U-Net Image Colorizer
von: He, Zhanhong, et al.
Veröffentlicht: (2025) -
Hierarchical Pooling Structure for Weakly Labeled Sound Event Detection
von: He, Ke-Xin, et al.
Veröffentlicht: (2019) -
From Weak to Strong Sound Event Labels using Adaptive Change-Point Detection and Active Learning
von: Martinsson, John, et al.
Veröffentlicht: (2024) -
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
von: He, Zhanhong, et al.
Veröffentlicht: (2025)