Sound event detection based on auxiliary decoder and maximum probability aggregation for DCASE Challenge 2024 Task 4
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Son, Sang Won, Park, Jongyeon, Kim, Hong Kook, Vesal, Sulaiman, Lim, Jeong Eun |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Sound Scene Synthesis at the DCASE 2024 Challenge
par: Lagrange, Mathieu, et autres
Publié: (2025)
par: Lagrange, Mathieu, et autres
Publié: (2025)
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
par: Cornell, Samuele, et autres
Publié: (2024)
par: Cornell, Samuele, et autres
Publié: (2024)
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
par: Yasuda, Masahiro, et autres
Publié: (2025)
par: Yasuda, Masahiro, et autres
Publié: (2025)
Performance improvement of spatial semantic segmentation with enriched audio features and agent-based error correction for DCASE 2025 Challenge Task 4
par: Park, Jongyeon, et autres
Publié: (2025)
par: Park, Jongyeon, et autres
Publié: (2025)
Performance Improvement of Language-Queried Audio Source Separation Based on Caption Augmentation From Large Language Models for DCASE Challenge 2024 Task 9
par: Lee, Do Hyun, et autres
Publié: (2024)
par: Lee, Do Hyun, et autres
Publié: (2024)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
par: Nishida, Tomoya, et autres
Publié: (2026)
par: Nishida, Tomoya, et autres
Publié: (2026)
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
par: Nishida, Tomoya, et autres
Publié: (2025)
par: Nishida, Tomoya, et autres
Publié: (2025)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
par: Xiao, Yang, et autres
Publié: (2024)
par: Xiao, Yang, et autres
Publié: (2024)
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
par: Schmid, Florian, et autres
Publié: (2024)
par: Schmid, Florian, et autres
Publié: (2024)
Description and Discussion on DCASE 2024 Challenge Task 2: First-Shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
par: Nishida, Tomoya, et autres
Publié: (2024)
par: Nishida, Tomoya, et autres
Publié: (2024)
Handling Domain Shifts for Anomalous Sound Detection: A Review of DCASE-Related Work
par: Wilkinghoff, Kevin, et autres
Publié: (2025)
par: Wilkinghoff, Kevin, et autres
Publié: (2025)
Description and analysis of novelties introduced in DCASE Task 4 2022 on the baseline system
par: Ronchini, Francesca, et autres
Publié: (2022)
par: Ronchini, Francesca, et autres
Publié: (2022)
Low-Complexity Acoustic Scene Classification with Device Information in the DCASE 2025 Challenge
par: Schmid, Florian, et autres
Publié: (2025)
par: Schmid, Florian, et autres
Publié: (2025)
BioDCASE 2026 Challenge Baseline for Cross-Domain Mosquito Species Classification
par: Hou, Yuanbo, et autres
Publié: (2026)
par: Hou, Yuanbo, et autres
Publié: (2026)
AISTAT lab system for DCASE2025 Task6: Language-based audio retrieval
par: Kim, Hyun Jun, et autres
Publié: (2025)
par: Kim, Hyun Jun, et autres
Publié: (2025)
A decade of DCASE: Achievements, practices, evaluations and future challenges
par: Mesaros, Annamaria, et autres
Publié: (2024)
par: Mesaros, Annamaria, et autres
Publié: (2024)
Sound event localization and detection based on crnn using rectangular filters and channel rotation data augmentation
par: Ronchini, Francesca, et autres
Publié: (2020)
par: Ronchini, Francesca, et autres
Publié: (2020)
Solution for Temporal Sound Localisation Task of ECCV Second Perception Test Challenge 2024
par: Gu, Haowei, et autres
Publié: (2024)
par: Gu, Haowei, et autres
Publié: (2024)
Patient Domain Supervised Contrastive Learning for Lung Sound Classification Using Mobile Phone
par: Jeong, Seung Gyu, et autres
Publié: (2025)
par: Jeong, Seung Gyu, et autres
Publié: (2025)
Patient-Aware Feature Alignment for Robust Lung Sound Classification:Cohesion-Separation and Global Alignment Losses
par: Jeong, Seung Gyu, et autres
Publié: (2025)
par: Jeong, Seung Gyu, et autres
Publié: (2025)
Frequency-aware convolution for sound event detection
par: Song, Tao, et autres
Publié: (2024)
par: Song, Tao, et autres
Publié: (2024)
DiffSound: Differentiable Modal Sound Rendering and Inverse Rendering for Diverse Inference Tasks
par: Jin, Xutong, et autres
Publié: (2024)
par: Jin, Xutong, et autres
Publié: (2024)
An Empirical Analysis of Task-Induced Encoder Bias in Fréchet Audio Distance
par: Jeong, Wonwoo
Publié: (2026)
par: Jeong, Wonwoo
Publié: (2026)
Onset and offset weighted loss function for sound event detection
par: Song, Tao
Publié: (2024)
par: Song, Tao
Publié: (2024)
Fine-tune the pretrained ATST model for sound event detection
par: Shao, Nian, et autres
Publié: (2023)
par: Shao, Nian, et autres
Publié: (2023)
XWSB: A Blend System Utilizing XLS-R and WavLM with SLS Classifier detection system for SVDD 2024 Challenge
par: Zhang, Qishan, et autres
Publié: (2024)
par: Zhang, Qishan, et autres
Publié: (2024)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Cinematic Demixing Track
par: Uhlich, Stefan, et autres
Publié: (2023)
par: Uhlich, Stefan, et autres
Publié: (2023)
Technical Report of Nomi Team in the Environmental Sound Deepfake Detection Challenge 2026
par: Mawalim, Candy Olivia, et autres
Publié: (2025)
par: Mawalim, Candy Olivia, et autres
Publié: (2025)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Music Demixing Track
par: Fabbro, Giorgio, et autres
Publié: (2023)
par: Fabbro, Giorgio, et autres
Publié: (2023)
Full-frequency dynamic convolution: a physical frequency-dependent convolution for sound event detection
par: Yue, Haobo, et autres
Publié: (2024)
par: Yue, Haobo, et autres
Publié: (2024)
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
par: Gao, Wenmiao, et autres
Publié: (2025)
par: Gao, Wenmiao, et autres
Publié: (2025)
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
par: Ronchini, Francesca, et autres
Publié: (2023)
par: Ronchini, Francesca, et autres
Publié: (2023)
Resnet-conformer network with shared weights and attention mechanism for sound event localization, detection, and distance estimation
par: Vo, Quoc Thinh, et autres
Publié: (2025)
par: Vo, Quoc Thinh, et autres
Publié: (2025)
Neurobench: DCASE 2020 Acoustic Scene Classification benchmark on XyloAudio 2
par: Ke, Weijie, et autres
Publié: (2024)
par: Ke, Weijie, et autres
Publié: (2024)
The Interspeech 2024 Challenge on Speech Processing Using Discrete Units
par: Chang, Xuankai, et autres
Publié: (2024)
par: Chang, Xuankai, et autres
Publié: (2024)
The IEEE-IS2 2024 Music Packet Loss Concealment Challenge
par: Mezza, Alessandro Ilic, et autres
Publié: (2024)
par: Mezza, Alessandro Ilic, et autres
Publié: (2024)
The ICASSP 2024 Audio Deep Packet Loss Concealment Challenge
par: Diener, Lorenz, et autres
Publié: (2024)
par: Diener, Lorenz, et autres
Publié: (2024)
The VoiceMOS Challenge 2024: Beyond Speech Quality Prediction
par: Huang, Wen-Chin, et autres
Publié: (2024)
par: Huang, Wen-Chin, et autres
Publié: (2024)
AdaProj: Adaptively Scaled Angular Margin Subspace Projections for Anomalous Sound Detection with Auxiliary Classification Tasks
par: Wilkinghoff, Kevin
Publié: (2024)
par: Wilkinghoff, Kevin
Publié: (2024)
Leveraging Sound Source Trajectories for Universal Sound Separation
par: Wu, Donghang, et autres
Publié: (2024)
par: Wu, Donghang, et autres
Publié: (2024)
Documents similaires
-
Sound Scene Synthesis at the DCASE 2024 Challenge
par: Lagrange, Mathieu, et autres
Publié: (2025) -
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
par: Cornell, Samuele, et autres
Publié: (2024) -
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
par: Yasuda, Masahiro, et autres
Publié: (2025) -
Performance improvement of spatial semantic segmentation with enriched audio features and agent-based error correction for DCASE 2025 Challenge Task 4
par: Park, Jongyeon, et autres
Publié: (2025) -
Performance Improvement of Language-Queried Audio Source Separation Based on Caption Augmentation From Large Language Models for DCASE Challenge 2024 Task 9
par: Lee, Do Hyun, et autres
Publié: (2024)