MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Yeow, Jun-Wei, Tan, Ee-Leng, Peksi, Santi, Gan, Woon-Seng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Improving Stereo 3D Sound Event Localization and Detection: Perceptual Features, Stereo-specific Data Augmentation, and Distance Normalization
por: Yeow, Jun-Wei, et al.
Publicado: (2025)
por: Yeow, Jun-Wei, et al.
Publicado: (2025)
Squeeze-and-Excite ResNet-Conformers for Sound Event Localization, Detection, and Distance Estimation for DCASE 2024 Challenge
por: Yeow, Jun Wei, et al.
Publicado: (2024)
por: Yeow, Jun Wei, et al.
Publicado: (2024)
Enhancing Situational Awareness in Wearable Audio Devices Using a Lightweight Sound Event Localization and Detection System
por: Yeow, Jun-Wei, et al.
Publicado: (2025)
por: Yeow, Jun-Wei, et al.
Publicado: (2025)
Joint Feature and Output Distillation for Low-complexity Acoustic Scene Classification
por: Li, Haowen, et al.
Publicado: (2025)
por: Li, Haowen, et al.
Publicado: (2025)
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction
por: Yeo, Jin Jie Sean, et al.
Publicado: (2024)
por: Yeo, Jin Jie Sean, et al.
Publicado: (2024)
Acoustic Scene Classification Using CNN-GRU Model Without Knowledge Distillation
por: Tan, Ee-Leng, et al.
Publicado: (2025)
por: Tan, Ee-Leng, et al.
Publicado: (2025)
Extracting Urban Sound Information for Residential Areas in Smart Cities Using an End-to-End IoT System
por: Tan, Ee-Leng, et al.
Publicado: (2024)
por: Tan, Ee-Leng, et al.
Publicado: (2024)
A Real-Time Platform for Portable and Scalable Active Noise Mitigation for Construction Machinery
por: Gan, Woon-Seng, et al.
Publicado: (2024)
por: Gan, Woon-Seng, et al.
Publicado: (2024)
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
por: Zhao, Shengkui, et al.
Publicado: (2022)
por: Zhao, Shengkui, et al.
Publicado: (2022)
Sub-band and Full-band Interactive U-Net with DPRNN for Demixing Cross-talk Stereo Music
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
por: Ooi, Kenneth, et al.
Publicado: (2023)
por: Ooi, Kenneth, et al.
Publicado: (2023)
A Stabilized Hybrid Active Noise Control Algorithm of GFANC and FxNLMS with Online Clustering
por: Luo, Zhengding, et al.
Publicado: (2026)
por: Luo, Zhengding, et al.
Publicado: (2026)
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
por: Yin, Jun, et al.
Publicado: (2025)
por: Yin, Jun, et al.
Publicado: (2025)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
por: Hu, Jinbo, et al.
Publicado: (2023)
por: Hu, Jinbo, et al.
Publicado: (2023)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
por: Imoto, Keisuke
Publicado: (2025)
por: Imoto, Keisuke
Publicado: (2025)
Zero- and Few-shot Sound Event Localization and Detection
por: Shimada, Kazuki, et al.
Publicado: (2023)
por: Shimada, Kazuki, et al.
Publicado: (2023)
Learning Magnitude Distribution of Sound Fields via Conditioned Autoencoder
por: Koyama, Shoichi, et al.
Publicado: (2025)
por: Koyama, Shoichi, et al.
Publicado: (2025)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
por: Chen, Yuanjian, et al.
Publicado: (2025)
por: Chen, Yuanjian, et al.
Publicado: (2025)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
por: Zhang, Xueping, et al.
Publicado: (2025)
por: Zhang, Xueping, et al.
Publicado: (2025)
Effective Pre-Training of Audio Transformers for Sound Event Detection
por: Schmid, Florian, et al.
Publicado: (2024)
por: Schmid, Florian, et al.
Publicado: (2024)
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
por: Ooi, Kenneth, et al.
Publicado: (2022)
por: Ooi, Kenneth, et al.
Publicado: (2022)
w2v-SELD: A Sound Event Localization and Detection Framework for Self-Supervised Spatial Audio Pre-Training
por: Santos, Orlem Lima dos, et al.
Publicado: (2023)
por: Santos, Orlem Lima dos, et al.
Publicado: (2023)
Mind the Gap: Detecting Cluster Exits for Robust Local Density-Based Score Normalization in Anomalous Sound Detection
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
por: Hu, Jinbo, et al.
Publicado: (2024)
por: Hu, Jinbo, et al.
Publicado: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
por: Lee, Gyeong-Tae, et al.
Publicado: (2025)
por: Lee, Gyeong-Tae, et al.
Publicado: (2025)
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
por: Roman, Adrian S., et al.
Publicado: (2025)
por: Roman, Adrian S., et al.
Publicado: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
por: Yu, Hogeon
Publicado: (2025)
por: Yu, Hogeon
Publicado: (2025)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
por: Lee, Gyeong-Tae
Publicado: (2025)
por: Lee, Gyeong-Tae
Publicado: (2025)
Improving Audio Spectrogram Transformers for Sound Event Detection Through Multi-Stage Training
por: Schmid, Florian, et al.
Publicado: (2024)
por: Schmid, Florian, et al.
Publicado: (2024)
Sound Event Bounding Boxes
por: Ebbers, Janek, et al.
Publicado: (2024)
por: Ebbers, Janek, et al.
Publicado: (2024)
Self Training and Ensembling Frequency Dependent Networks with Coarse Prediction Pooling and Sound Event Bounding Boxes
por: Nam, Hyeonuk, et al.
Publicado: (2024)
por: Nam, Hyeonuk, et al.
Publicado: (2024)
AudioSpa: Spatializing Sound Events with Text
por: Feng, Linfeng, et al.
Publicado: (2025)
por: Feng, Linfeng, et al.
Publicado: (2025)
Frequency Dynamic Convolutions for Sound Event Detection
por: Nam, Hyeonuk
Publicado: (2025)
por: Nam, Hyeonuk
Publicado: (2025)
Temporal Pooling Strategies for Training-Free Anomalous Sound Detection with Self-Supervised Audio Embeddings
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
ConSep: a Noise- and Reverberation-Robust Speech Separation Framework by Magnitude Conditioning
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
Sound Zone Control Robust To Sound Speed Change
por: Bhattacharjee, Sankha Subhra, et al.
Publicado: (2024)
por: Bhattacharjee, Sankha Subhra, et al.
Publicado: (2024)
Sound Event Detection and Localization with Distance Estimation
por: Krause, Daniel Aleksander, et al.
Publicado: (2024)
por: Krause, Daniel Aleksander, et al.
Publicado: (2024)
Retrieval-Augmented Approach for Unsupervised Anomalous Sound Detection and Captioning without Model Training
por: Ogura, Ryoya, et al.
Publicado: (2024)
por: Ogura, Ryoya, et al.
Publicado: (2024)
Ejemplares similares
-
Improving Stereo 3D Sound Event Localization and Detection: Perceptual Features, Stereo-specific Data Augmentation, and Distance Normalization
por: Yeow, Jun-Wei, et al.
Publicado: (2025) -
Squeeze-and-Excite ResNet-Conformers for Sound Event Localization, Detection, and Distance Estimation for DCASE 2024 Challenge
por: Yeow, Jun Wei, et al.
Publicado: (2024) -
Enhancing Situational Awareness in Wearable Audio Devices Using a Lightweight Sound Event Localization and Detection System
por: Yeow, Jun-Wei, et al.
Publicado: (2025) -
Joint Feature and Output Distillation for Low-complexity Acoustic Scene Classification
por: Li, Haowen, et al.
Publicado: (2025) -
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction
por: Yeo, Jin Jie Sean, et al.
Publicado: (2024)