Learning Multi-Target TDOA Features for Sound Event Localization and Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Berg, Axel, Engman, Johanna, Gulin, Jens, Åström, Karl, Oskarsson, Magnus |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
wav2pos: Sound Source Localization using Masked Autoencoders
von: Berg, Axel, et al.
Veröffentlicht: (2024)
von: Berg, Axel, et al.
Veröffentlicht: (2024)
Sound Event Detection and Localization with Distance Estimation
von: Krause, Daniel Aleksander, et al.
Veröffentlicht: (2024)
von: Krause, Daniel Aleksander, et al.
Veröffentlicht: (2024)
The Impact of Semi-Supervised Learning on Line Segment Detection
von: Engman, Johanna, et al.
Veröffentlicht: (2024)
von: Engman, Johanna, et al.
Veröffentlicht: (2024)
SONNET: Enhancing Time Delay Estimation by Leveraging Simulated Audio
von: Tegler, Erik, et al.
Veröffentlicht: (2024)
von: Tegler, Erik, et al.
Veröffentlicht: (2024)
An Experimental Study on Joint Modeling for Sound Event Localization and Detection with Source Distance Estimation
von: Dong, Yuxuan, et al.
Veröffentlicht: (2025)
von: Dong, Yuxuan, et al.
Veröffentlicht: (2025)
Spatial Scaper: A Library to Simulate and Augment Soundscapes for Sound Event Localization and Detection in Realistic Rooms
von: Roman, Iran R., et al.
Veröffentlicht: (2024)
von: Roman, Iran R., et al.
Veröffentlicht: (2024)
Energy Consumption Trends in Sound Event Detection Systems
von: Douwes, Constance, et al.
Veröffentlicht: (2024)
von: Douwes, Constance, et al.
Veröffentlicht: (2024)
Feature Aggregation in Joint Sound Classification and Localization Neural Networks
von: Healy, Brendan, et al.
Veröffentlicht: (2023)
von: Healy, Brendan, et al.
Veröffentlicht: (2023)
RCT: Random Consistency Training for Semi-supervised Sound Event Detection
von: Shao, Nian, et al.
Veröffentlicht: (2021)
von: Shao, Nian, et al.
Veröffentlicht: (2021)
Exploring Performance-Complexity Trade-Offs in Sound Event Detection Models
von: Morocutti, Tobias, et al.
Veröffentlicht: (2025)
von: Morocutti, Tobias, et al.
Veröffentlicht: (2025)
From Weak to Strong Sound Event Labels using Adaptive Change-Point Detection and Active Learning
von: Martinsson, John, et al.
Veröffentlicht: (2024)
von: Martinsson, John, et al.
Veröffentlicht: (2024)
Text-Queried Target Sound Event Localization
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2024)
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2024)
Cross-Referencing Self-Training Network for Sound Event Detection in Audio Mixtures
von: Park, Sangwook, et al.
Veröffentlicht: (2021)
von: Park, Sangwook, et al.
Veröffentlicht: (2021)
Class-Incremental Learning for Sound Event Localization and Detection
von: Pandey, Ruchi, et al.
Veröffentlicht: (2024)
von: Pandey, Ruchi, et al.
Veröffentlicht: (2024)
Spatial and Semantic Embedding Integration for Stereo Sound Event Localization and Detection in Regular Videos
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
Aggregation Strategies for Efficient Annotation of Bioacoustic Sound Events Using Active Learning
von: Lindholm, Richard, et al.
Veröffentlicht: (2025)
von: Lindholm, Richard, et al.
Veröffentlicht: (2025)
MTDA-HSED: Mutual-Assistance Tuning and Dual-Branch Aggregating for Heterogeneous Sound Event Detection
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
SoundSculpt: Direction and Semantics Driven Ambisonic Target Sound Extraction
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
DOA-Aware Audio-Visual Self-Supervised Learning for Sound Event Localization and Detection
von: Fujita, Yoto, et al.
Veröffentlicht: (2024)
von: Fujita, Yoto, et al.
Veröffentlicht: (2024)
An Enhanced Audio Feature Tailored for Anomalous Sound Detection Based on Pre-trained Models
von: Zhong, Guirui, et al.
Veröffentlicht: (2025)
von: Zhong, Guirui, et al.
Veröffentlicht: (2025)
Hybrid Disagreement-Diversity Active Learning for Bioacoustic Sound Event Detection
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
HiSSNet: Sound Event Detection and Speaker Identification via Hierarchical Prototypical Networks for Low-Resource Headphones
von: Shashaank, N, et al.
Veröffentlicht: (2023)
von: Shashaank, N, et al.
Veröffentlicht: (2023)
Advanced Framework for Animal Sound Classification With Features Optimization
von: Yang, Qiang, et al.
Veröffentlicht: (2024)
von: Yang, Qiang, et al.
Veröffentlicht: (2024)
Respiratory Inhaler Sound Event Classification Using Self-Supervised Learning
von: Panah, Davoud Shariat, et al.
Veröffentlicht: (2025)
von: Panah, Davoud Shariat, et al.
Veröffentlicht: (2025)
Audio Simulation for Sound Source Localization in Virtual Evironment
von: Di Yuan, Yi, et al.
Veröffentlicht: (2024)
von: Di Yuan, Yi, et al.
Veröffentlicht: (2024)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds
von: Chang, Andrew, et al.
Veröffentlicht: (2025)
von: Chang, Andrew, et al.
Veröffentlicht: (2025)
ToS: A Team of Specialists ensemble framework for Stereo Sound Event Localization and Detection with distance estimation in Video
von: Berghi, Davide, et al.
Veröffentlicht: (2026)
von: Berghi, Davide, et al.
Veröffentlicht: (2026)
MVANet: Multi-Stage Video Attention Network for Sound Event Localization and Detection with Source Distance Estimation
von: Hong, Hengyi, et al.
Veröffentlicht: (2024)
von: Hong, Hengyi, et al.
Veröffentlicht: (2024)
Microphone Conversion: Mitigating Device Variability in Sound Event Classification
von: Ryu, Myeonghoon, et al.
Veröffentlicht: (2024)
von: Ryu, Myeonghoon, et al.
Veröffentlicht: (2024)
Activity-Guided Industrial Anomalous Sound Detection against Interferences
von: Lee, Yunjoo, et al.
Veröffentlicht: (2024)
von: Lee, Yunjoo, et al.
Veröffentlicht: (2024)
A Review on Sound Source Localization in Robotics: Focusing on Deep Learning Methods
von: Jalayer, Reza, et al.
Veröffentlicht: (2025)
von: Jalayer, Reza, et al.
Veröffentlicht: (2025)
Improving the Speaker Anonymization Evaluation's Robustness to Target Speakers with Adversarial Learning
von: Franzreb, Carlos, et al.
Veröffentlicht: (2025)
von: Franzreb, Carlos, et al.
Veröffentlicht: (2025)
Mixture of Mixups for Multi-label Classification of Rare Anuran Sounds
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Improving Stereo 3D Sound Event Localization and Detection: Perceptual Features, Stereo-specific Data Augmentation, and Distance Normalization
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
Regularized Contrastive Pre-training for Few-shot Bioacoustic Sound Detection
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
Patient-Aware Feature Alignment for Robust Lung Sound Classification:Cohesion-Separation and Global Alignment Losses
von: Jeong, Seung Gyu, et al.
Veröffentlicht: (2025)
von: Jeong, Seung Gyu, et al.
Veröffentlicht: (2025)
Zero- and Few-shot Sound Event Localization and Detection
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
von: Shimada, Kazuki, et al.
Veröffentlicht: (2023)
Targeted Augmented Data for Audio Deepfake Detection
von: Astrid, Marcella, et al.
Veröffentlicht: (2024)
von: Astrid, Marcella, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
wav2pos: Sound Source Localization using Masked Autoencoders
von: Berg, Axel, et al.
Veröffentlicht: (2024) -
Sound Event Detection and Localization with Distance Estimation
von: Krause, Daniel Aleksander, et al.
Veröffentlicht: (2024) -
The Impact of Semi-Supervised Learning on Line Segment Detection
von: Engman, Johanna, et al.
Veröffentlicht: (2024) -
SONNET: Enhancing Time Delay Estimation by Leveraging Simulated Audio
von: Tegler, Erik, et al.
Veröffentlicht: (2024) -
An Experimental Study on Joint Modeling for Sound Event Localization and Detection with Source Distance Estimation
von: Dong, Yuxuan, et al.
Veröffentlicht: (2025)