PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hu, Jinbo, Cao, Yin, Wu, Ming, Kang, Fang, Yang, Feiran, Wang, Wenwu, Plumbley, Mark D., Yang, Jun |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
par: Hu, Jinbo, et autres
Publié: (2023)
par: Hu, Jinbo, et autres
Publié: (2023)
Leveraging Pre-trained AudioLDM for Sound Generation: A Benchmark Study
par: Yuan, Yi, et autres
Publié: (2023)
par: Yuan, Yi, et autres
Publié: (2023)
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
par: Bibbó, Gabriel, et autres
Publié: (2024)
par: Bibbó, Gabriel, et autres
Publié: (2024)
FlowSep: Language-Queried Sound Separation with Rectified Flow Matching
par: Yuan, Yi, et autres
Publié: (2024)
par: Yuan, Yi, et autres
Publié: (2024)
Exploring the User Experience of AI-Assisted Sound Searching Systems for Creative Workflows
par: Liu, Haohe, et autres
Publié: (2025)
par: Liu, Haohe, et autres
Publié: (2025)
EnvSDD: Benchmarking Environmental Sound Deepfake Detection
par: Yin, Han, et autres
Publié: (2025)
par: Yin, Han, et autres
Publié: (2025)
SALM: Spatial Audio Language Model with Structured Embeddings for Understanding and Editing
par: Hu, Jinbo, et autres
Publié: (2025)
par: Hu, Jinbo, et autres
Publié: (2025)
Region-Specific Audio Tagging for Spatial Sound
par: Zhao, Jinzheng, et autres
Publié: (2025)
par: Zhao, Jinzheng, et autres
Publié: (2025)
ASiT: Local-Global Audio Spectrogram vIsion Transformer for Event Classification
par: Atito, Sara, et autres
Publié: (2022)
par: Atito, Sara, et autres
Publié: (2022)
Efficient Audio Captioning with Encoder-Level Knowledge Distillation
par: Xu, Xuenan, et autres
Publié: (2024)
par: Xu, Xuenan, et autres
Publié: (2024)
Universal Sound Separation with Self-Supervised Audio Masked Autoencoder
par: Zhao, Junqi, et autres
Publié: (2024)
par: Zhao, Junqi, et autres
Publié: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
par: Xiao, Yang, et autres
Publié: (2024)
par: Xiao, Yang, et autres
Publié: (2024)
Sound-VECaps: Improving Audio Generation with Visual Enhanced Captions
par: Yuan, Yi, et autres
Publié: (2024)
par: Yuan, Yi, et autres
Publié: (2024)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
par: Chen, Yuanjian, et autres
Publié: (2025)
par: Chen, Yuanjian, et autres
Publié: (2025)
WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
par: Xiao, Yang, et autres
Publié: (2024)
par: Xiao, Yang, et autres
Publié: (2024)
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
par: Yin, Jun, et autres
Publié: (2025)
par: Yin, Jun, et autres
Publié: (2025)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
par: Lee, Gyeong-Tae
Publié: (2025)
par: Lee, Gyeong-Tae
Publié: (2025)
Effective Pre-Training of Audio Transformers for Sound Event Detection
par: Schmid, Florian, et autres
Publié: (2024)
par: Schmid, Florian, et autres
Publié: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
par: Yin, Han, et autres
Publié: (2024)
par: Yin, Han, et autres
Publié: (2024)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
par: Yang, Ningyuan, et autres
Publié: (2026)
par: Yang, Ningyuan, et autres
Publié: (2026)
MAT-SED: A Masked Audio Transformer with Masked-Reconstruction Based Pre-training for Sound Event Detection
par: Cai, Pengfei, et autres
Publié: (2024)
par: Cai, Pengfei, et autres
Publié: (2024)
BioDCASE 2026 Challenge Baseline for Cross-Domain Mosquito Species Classification
par: Hou, Yuanbo, et autres
Publié: (2026)
par: Hou, Yuanbo, et autres
Publié: (2026)
Pre-training Autoencoder for Acoustic Event Classification via Blinky
par: Liu, Xiaoyang, et autres
Publié: (2025)
par: Liu, Xiaoyang, et autres
Publié: (2025)
Zero- and Few-shot Sound Event Localization and Detection
par: Shimada, Kazuki, et autres
Publié: (2023)
par: Shimada, Kazuki, et autres
Publié: (2023)
w2v-SELD: A Sound Event Localization and Detection Framework for Self-Supervised Spatial Audio Pre-Training
par: Santos, Orlem Lima dos, et autres
Publié: (2023)
par: Santos, Orlem Lima dos, et autres
Publié: (2023)
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
par: Chen, Yuanjian, et autres
Publié: (2025)
par: Chen, Yuanjian, et autres
Publié: (2025)
An Enhanced Audio Feature Tailored for Anomalous Sound Detection Based on Pre-trained Models
par: Zhong, Guirui, et autres
Publié: (2025)
par: Zhong, Guirui, et autres
Publié: (2025)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
par: Zhang, Xueping, et autres
Publié: (2025)
par: Zhang, Xueping, et autres
Publié: (2025)
Exploring Differences between Human Perception and Model Inference in Audio Event Recognition
par: Tan, Yizhou, et autres
Publié: (2024)
par: Tan, Yizhou, et autres
Publié: (2024)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
par: Yeow, Jun-Wei, et autres
Publié: (2025)
par: Yeow, Jun-Wei, et autres
Publié: (2025)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
par: Gao, Wenmiao, et autres
Publié: (2025)
par: Gao, Wenmiao, et autres
Publié: (2025)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
par: Yin, Han, et autres
Publié: (2024)
par: Yin, Han, et autres
Publié: (2024)
Environmental Sound Classification on An Embedded Hardware Platform
par: Bibbo, Gabriel, et autres
Publié: (2023)
par: Bibbo, Gabriel, et autres
Publié: (2023)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
par: Xiao, Yang, et autres
Publié: (2024)
par: Xiao, Yang, et autres
Publié: (2024)
SemantiCodec: An Ultra Low Bitrate Semantic Audio Codec for General Sound
par: Liu, Haohe, et autres
Publié: (2024)
par: Liu, Haohe, et autres
Publié: (2024)
Towards Generating Diverse Audio Captions via Adversarial Training
par: Mei, Xinhao, et autres
Publié: (2022)
par: Mei, Xinhao, et autres
Publié: (2022)
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
par: Koga, Naoki, et autres
Publié: (2024)
par: Koga, Naoki, et autres
Publié: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
par: Lee, Gyeong-Tae, et autres
Publié: (2025)
par: Lee, Gyeong-Tae, et autres
Publié: (2025)
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
par: Roman, Adrian S., et autres
Publié: (2025)
par: Roman, Adrian S., et autres
Publié: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
par: Yu, Hogeon
Publié: (2025)
par: Yu, Hogeon
Publié: (2025)
Documents similaires
-
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
par: Hu, Jinbo, et autres
Publié: (2023) -
Leveraging Pre-trained AudioLDM for Sound Generation: A Benchmark Study
par: Yuan, Yi, et autres
Publié: (2023) -
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
par: Bibbó, Gabriel, et autres
Publié: (2024) -
FlowSep: Language-Queried Sound Separation with Rectified Flow Matching
par: Yuan, Yi, et autres
Publié: (2024) -
Exploring the User Experience of AI-Assisted Sound Searching Systems for Creative Workflows
par: Liu, Haohe, et autres
Publié: (2025)