6DoF SELD: Sound Event Localization and Detection Using Microphones and Motion Tracking Sensors on self-motioning human
Fuente:
arXiv
Salvato in:
| Autori principali: | Yasuda, Masahiro, Saito, Shoichiro, Nakayama, Akira, Harada, Noboru |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Guided Masked Self-Distillation Modeling for Distributed Multimedia Sensor Event Analysis
di: Yasuda, Masahiro, et al.
Pubblicazione: (2024)
di: Yasuda, Masahiro, et al.
Pubblicazione: (2024)
Baseline Systems and Evaluation Metrics for Spatial Semantic Segmentation of Sound Scenes
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2025)
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2025)
w2v-SELD: A Sound Event Localization and Detection Framework for Self-Supervised Spatial Audio Pre-Training
di: Santos, Orlem Lima dos, et al.
Pubblicazione: (2023)
di: Santos, Orlem Lima dos, et al.
Pubblicazione: (2023)
SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation
di: Mu, Da, et al.
Pubblicazione: (2024)
di: Mu, Da, et al.
Pubblicazione: (2024)
Assessing the Utility of Audio Foundation Models for Heart and Respiratory Sound Analysis
di: Niizumi, Daisuke, et al.
Pubblicazione: (2025)
di: Niizumi, Daisuke, et al.
Pubblicazione: (2025)
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
CLAP-ART: Automated Audio Captioning with Semantic-rich Audio Representation Tokenizer
di: Takeuchi, Daiki, et al.
Pubblicazione: (2025)
di: Takeuchi, Daiki, et al.
Pubblicazione: (2025)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
di: Wang, Jiang, et al.
Pubblicazione: (2025)
di: Wang, Jiang, et al.
Pubblicazione: (2025)
Microphone Conversion: Mitigating Device Variability in Sound Event Classification
di: Ryu, Myeonghoon, et al.
Pubblicazione: (2024)
di: Ryu, Myeonghoon, et al.
Pubblicazione: (2024)
Blind Localization of Early Room Reflections with Arbitrary Microphone Array
di: Hadadi, Yogev, et al.
Pubblicazione: (2024)
di: Hadadi, Yogev, et al.
Pubblicazione: (2024)
M2D-CLAP: Masked Modeling Duo Meets CLAP for Learning General-purpose Audio-Language Representation
di: Niizumi, Daisuke, et al.
Pubblicazione: (2024)
di: Niizumi, Daisuke, et al.
Pubblicazione: (2024)
Zero- and Few-shot Sound Event Localization and Detection
di: Shimada, Kazuki, et al.
Pubblicazione: (2023)
di: Shimada, Kazuki, et al.
Pubblicazione: (2023)
Enhancing 1-Second 3D SELD Performance with Filter Bank Analysis and SCConv Integration in CST-Former
di: Zhang, Zhehui
Pubblicazione: (2024)
di: Zhang, Zhehui
Pubblicazione: (2024)
Towards Pre-training an Effective Respiratory Audio Foundation Model
di: Niizumi, Daisuke, et al.
Pubblicazione: (2025)
di: Niizumi, Daisuke, et al.
Pubblicazione: (2025)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
di: Zhang, Xueping, et al.
Pubblicazione: (2025)
di: Zhang, Xueping, et al.
Pubblicazione: (2025)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2026)
di: Nishida, Tomoya, et al.
Pubblicazione: (2026)
SonicBoom: Contact Localization Using Array of Microphones
di: Lee, Moonyoung, et al.
Pubblicazione: (2024)
di: Lee, Moonyoung, et al.
Pubblicazione: (2024)
Unrestricted Global Phase Bias-Aware Single-channel Speech Enhancement with Conformer-based Metric GAN
di: Zhang, Shiqi, et al.
Pubblicazione: (2024)
di: Zhang, Shiqi, et al.
Pubblicazione: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
di: Roman, Adrian S., et al.
Pubblicazione: (2025)
di: Roman, Adrian S., et al.
Pubblicazione: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
di: Yu, Hogeon
Pubblicazione: (2025)
di: Yu, Hogeon
Pubblicazione: (2025)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
di: Hu, Jinbo, et al.
Pubblicazione: (2023)
di: Hu, Jinbo, et al.
Pubblicazione: (2023)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization
di: Yang, Bing, et al.
Pubblicazione: (2024)
di: Yang, Bing, et al.
Pubblicazione: (2024)
Sound Event Bounding Boxes
di: Ebbers, Janek, et al.
Pubblicazione: (2024)
di: Ebbers, Janek, et al.
Pubblicazione: (2024)
Direction Estimation of Sound Sources Using Microphone Arrays and Signal Strength
di: Pour, Mahdi Ali, et al.
Pubblicazione: (2025)
di: Pour, Mahdi Ali, et al.
Pubblicazione: (2025)
Efficient and Microphone-Fault-Tolerant 3D Sound Source Localization
di: Yang, Yiyuan, et al.
Pubblicazione: (2025)
di: Yang, Yiyuan, et al.
Pubblicazione: (2025)
Deep Learning Based Stage-wise Two-dimensional Speaker Localization with Large Ad-hoc Microphone Arrays
di: Liu, Shupei, et al.
Pubblicazione: (2022)
di: Liu, Shupei, et al.
Pubblicazione: (2022)
Refining Knowledge Transfer on Audio-Image Temporal Agreement for Audio-Text Cross Retrieval
di: Tsubaki, Shunsuke, et al.
Pubblicazione: (2024)
di: Tsubaki, Shunsuke, et al.
Pubblicazione: (2024)
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2025)
di: Nishida, Tomoya, et al.
Pubblicazione: (2025)
Frequency-Domain Sound Field from the Perspective of Band-Limited Functions
di: Iwami, Takahiro, et al.
Pubblicazione: (2024)
di: Iwami, Takahiro, et al.
Pubblicazione: (2024)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
di: Imoto, Keisuke
Pubblicazione: (2025)
di: Imoto, Keisuke
Pubblicazione: (2025)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
di: Yeow, Jun-Wei, et al.
Pubblicazione: (2025)
di: Yeow, Jun-Wei, et al.
Pubblicazione: (2025)
Applying Automatic Differentiation to Optimize Differential Microphone Array Designs
di: Galougah, Siminfar Samakoush, et al.
Pubblicazione: (2024)
di: Galougah, Siminfar Samakoush, et al.
Pubblicazione: (2024)
Asynchronous Microphone Array Calibration using Hybrid TDOA Information
di: Zhang, Chengjie, et al.
Pubblicazione: (2024)
di: Zhang, Chengjie, et al.
Pubblicazione: (2024)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Music Demixing Track
di: Fabbro, Giorgio, et al.
Pubblicazione: (2023)
di: Fabbro, Giorgio, et al.
Pubblicazione: (2023)
AudioSpa: Spatializing Sound Events with Text
di: Feng, Linfeng, et al.
Pubblicazione: (2025)
di: Feng, Linfeng, et al.
Pubblicazione: (2025)
Frequency Dynamic Convolutions for Sound Event Detection
di: Nam, Hyeonuk
Pubblicazione: (2025)
di: Nam, Hyeonuk
Pubblicazione: (2025)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
di: Hu, Jinbo, et al.
Pubblicazione: (2024)
di: Hu, Jinbo, et al.
Pubblicazione: (2024)
Sound Event Detection and Localization with Distance Estimation
di: Krause, Daniel Aleksander, et al.
Pubblicazione: (2024)
di: Krause, Daniel Aleksander, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Guided Masked Self-Distillation Modeling for Distributed Multimedia Sensor Event Analysis
di: Yasuda, Masahiro, et al.
Pubblicazione: (2024) -
Baseline Systems and Evaluation Metrics for Spatial Semantic Segmentation of Sound Scenes
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2025) -
w2v-SELD: A Sound Event Localization and Detection Framework for Self-Supervised Spatial Audio Pre-Training
di: Santos, Orlem Lima dos, et al.
Pubblicazione: (2023) -
SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation
di: Mu, Da, et al.
Pubblicazione: (2024) -
Assessing the Utility of Audio Foundation Models for Heart and Respiratory Sound Analysis
di: Niizumi, Daisuke, et al.
Pubblicazione: (2025)