Cross-domain Sound Recognition for Efficient Underwater Data Analysis

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Park, Jeongsoo, Han, Dong-Gyun, La, Hyoung Sul, Lee, Sangmin, Han, Yoonchang, Yang, Eun-Jin
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917594698612736
author Park, Jeongsoo
Han, Dong-Gyun
La, Hyoung Sul
Lee, Sangmin
Han, Yoonchang
Yang, Eun-Jin
author_facet Park, Jeongsoo
Han, Dong-Gyun
La, Hyoung Sul
Lee, Sangmin
Han, Yoonchang
Yang, Eun-Jin
contents This paper presents a novel deep learning approach for analyzing massive underwater acoustic data by leveraging a model trained on a broad spectrum of non-underwater (aerial) sounds. Recognizing the challenge in labeling vast amounts of underwater data, we propose a two-fold methodology to accelerate this labor-intensive procedure. The first part of our approach involves PCA and UMAP visualization of the underwater data using the feature vectors of an aerial sound recognition model. This enables us to cluster the data in a two dimensional space and listen to points within these clusters to understand their defining characteristics. This innovative method simplifies the process of selecting candidate labels for further training. In the second part, we train a neural network model using both the selected underwater data and the non-underwater dataset. We conducted a quantitative analysis to measure the precision, recall, and F1 score of our model for recognizing airgun sounds, a common type of underwater sound. The F1 score achieved by our model exceeded 84.3%, demonstrating the effectiveness of our approach in analyzing underwater acoustic data. The methodology presented in this paper holds significant potential to reduce the amount of labor required in underwater data analysis and opens up new possibilities for further research in the field of cross-domain data analysis.
format Preprint
id arxiv_https___arxiv_org_abs_2309_03451
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Cross-domain Sound Recognition for Efficient Underwater Data Analysis
Park, Jeongsoo
Han, Dong-Gyun
La, Hyoung Sul
Lee, Sangmin
Han, Yoonchang
Yang, Eun-Jin
Sound
Machine Learning
Audio and Speech Processing
This paper presents a novel deep learning approach for analyzing massive underwater acoustic data by leveraging a model trained on a broad spectrum of non-underwater (aerial) sounds. Recognizing the challenge in labeling vast amounts of underwater data, we propose a two-fold methodology to accelerate this labor-intensive procedure. The first part of our approach involves PCA and UMAP visualization of the underwater data using the feature vectors of an aerial sound recognition model. This enables us to cluster the data in a two dimensional space and listen to points within these clusters to understand their defining characteristics. This innovative method simplifies the process of selecting candidate labels for further training. In the second part, we train a neural network model using both the selected underwater data and the non-underwater dataset. We conducted a quantitative analysis to measure the precision, recall, and F1 score of our model for recognizing airgun sounds, a common type of underwater sound. The F1 score achieved by our model exceeded 84.3%, demonstrating the effectiveness of our approach in analyzing underwater acoustic data. The methodology presented in this paper holds significant potential to reduce the amount of labor required in underwater data analysis and opens up new possibilities for further research in the field of cross-domain data analysis.
title Cross-domain Sound Recognition for Efficient Underwater Data Analysis
topic Sound
Machine Learning
Audio and Speech Processing
url https://arxiv.org/abs/2309.03451