Environmental Sound Classification on An Embedded Hardware Platform

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bibbo, Gabriel, Singh, Arshdeep, Plumbley, Mark D.
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909723744272384
author Bibbo, Gabriel
Singh, Arshdeep
Plumbley, Mark D.
author_facet Bibbo, Gabriel
Singh, Arshdeep
Plumbley, Mark D.
contents Convolutional neural networks (CNNs) have exhibited state-of-the-art performance in various audio classification tasks. However, their real-time deployment remains a challenge on resource constrained devices such as embedded systems. In this paper, we analyze how the performance of large-scale pre-trained audio neural networks designed for audio pattern recognition changes when deployed on a hardware such as a Raspberry Pi. We empirically study the role of CPU temperature, microphone quality and audio signal volume on performance. Our experiments reveal that the continuous CPU usage results in an increased temperature that can trigger an automated slowdown mechanism in the Raspberry Pi, impacting inference latency. The quality of a microphone, specifically with affordable devices such as the Google AIY Voice Kit, and audio signal volume, all affect the system performance. In the course of our investigation, we encounter substantial complications linked to library compatibility and the unique processor architecture requirements of the Raspberry Pi, making the process less straightforward compared to conventional computers (PCs). Our observations, while presenting challenges, pave the way for future researchers to develop more compact machine learning models, design heat-dissipative hardware, and select appropriate microphones when AI models are deployed for real-time applications on edge devices.
format Preprint
id arxiv_https___arxiv_org_abs_2306_09106
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Environmental Sound Classification on An Embedded Hardware Platform
Bibbo, Gabriel
Singh, Arshdeep
Plumbley, Mark D.
Sound
Artificial Intelligence
Systems and Control
Audio and Speech Processing
Convolutional neural networks (CNNs) have exhibited state-of-the-art performance in various audio classification tasks. However, their real-time deployment remains a challenge on resource constrained devices such as embedded systems. In this paper, we analyze how the performance of large-scale pre-trained audio neural networks designed for audio pattern recognition changes when deployed on a hardware such as a Raspberry Pi. We empirically study the role of CPU temperature, microphone quality and audio signal volume on performance. Our experiments reveal that the continuous CPU usage results in an increased temperature that can trigger an automated slowdown mechanism in the Raspberry Pi, impacting inference latency. The quality of a microphone, specifically with affordable devices such as the Google AIY Voice Kit, and audio signal volume, all affect the system performance. In the course of our investigation, we encounter substantial complications linked to library compatibility and the unique processor architecture requirements of the Raspberry Pi, making the process less straightforward compared to conventional computers (PCs). Our observations, while presenting challenges, pave the way for future researchers to develop more compact machine learning models, design heat-dissipative hardware, and select appropriate microphones when AI models are deployed for real-time applications on edge devices.
title Environmental Sound Classification on An Embedded Hardware Platform
topic Sound
Artificial Intelligence
Systems and Control
Audio and Speech Processing
url https://arxiv.org/abs/2306.09106