ENACT-Heart -- ENsemble-based Assessment Using CNN and Transformer on Heart Sounds

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Han, Jiho, Shaout, Adnan
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866929728759267328
author Han, Jiho
Shaout, Adnan
author_facet Han, Jiho
Shaout, Adnan
contents This study explores the application of Vision Transformer (ViT) principles in audio analysis, specifically focusing on heart sounds. This paper introduces ENACT-Heart - a novel ensemble approach that leverages the complementary strengths of Convolutional Neural Networks (CNN) and ViT through a Mixture of Experts (MoE) framework, achieving a remarkable classification accuracy of 97.52%. This outperforms the individual contributions of ViT (93.88%) and CNN (95.45%), demonstrating the potential for enhanced diagnostic accuracy in cardiovascular health monitoring. These results demonstrate the potential of ensemble methods in enhancing classification performance for cardiovascular health monitoring and diagnosis.
format Preprint
id arxiv_https___arxiv_org_abs_2502_16914
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle ENACT-Heart -- ENsemble-based Assessment Using CNN and Transformer on Heart Sounds
Han, Jiho
Shaout, Adnan
Sound
Artificial Intelligence
Audio and Speech Processing
This study explores the application of Vision Transformer (ViT) principles in audio analysis, specifically focusing on heart sounds. This paper introduces ENACT-Heart - a novel ensemble approach that leverages the complementary strengths of Convolutional Neural Networks (CNN) and ViT through a Mixture of Experts (MoE) framework, achieving a remarkable classification accuracy of 97.52%. This outperforms the individual contributions of ViT (93.88%) and CNN (95.45%), demonstrating the potential for enhanced diagnostic accuracy in cardiovascular health monitoring. These results demonstrate the potential of ensemble methods in enhancing classification performance for cardiovascular health monitoring and diagnosis.
title ENACT-Heart -- ENsemble-based Assessment Using CNN and Transformer on Heart Sounds
topic Sound
Artificial Intelligence
Audio and Speech Processing
url https://arxiv.org/abs/2502.16914