Scaling to Multimodal and Multichannel Heart Sound Classification with Synthetic and Augmented Biosignals
Fuente:
arXiv
Saved in:
| Main Authors: | Marocchi, Milan, Fynn, Matthew, Mandana, Kayapanda, Rong, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio
by: Abbott, Leigh, et al.
Published: (2024)
by: Abbott, Leigh, et al.
Published: (2024)
Noise-Robust Contrastive Learning with an MFCC-Conformer For Coronary Artery Disease Detection
by: Marocchi, Milan, et al.
Published: (2026)
by: Marocchi, Milan, et al.
Published: (2026)
Respiratory Disease Classification and Biometric Analysis Using Biosignals from Digital Stethoscopes
by: Casado, Constantino Álvarez, et al.
Published: (2023)
by: Casado, Constantino Álvarez, et al.
Published: (2023)
Practicality meets precision: Wearable vest with integrated multi-channel PCG sensors for effective coronary artery disease pre-screening
by: Fynn, Matthew, et al.
Published: (2024)
by: Fynn, Matthew, et al.
Published: (2024)
BUET Multi-disease Heart Sound Dataset: A Comprehensive Auscultation Dataset for Developing Computer-Aided Diagnostic Systems
by: Ali, Shams Nafisa, et al.
Published: (2024)
by: Ali, Shams Nafisa, et al.
Published: (2024)
Reconstruction of Sound Field through Diffusion Models
by: Miotello, Federico, et al.
Published: (2023)
by: Miotello, Federico, et al.
Published: (2023)
XAI-Driven Spectral Analysis of Cough Sounds for Respiratory Disease Characterization
by: Amado-Caballero, Patricia, et al.
Published: (2025)
by: Amado-Caballero, Patricia, et al.
Published: (2025)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
by: Diaz-Guerra, David, et al.
Published: (2023)
by: Diaz-Guerra, David, et al.
Published: (2023)
FunnelNet: An End-to-End Deep Learning Framework to Monitor Digital Heart Murmur in Real-Time
by: Jobayer, Md, et al.
Published: (2024)
by: Jobayer, Md, et al.
Published: (2024)
Remixing Music for Hearing Aids Using Ensemble of Fine-Tuned Source Separators
by: Daly, Matthew
Published: (2024)
by: Daly, Matthew
Published: (2024)
EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
by: Manivannan, Mithun, et al.
Published: (2024)
by: Manivannan, Mithun, et al.
Published: (2024)
CochCeps-Augment: A Novel Self-Supervised Contrastive Learning Using Cochlear Cepstrum-based Masking for Speech Emotion Recognition
by: Ziogas, Ioannis, et al.
Published: (2024)
by: Ziogas, Ioannis, et al.
Published: (2024)
Masked Symbol Modeling for Demodulation of Oversampled Baseband Communication Signals in Impulsive Noise-Dominated Channels
by: Bedir, Oguz, et al.
Published: (2025)
by: Bedir, Oguz, et al.
Published: (2025)
EEG-to-Voice Decoding of Spoken and Imagined speech Using Non-Invasive EEG
by: Park, Hanbeot, et al.
Published: (2025)
by: Park, Hanbeot, et al.
Published: (2025)
Real-Time Streamable Generative Speech Restoration with Flow Matching
by: Welker, Simon, et al.
Published: (2025)
by: Welker, Simon, et al.
Published: (2025)
O-EENC-SD: Efficient Online End-to-End Neural Clustering for Speaker Diarization
by: Gruttadauria, Elio, et al.
Published: (2025)
by: Gruttadauria, Elio, et al.
Published: (2025)
Brain-to-Speech: Prosody Feature Engineering and Transformer-Based Reconstruction
by: Al-Radhi, Mohammed Salah, et al.
Published: (2026)
by: Al-Radhi, Mohammed Salah, et al.
Published: (2026)
Time Series Diffusion Method: A Denoising Diffusion Probabilistic Model for Vibration Signal Generation
by: Yi, Haiming, et al.
Published: (2023)
by: Yi, Haiming, et al.
Published: (2023)
Automated Dysphagia Screening Using Noninvasive Neck Acoustic Sensing
by: Chng, Jade, et al.
Published: (2026)
by: Chng, Jade, et al.
Published: (2026)
Investigation of Feature Selection and Pooling Methods for Environmental Sound Classification
by: Dehaghani, Parinaz Binandeh, et al.
Published: (2025)
by: Dehaghani, Parinaz Binandeh, et al.
Published: (2025)
Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement
by: Yang, Yujie, et al.
Published: (2025)
by: Yang, Yujie, et al.
Published: (2025)
IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS
by: Sankar, Ashwin, et al.
Published: (2024)
by: Sankar, Ashwin, et al.
Published: (2024)
Pitch-Conditioned Instrument Sound Synthesis From an Interactive Timbre Latent Space
by: Limberg, Christian, et al.
Published: (2025)
by: Limberg, Christian, et al.
Published: (2025)
FlowDec: A flow-based full-band general audio codec with high perceptual quality
by: Welker, Simon, et al.
Published: (2025)
by: Welker, Simon, et al.
Published: (2025)
Expressive Acoustic Guitar Sound Synthesis with an Instrument-Specific Input Representation and Diffusion Outpainting
by: Kim, Hounsu, et al.
Published: (2024)
by: Kim, Hounsu, et al.
Published: (2024)
Classification of Heart Sounds Using Multi-Branch Deep Convolutional Network and LSTM-CNN
by: Latifi, Seyed Amir, et al.
Published: (2024)
by: Latifi, Seyed Amir, et al.
Published: (2024)
T-FOLEY: A Controllable Waveform-Domain Diffusion Model for Temporal-Event-Guided Foley Sound Synthesis
by: Chung, Yoonjin, et al.
Published: (2024)
by: Chung, Yoonjin, et al.
Published: (2024)
When Humans Growl and Birds Speak: High-Fidelity Voice Conversion from Human to Animal and Designed Sounds
by: Kang, Minsu, et al.
Published: (2025)
by: Kang, Minsu, et al.
Published: (2025)
ToS: A Team of Specialists ensemble framework for Stereo Sound Event Localization and Detection with distance estimation in Video
by: Berghi, Davide, et al.
Published: (2026)
by: Berghi, Davide, et al.
Published: (2026)
QiVC-Net: Quantum-Inspired Variational Convolutional Network, with Application to Biosignal Classification
by: Golnari, Amin, et al.
Published: (2025)
by: Golnari, Amin, et al.
Published: (2025)
Gaussian Process Regression of Steering Vectors With Physics-Aware Deep Composite Kernels for Augmented Listening
by: Di Carlo, Diego, et al.
Published: (2025)
by: Di Carlo, Diego, et al.
Published: (2025)
Frequency-Aware Masked Autoencoders for Multimodal Pretraining on Biosignals
by: Liu, Ran, et al.
Published: (2023)
by: Liu, Ran, et al.
Published: (2023)
Attentive Fusion: A Transformer-based Approach to Multimodal Hate Speech Detection
by: Mandal, Atanu, et al.
Published: (2024)
by: Mandal, Atanu, et al.
Published: (2024)
Scaling Transformers for Low-Bitrate High-Quality Speech Coding
by: Parker, Julian D, et al.
Published: (2024)
by: Parker, Julian D, et al.
Published: (2024)
Resampling Filter Design for Multirate Neural Audio Effect Processing
by: Carson, Alistair, et al.
Published: (2025)
by: Carson, Alistair, et al.
Published: (2025)
Joint Source-Environment Adaptation for Deep Learning-Based Underwater Acoustic Source Ranging
by: Kari, Dariush, et al.
Published: (2025)
by: Kari, Dariush, et al.
Published: (2025)
Adaptive Control Attention Network for Underwater Acoustic Localization and Domain Adaptation
by: Vo, Quoc Thinh, et al.
Published: (2025)
by: Vo, Quoc Thinh, et al.
Published: (2025)
Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models
by: Baoueb, Teysir, et al.
Published: (2025)
by: Baoueb, Teysir, et al.
Published: (2025)
Mismatch-Robust Underwater Acoustic Localization Using A Differentiable Modular Forward Model
by: Kari, Dariush, et al.
Published: (2025)
by: Kari, Dariush, et al.
Published: (2025)
Lightweight DNN for Full-Band Speech Denoising on Mobile Devices: Exploiting Long and Short Temporal Patterns
by: Drossos, Konstantinos, et al.
Published: (2025)
by: Drossos, Konstantinos, et al.
Published: (2025)
Similar Items
-
Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio
by: Abbott, Leigh, et al.
Published: (2024) -
Noise-Robust Contrastive Learning with an MFCC-Conformer For Coronary Artery Disease Detection
by: Marocchi, Milan, et al.
Published: (2026) -
Respiratory Disease Classification and Biometric Analysis Using Biosignals from Digital Stethoscopes
by: Casado, Constantino Álvarez, et al.
Published: (2023) -
Practicality meets precision: Wearable vest with integrated multi-channel PCG sensors for effective coronary artery disease pre-screening
by: Fynn, Matthew, et al.
Published: (2024) -
BUET Multi-disease Heart Sound Dataset: A Comprehensive Auscultation Dataset for Developing Computer-Aided Diagnostic Systems
by: Ali, Shams Nafisa, et al.
Published: (2024)