Transformer Architectures for Respiratory Sound Analysis and Multimodal Diagnosis
Fuente:
arXiv
Saved in:
| Main Authors: | Aptekarev, Theodore, Sokolovsky, Vladimir, Furman, Gregory |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Analysis of Multi‐Component Echo Decay in Achilles Tendon by NMR Spectroscopy
by: Theodore Aptekarev, et al.
Published: (2026)
by: Theodore Aptekarev, et al.
Published: (2026)
Waveform-Logmel Audio Neural Networks for Respiratory Sound Classification
by: Xie, Jiadong, et al.
Published: (2025)
by: Xie, Jiadong, et al.
Published: (2025)
Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles
by: Toikkanen, Miika, et al.
Published: (2025)
by: Toikkanen, Miika, et al.
Published: (2025)
Multi-View Spectrogram Transformer for Respiratory Sound Classification
by: He, Wentao, et al.
Published: (2023)
by: He, Wentao, et al.
Published: (2023)
Abnormal Respiratory Sound Identification Using Audio-Spectrogram Vision Transformer
by: Ariyanti, Whenty, et al.
Published: (2024)
by: Ariyanti, Whenty, et al.
Published: (2024)
Improving Deep Learning-based Respiratory Sound Analysis with Frequency Selection and Attention Mechanism
by: Fraihi, Nouhaila, et al.
Published: (2025)
by: Fraihi, Nouhaila, et al.
Published: (2025)
Adaptive Differential Denoising for Respiratory Sounds Classification
by: Dong, Gaoyang, et al.
Published: (2025)
by: Dong, Gaoyang, et al.
Published: (2025)
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification
by: Bae, Sangmin, et al.
Published: (2023)
by: Bae, Sangmin, et al.
Published: (2023)
C2GA: A Class-Controllable Generative Augmentation Framework for Respiratory Sound Classification
by: Ma, Ziqi, et al.
Published: (2026)
by: Ma, Ziqi, et al.
Published: (2026)
Disentangling Dual-Encoder Masked Autoencoder for Respiratory Sound Classification
by: Wei, Peidong, et al.
Published: (2025)
by: Wei, Peidong, et al.
Published: (2025)
BTS: Bridging Text and Sound Modalities for Metadata-Aided Respiratory Sound Classification
by: Kim, June-Woo, et al.
Published: (2024)
by: Kim, June-Woo, et al.
Published: (2024)
Assessing the Utility of Audio Foundation Models for Heart and Respiratory Sound Analysis
by: Niizumi, Daisuke, et al.
Published: (2025)
by: Niizumi, Daisuke, et al.
Published: (2025)
XAI-Driven Spectral Analysis of Cough Sounds for Respiratory Disease Characterization
by: Amado-Caballero, Patricia, et al.
Published: (2025)
by: Amado-Caballero, Patricia, et al.
Published: (2025)
Pediatric Asthma Detection with Googles HeAR Model: An AI-Driven Respiratory Sound Classifier
by: Ehtesham, Abul, et al.
Published: (2025)
by: Ehtesham, Abul, et al.
Published: (2025)
HSDreport: Heart Sound Diagnosis with Echocardiography Reports
by: Zhao, Zihan, et al.
Published: (2024)
by: Zhao, Zihan, et al.
Published: (2024)
VoxMed: One-Step Respiratory Disease Classifier using Digital Stethoscope Sounds
by: Mundra, Paridhi, et al.
Published: (2024)
by: Mundra, Paridhi, et al.
Published: (2024)
Resp-Agent: An Agent-Based System for Multimodal Respiratory Sound Generation and Disease Diagnosis
by: Zhang, Pengfei, et al.
Published: (2026)
by: Zhang, Pengfei, et al.
Published: (2026)
Geometry-Aware Optimization for Respiratory Sound Classification: Enhancing Sensitivity with SAM-Optimized Audio Spectrogram Transformers
by: Işık, Atakan, et al.
Published: (2025)
by: Işık, Atakan, et al.
Published: (2025)
Lungmix: A Mixup-Based Strategy for Generalization in Respiratory Sound Classification
by: Ge, Shijia, et al.
Published: (2024)
by: Ge, Shijia, et al.
Published: (2024)
Tri-MTL: A Triple Multitask Learning Approach for Respiratory Disease Diagnosis
by: Kim, June-Woo, et al.
Published: (2025)
by: Kim, June-Woo, et al.
Published: (2025)
Metric Analysis for Spatial Semantic Segmentation of Sound Scenes
by: Mishra, Mayank, et al.
Published: (2025)
by: Mishra, Mayank, et al.
Published: (2025)
MARS-Sep: Multimodal-Aligned Reinforced Sound Separation
by: Zhang, Zihan, et al.
Published: (2025)
by: Zhang, Zihan, et al.
Published: (2025)
RepAugment: Input-Agnostic Representation-Level Augmentation for Respiratory Sound Classification
by: Kim, June-Woo, et al.
Published: (2024)
by: Kim, June-Woo, et al.
Published: (2024)
Estimating Respiratory Effort from Nocturnal Breathing Sounds for Obstructive Sleep Apnoea Screening
by: Xu, Xiaolei, et al.
Published: (2025)
by: Xu, Xiaolei, et al.
Published: (2025)
Multimodal Laryngoscopic Video Analysis for Assisted Diagnosis of Vocal Fold Paralysis
by: Zhang, Yucong, et al.
Published: (2024)
by: Zhang, Yucong, et al.
Published: (2024)
Vision Transformer Segmentation for Visual Bird Sound Denoising
by: Kumar, Sahil, et al.
Published: (2024)
by: Kumar, Sahil, et al.
Published: (2024)
SynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
by: Brade, Stephen, et al.
Published: (2023)
by: Brade, Stephen, et al.
Published: (2023)
PulmoVec: A Two-Stage Stacking Meta-Learning Architecture Built on the HeAR Foundation Model for Multi-Task Classification of Pediatric Respiratory Sounds
by: Akbasli, Izzet Turkalp, et al.
Published: (2026)
by: Akbasli, Izzet Turkalp, et al.
Published: (2026)
Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions
by: Koo, Heejoon, et al.
Published: (2026)
by: Koo, Heejoon, et al.
Published: (2026)
LISTEN: Lightweight Industrial Sound-representable Transformer for Edge Notification
by: Han, Changheon, et al.
Published: (2025)
by: Han, Changheon, et al.
Published: (2025)
Effective Pre-Training of Audio Transformers for Sound Event Detection
by: Schmid, Florian, et al.
Published: (2024)
by: Schmid, Florian, et al.
Published: (2024)
CycleGuardian: A Framework for Automatic RespiratorySound classification Based on Improved Deep clustering and Contrastive Learning
by: Chu, Yun, et al.
Published: (2025)
by: Chu, Yun, et al.
Published: (2025)
Sound Field Translation and Mixed Source Model for Virtual Applications with Perceptual Validation
by: Birnie, Lachlan, et al.
Published: (2020)
by: Birnie, Lachlan, et al.
Published: (2020)
Thinking with Sound: Audio Chain-of-Thought Enables Multimodal Reasoning in Large Audio-Language Models
by: Xiong, Zhen, et al.
Published: (2025)
by: Xiong, Zhen, et al.
Published: (2025)
Fine-Grained Engine Fault Sound Event Detection Using Multimodal Signals
by: Fedorishin, Dennis, et al.
Published: (2024)
by: Fedorishin, Dennis, et al.
Published: (2024)
Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases
by: Zhang, Pengfei, et al.
Published: (2024)
by: Zhang, Pengfei, et al.
Published: (2024)
Scaling to Multimodal and Multichannel Heart Sound Classification with Synthetic and Augmented Biosignals
by: Marocchi, Milan, et al.
Published: (2025)
by: Marocchi, Milan, et al.
Published: (2025)
Respiratory Disease Classification and Biometric Analysis Using Biosignals from Digital Stethoscopes
by: Casado, Constantino Álvarez, et al.
Published: (2023)
by: Casado, Constantino Álvarez, et al.
Published: (2023)
Instantaneous Spectra Analysis of Pulse Series -- Application to Lung Sounds with Abnormalities
by: Ishiyama, Fumihiko
Published: (2026)
by: Ishiyama, Fumihiko
Published: (2026)
Environmental Sound Deepfake Detection Challenge: An Overview
by: Yin, Han, et al.
Published: (2025)
by: Yin, Han, et al.
Published: (2025)
Similar Items
-
Analysis of Multi‐Component Echo Decay in Achilles Tendon by NMR Spectroscopy
by: Theodore Aptekarev, et al.
Published: (2026) -
Waveform-Logmel Audio Neural Networks for Respiratory Sound Classification
by: Xie, Jiadong, et al.
Published: (2025) -
Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles
by: Toikkanen, Miika, et al.
Published: (2025) -
Multi-View Spectrogram Transformer for Respiratory Sound Classification
by: He, Wentao, et al.
Published: (2023) -
Abnormal Respiratory Sound Identification Using Audio-Spectrogram Vision Transformer
by: Ariyanti, Whenty, et al.
Published: (2024)