Empowering Multimodal Respiratory Sound Classification with Counterfactual Adversarial Debiasing for Out-of-Distribution Robustness
Fuente:
arXiv
Salvato in:
| Autori principali: | Koo, Heejoon, Toikkanen, Miika, Kim, Yoon Tae, Kim, Soo Yong, Kim, June-Woo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions
di: Koo, Heejoon, et al.
Pubblicazione: (2026)
di: Koo, Heejoon, et al.
Pubblicazione: (2026)
Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles
di: Toikkanen, Miika, et al.
Pubblicazione: (2025)
di: Toikkanen, Miika, et al.
Pubblicazione: (2025)
RepAugment: Input-Agnostic Representation-Level Augmentation for Respiratory Sound Classification
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
BTS: Bridging Text and Sound Modalities for Metadata-Aided Respiratory Sound Classification
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
Can Large Audio Language Models Ignore Multilingual Distractors? An Evaluation of Their Selective Auditory Attention Capabilities
di: Koo, Heejoon
Pubblicazione: (2026)
di: Koo, Heejoon
Pubblicazione: (2026)
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification
di: Bae, Sangmin, et al.
Pubblicazione: (2023)
di: Bae, Sangmin, et al.
Pubblicazione: (2023)
Meta-Ensemble Learning with Diverse Data Splits for Improved Respiratory Sound Classification
di: Kim, June-Woo, et al.
Pubblicazione: (2026)
di: Kim, June-Woo, et al.
Pubblicazione: (2026)
Self-Guided Target Sound Extraction and Classification Through Universal Sound Separation Model and Multiple Clues
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
Adversarial Multi-Task Learning for Disentangling Timbre and Pitch in Singing Voice Synthesis
di: Kim, Tae-Woo, et al.
Pubblicazione: (2022)
di: Kim, Tae-Woo, et al.
Pubblicazione: (2022)
AdaMER-CTC: Connectionist Temporal Classification with Adaptive Maximum Entropy Regularization for Automatic Speech Recognition
di: Eom, SooHwan, et al.
Pubblicazione: (2024)
di: Eom, SooHwan, et al.
Pubblicazione: (2024)
Adaptive Differential Denoising for Respiratory Sounds Classification
di: Dong, Gaoyang, et al.
Pubblicazione: (2025)
di: Dong, Gaoyang, et al.
Pubblicazione: (2025)
Sound Separation and Classification with Object and Semantic Guidance
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
FastEnhancer: Speed-Optimized Streaming Neural Speech Enhancement
di: Ahn, Sunghwan, et al.
Pubblicazione: (2025)
di: Ahn, Sunghwan, et al.
Pubblicazione: (2025)
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition
di: Yoon, Ji Won, et al.
Pubblicazione: (2022)
di: Yoon, Ji Won, et al.
Pubblicazione: (2022)
Disentangling Dual-Encoder Masked Autoencoder for Respiratory Sound Classification
di: Wei, Peidong, et al.
Pubblicazione: (2025)
di: Wei, Peidong, et al.
Pubblicazione: (2025)
Improving the Robustness and Clinical Applicability of Automatic Respiratory Sound Classification Using Deep Learning-Based Audio Enhancement: Algorithm Development and Validation
di: Tzeng, Jing-Tong, et al.
Pubblicazione: (2024)
di: Tzeng, Jing-Tong, et al.
Pubblicazione: (2024)
MF-PAM: Accurate Pitch Estimation through Periodicity Analysis and Multi-level Feature Fusion
di: Chung, Woo-Jin, et al.
Pubblicazione: (2023)
di: Chung, Woo-Jin, et al.
Pubblicazione: (2023)
Towards Maximum Likelihood Training for Transducer-based Streaming Speech Recognition
di: Lee, Hyeonseung, et al.
Pubblicazione: (2024)
di: Lee, Hyeonseung, et al.
Pubblicazione: (2024)
Patient-Aware Feature Alignment for Robust Lung Sound Classification:Cohesion-Separation and Global Alignment Losses
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2025)
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2025)
Paralinguistics-Aware Speech-Empowered Large Language Models for Natural Conversation
di: Kim, Heeseung, et al.
Pubblicazione: (2024)
di: Kim, Heeseung, et al.
Pubblicazione: (2024)
Raon-OpenTTS: Open Models and Data for Robust Text-to-Speech
di: Kim, Semin, et al.
Pubblicazione: (2026)
di: Kim, Semin, et al.
Pubblicazione: (2026)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
di: Lee, Dongheon, et al.
Pubblicazione: (2024)
di: Lee, Dongheon, et al.
Pubblicazione: (2024)
DISPATCH: Distilling Selective Patches for Speech Enhancement
di: Kim, Dohwan, et al.
Pubblicazione: (2025)
di: Kim, Dohwan, et al.
Pubblicazione: (2025)
VoiceGuider: Enhancing Out-of-Domain Performance in Parameter-Efficient Speaker-Adaptive Text-to-Speech via Autoguidance
di: Yeom, Jiheum, et al.
Pubblicazione: (2024)
di: Yeom, Jiheum, et al.
Pubblicazione: (2024)
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
FADEL: Uncertainty-aware Fake Audio Detection with Evidential Deep Learning
di: Kang, Ju Yeon, et al.
Pubblicazione: (2025)
di: Kang, Ju Yeon, et al.
Pubblicazione: (2025)
Towards Understanding of Frequency Dependence on Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
di: Nam, Hyeonuk, et al.
Pubblicazione: (2025)
Diversifying and Expanding Frequency-Adaptive Convolution Kernels for Sound Event Detection
di: Nam, Hyeonuk, et al.
Pubblicazione: (2024)
di: Nam, Hyeonuk, et al.
Pubblicazione: (2024)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
HILCodec: High-Fidelity and Lightweight Neural Audio Codec
di: Ahn, Sunghwan, et al.
Pubblicazione: (2024)
di: Ahn, Sunghwan, et al.
Pubblicazione: (2024)
Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation
di: Hsieh, Tsun-An, et al.
Pubblicazione: (2024)
di: Hsieh, Tsun-An, et al.
Pubblicazione: (2024)
Adapting a Text-to-Audio Model for Room Impulse Response Generation
di: Kim, Kirak, et al.
Pubblicazione: (2026)
di: Kim, Kirak, et al.
Pubblicazione: (2026)
Team HYU ASML ROBOVOX SP Cup 2024 System Description
di: Choi, Jeong-Hwan, et al.
Pubblicazione: (2024)
di: Choi, Jeong-Hwan, et al.
Pubblicazione: (2024)
High Fidelity Text-to-Speech Via Discrete Tokens Using Token Transducer and Group Masked Language Model
di: Lee, Joun Yeop, et al.
Pubblicazione: (2024)
di: Lee, Joun Yeop, et al.
Pubblicazione: (2024)
Patient Domain Supervised Contrastive Learning for Lung Sound Classification Using Mobile Phone
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2025)
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2025)
Lungmix: A Mixup-Based Strategy for Generalization in Respiratory Sound Classification
di: Ge, Shijia, et al.
Pubblicazione: (2024)
di: Ge, Shijia, et al.
Pubblicazione: (2024)
Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence Models
di: Lim, Yunkyu, et al.
Pubblicazione: (2025)
di: Lim, Yunkyu, et al.
Pubblicazione: (2025)
REVERB-FL: Server-Side Adversarial and Reserve-Enhanced Federated Learning for Robust Audio Classification
di: Peechara, Sathwika, et al.
Pubblicazione: (2025)
di: Peechara, Sathwika, et al.
Pubblicazione: (2025)
Noise-Agnostic Multitask Whisper Training for Reducing False Alarm Errors in Call-for-Help Detection
di: Ryu, Myeonghoon, et al.
Pubblicazione: (2025)
di: Ryu, Myeonghoon, et al.
Pubblicazione: (2025)
Inter-channel Conv-TasNet for multichannel speech enhancement
di: Lee, Dongheon, et al.
Pubblicazione: (2021)
di: Lee, Dongheon, et al.
Pubblicazione: (2021)
Documenti analoghi
-
Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions
di: Koo, Heejoon, et al.
Pubblicazione: (2026) -
Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles
di: Toikkanen, Miika, et al.
Pubblicazione: (2025) -
RepAugment: Input-Agnostic Representation-Level Augmentation for Respiratory Sound Classification
di: Kim, June-Woo, et al.
Pubblicazione: (2024) -
BTS: Bridging Text and Sound Modalities for Metadata-Aided Respiratory Sound Classification
di: Kim, June-Woo, et al.
Pubblicazione: (2024) -
Can Large Audio Language Models Ignore Multilingual Distractors? An Evaluation of Their Selective Auditory Attention Capabilities
di: Koo, Heejoon
Pubblicazione: (2026)