Disentangling Dual-Encoder Masked Autoencoder for Respiratory Sound Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Peidong, Miao, Shiyu, Li, Lin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Differential Denoising for Respiratory Sounds Classification
von: Dong, Gaoyang, et al.
Veröffentlicht: (2025)
von: Dong, Gaoyang, et al.
Veröffentlicht: (2025)
Universal Sound Separation with Self-Supervised Audio Masked Autoencoder
von: Zhao, Junqi, et al.
Veröffentlicht: (2024)
von: Zhao, Junqi, et al.
Veröffentlicht: (2024)
Learning Multidimensional Disentangled Representations of Instrumental Sounds for Musical Similarity Assessment
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024)
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024)
BTS: Bridging Text and Sound Modalities for Metadata-Aided Respiratory Sound Classification
von: Kim, June-Woo, et al.
Veröffentlicht: (2024)
von: Kim, June-Woo, et al.
Veröffentlicht: (2024)
wav2pos: Sound Source Localization using Masked Autoencoders
von: Berg, Axel, et al.
Veröffentlicht: (2024)
von: Berg, Axel, et al.
Veröffentlicht: (2024)
Ideal-LLM: Integrating Dual Encoders and Language-Adapted LLM for Multilingual Speech-to-Text
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
Genuine-Focused Learning using Mask AutoEncoder for Generalized Fake Audio Detection
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
Learning Magnitude Distribution of Sound Fields via Conditioned Autoencoder
von: Koyama, Shoichi, et al.
Veröffentlicht: (2025)
von: Koyama, Shoichi, et al.
Veröffentlicht: (2025)
Audible Networks: Deconstructing and Manipulating Sounds with Deep Non-Negative Autoencoders
von: Burred, Juan José, et al.
Veröffentlicht: (2025)
von: Burred, Juan José, et al.
Veröffentlicht: (2025)
Disentangling Hierarchical Features for Anomalous Sound Detection Under Domain Shift
von: Guan, Jian, et al.
Veröffentlicht: (2025)
von: Guan, Jian, et al.
Veröffentlicht: (2025)
Lungmix: A Mixup-Based Strategy for Generalization in Respiratory Sound Classification
von: Ge, Shijia, et al.
Veröffentlicht: (2024)
von: Ge, Shijia, et al.
Veröffentlicht: (2024)
VoxMed: One-Step Respiratory Disease Classifier using Digital Stethoscope Sounds
von: Mundra, Paridhi, et al.
Veröffentlicht: (2024)
von: Mundra, Paridhi, et al.
Veröffentlicht: (2024)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
von: Yao, Shengshi, et al.
Veröffentlicht: (2025)
von: Yao, Shengshi, et al.
Veröffentlicht: (2025)
Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles
von: Toikkanen, Miika, et al.
Veröffentlicht: (2025)
von: Toikkanen, Miika, et al.
Veröffentlicht: (2025)
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
RepAugment: Input-Agnostic Representation-Level Augmentation for Respiratory Sound Classification
von: Kim, June-Woo, et al.
Veröffentlicht: (2024)
von: Kim, June-Woo, et al.
Veröffentlicht: (2024)
Pre-training Autoencoder for Acoustic Event Classification via Blinky
von: Liu, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoyang, et al.
Veröffentlicht: (2025)
ReFlow-TTS: A Rectified Flow Model for High-fidelity Text-to-Speech
von: Guan, Wenhao, et al.
Veröffentlicht: (2023)
von: Guan, Wenhao, et al.
Veröffentlicht: (2023)
Trainingless Adaptation of Pretrained Models for Environmental Sound Classification
von: Tonami, Noriyuki, et al.
Veröffentlicht: (2024)
von: Tonami, Noriyuki, et al.
Veröffentlicht: (2024)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
Adapting General Disentanglement-Based Speaker Anonymization for Enhanced Emotion Preservation
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
Learning Marmoset Vocal Patterns with a Masked Autoencoder for Robust Call Segmentation, Classification, and Caller Identification
von: Wu, Bin, et al.
Veröffentlicht: (2024)
von: Wu, Bin, et al.
Veröffentlicht: (2024)
Can Masked Autoencoders Also Listen to Birds?
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
Addressing Index Collapse of Large-Codebook Speech Tokenizer with Dual-Decoding Product-Quantized Variational Auto-Encoder
von: Guo, Haohan, et al.
Veröffentlicht: (2024)
von: Guo, Haohan, et al.
Veröffentlicht: (2024)
Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions
von: Koo, Heejoon, et al.
Veröffentlicht: (2026)
von: Koo, Heejoon, et al.
Veröffentlicht: (2026)
AFT: An Exemplar-Free Class Incremental Learning Method for Environmental Sound Classification
von: Chen, Xinyi, et al.
Veröffentlicht: (2025)
von: Chen, Xinyi, et al.
Veröffentlicht: (2025)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
Multi-View Spectrogram Transformer for Respiratory Sound Classification
von: He, Wentao, et al.
Veröffentlicht: (2023)
von: He, Wentao, et al.
Veröffentlicht: (2023)
DualVC 2: Dynamic Masked Convolution for Unified Streaming and Non-Streaming Voice Conversion
von: Ning, Ziqian, et al.
Veröffentlicht: (2023)
von: Ning, Ziqian, et al.
Veröffentlicht: (2023)
Evaluating CNN with Stacked Feature Representations and Audio Spectrogram Transformer Models for Sound Classification
von: Dehaghania, Parinaz Binandeh, et al.
Veröffentlicht: (2026)
von: Dehaghania, Parinaz Binandeh, et al.
Veröffentlicht: (2026)
Automatic Sound Event Detection and Classification of Great Ape Calls Using Neural Networks
von: Jiang, Zifan, et al.
Veröffentlicht: (2023)
von: Jiang, Zifan, et al.
Veröffentlicht: (2023)
Sound-Based Spin Estimation in Table Tennis: Dataset and Real-Time Classification Pipeline
von: Gossard, Thomas, et al.
Veröffentlicht: (2024)
von: Gossard, Thomas, et al.
Veröffentlicht: (2024)
Fully Few-shot Class-incremental Audio Classification Using Expandable Dual-embedding Extractor
von: Si, Yongjie, et al.
Veröffentlicht: (2024)
von: Si, Yongjie, et al.
Veröffentlicht: (2024)
Auden-Voice: General-Purpose Voice Encoder for Speech and Language Understanding
von: Huo, Mingyue, et al.
Veröffentlicht: (2025)
von: Huo, Mingyue, et al.
Veröffentlicht: (2025)
Disentangled Dual-Branch Graph Learning for Conversational Emotion Recognition
von: Guo, Chengling, et al.
Veröffentlicht: (2026)
von: Guo, Chengling, et al.
Veröffentlicht: (2026)
The ICME 2025 Audio Encoder Capability Challenge
von: Zhang, Junbo, et al.
Veröffentlicht: (2025)
von: Zhang, Junbo, et al.
Veröffentlicht: (2025)
Noisy Disentanglement with Tri-stage Training for Noise-Robust Speech Recognition
von: Chen, Shuangyuan, et al.
Veröffentlicht: (2025)
von: Chen, Shuangyuan, et al.
Veröffentlicht: (2025)
Residual Learning for Neural Ambisonics Encoders
von: Deppisch, Thomas, et al.
Veröffentlicht: (2026)
von: Deppisch, Thomas, et al.
Veröffentlicht: (2026)
SEF-PNet: Speaker Encoder-Free Personalized Speech Enhancement with Local and Global Contexts Aggregation
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
Speaker Disentanglement of Speech Pre-trained Model Based on Interpretability
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adaptive Differential Denoising for Respiratory Sounds Classification
von: Dong, Gaoyang, et al.
Veröffentlicht: (2025) -
Universal Sound Separation with Self-Supervised Audio Masked Autoencoder
von: Zhao, Junqi, et al.
Veröffentlicht: (2024) -
Learning Multidimensional Disentangled Representations of Instrumental Sounds for Musical Similarity Assessment
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024) -
BTS: Bridging Text and Sound Modalities for Metadata-Aided Respiratory Sound Classification
von: Kim, June-Woo, et al.
Veröffentlicht: (2024) -
wav2pos: Sound Source Localization using Masked Autoencoders
von: Berg, Axel, et al.
Veröffentlicht: (2024)