Sound Signal Synthesis with Auxiliary Classifier GAN, COVID-19 cough as an example
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Saleh, Yahya Sherif Solayman Mohamed, Dabbous, Ahmed Mohammed, Alkhaled, Lama, Chai, Hum Yan, Rana, Muhammad Ehsan, Mokayed, Hamam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vehicle Detection Performance in Nordic Region
von: Mokayed, Hamam, et al.
Veröffentlicht: (2024)
von: Mokayed, Hamam, et al.
Veröffentlicht: (2024)
Findings of MEGA: Maths Explanation with LLMs using the Socratic Method for Active Learning
von: Adewumi, Tosin, et al.
Veröffentlicht: (2025)
von: Adewumi, Tosin, et al.
Veröffentlicht: (2025)
Counterargument for Critical Thinking as Judged by AI and Humans
von: Adewumi, Tosin, et al.
Veröffentlicht: (2026)
von: Adewumi, Tosin, et al.
Veröffentlicht: (2026)
CMGAN: Conformer-based Metric GAN for Speech Enhancement
von: Cao, Ruizhe, et al.
Veröffentlicht: (2022)
von: Cao, Ruizhe, et al.
Veröffentlicht: (2022)
CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement
von: Abdulatif, Sherif, et al.
Veröffentlicht: (2022)
von: Abdulatif, Sherif, et al.
Veröffentlicht: (2022)
Uncertainty Calibration of Multi-Label Bird Sound Classifiers
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
TLDiffGAN: A Latent Diffusion-GAN Framework with Temporal Information Fusion for Anomalous Sound Detection
von: Ma, Chengyuan, et al.
Veröffentlicht: (2026)
von: Ma, Chengyuan, et al.
Veröffentlicht: (2026)
Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB)
von: Allauzen, Cyril, et al.
Veröffentlicht: (2026)
von: Allauzen, Cyril, et al.
Veröffentlicht: (2026)
mmWave Radar Aware Dual-Conditioned GAN for Speech Reconstruction of Signals With Low SNR
von: Karani, Jash, et al.
Veröffentlicht: (2026)
von: Karani, Jash, et al.
Veröffentlicht: (2026)
Massive Sound Embedding Benchmark (MSEB)
von: Heigold, Georg, et al.
Veröffentlicht: (2026)
von: Heigold, Georg, et al.
Veröffentlicht: (2026)
AdaProj: Adaptively Scaled Angular Margin Subspace Projections for Anomalous Sound Detection with Auxiliary Classification Tasks
von: Wilkinghoff, Kevin
Veröffentlicht: (2024)
von: Wilkinghoff, Kevin
Veröffentlicht: (2024)
Pediatric Asthma Detection with Googles HeAR Model: An AI-Driven Respiratory Sound Classifier
von: Ehtesham, Abul, et al.
Veröffentlicht: (2025)
von: Ehtesham, Abul, et al.
Veröffentlicht: (2025)
Sound Field Synthesis with Acoustic Waves
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
Instrumental Text-to-Music Generation with Auxiliary Conditioning Branches
von: Koh, Junyoung
Veröffentlicht: (2026)
von: Koh, Junyoung
Veröffentlicht: (2026)
VoxMed: One-Step Respiratory Disease Classifier using Digital Stethoscope Sounds
von: Mundra, Paridhi, et al.
Veröffentlicht: (2024)
von: Mundra, Paridhi, et al.
Veröffentlicht: (2024)
Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
Sign-to-Speech Prosody Transfer via Sign Reconstruction-based GAN
von: Manabe, Toranosuke, et al.
Veröffentlicht: (2026)
von: Manabe, Toranosuke, et al.
Veröffentlicht: (2026)
RankUp: Boosting Semi-Supervised Regression with an Auxiliary Ranking Classifier
von: Huang, Pin-Yen, et al.
Veröffentlicht: (2024)
von: Huang, Pin-Yen, et al.
Veröffentlicht: (2024)
Cross-Modal Binary Attention: An Energy-Efficient Fusion Framework for Audio-Visual Learning
von: Saleh, Mohamed, et al.
Veröffentlicht: (2026)
von: Saleh, Mohamed, et al.
Veröffentlicht: (2026)
Elastic Net Regularization and Gabor Dictionary for Classification of Heart Sound Signals using Deep Learning
von: Fakhry, Mahmoud, et al.
Veröffentlicht: (2026)
von: Fakhry, Mahmoud, et al.
Veröffentlicht: (2026)
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds
von: Chang, Andrew, et al.
Veröffentlicht: (2025)
von: Chang, Andrew, et al.
Veröffentlicht: (2025)
Fine-Grained Engine Fault Sound Event Detection Using Multimodal Signals
von: Fedorishin, Dennis, et al.
Veröffentlicht: (2024)
von: Fedorishin, Dennis, et al.
Veröffentlicht: (2024)
Towards reliable respiratory disease diagnosis based on cough sounds and vision transformers
von: Wang, Qian, et al.
Veröffentlicht: (2024)
von: Wang, Qian, et al.
Veröffentlicht: (2024)
Maximum Likelihood Estimation of the Direction of Sound In A Reverberant Noisy Environment
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
Sketch2Sound: Controllable Audio Generation via Time-Varying Signals and Sonic Imitations
von: García, Hugo Flores, et al.
Veröffentlicht: (2024)
von: García, Hugo Flores, et al.
Veröffentlicht: (2024)
From Diet to Free Lunch: Estimating Auxiliary Signal Properties using Dynamic Pruning Masks in Speech Enhancement Networks
von: Miccini, Riccardo, et al.
Veröffentlicht: (2026)
von: Miccini, Riccardo, et al.
Veröffentlicht: (2026)
Direction Estimation of Sound Sources Using Microphone Arrays and Signal Strength
von: Pour, Mahdi Ali, et al.
Veröffentlicht: (2025)
von: Pour, Mahdi Ali, et al.
Veröffentlicht: (2025)
Environmental Sound Deepfake Detection Challenge: An Overview
von: Yin, Han, et al.
Veröffentlicht: (2025)
von: Yin, Han, et al.
Veröffentlicht: (2025)
Transformer Architectures for Respiratory Sound Analysis and Multimodal Diagnosis
von: Aptekarev, Theodore, et al.
Veröffentlicht: (2026)
von: Aptekarev, Theodore, et al.
Veröffentlicht: (2026)
Localization of Sound Sources in a Room with One Microphone
von: Tukuljac, Helena Peic, et al.
Veröffentlicht: (2017)
von: Tukuljac, Helena Peic, et al.
Veröffentlicht: (2017)
Metric Analysis for Spatial Semantic Segmentation of Sound Scenes
von: Mishra, Mayank, et al.
Veröffentlicht: (2025)
von: Mishra, Mayank, et al.
Veröffentlicht: (2025)
Latent Multi-view Learning for Robust Environmental Sound Representations
von: Ding, Sivan, et al.
Veröffentlicht: (2025)
von: Ding, Sivan, et al.
Veröffentlicht: (2025)
Multi-scale Scanning Network for Machine Anomalous Sound Detection
von: Zhang, Yucong, et al.
Veröffentlicht: (2025)
von: Zhang, Yucong, et al.
Veröffentlicht: (2025)
Waveform-Logmel Audio Neural Networks for Respiratory Sound Classification
von: Xie, Jiadong, et al.
Veröffentlicht: (2025)
von: Xie, Jiadong, et al.
Veröffentlicht: (2025)
Enhanced Heart Sound Classification Using Mel Frequency Cepstral Coefficients and Comparative Analysis of Single vs. Ensemble Classifier Strategies
von: Rahmani, Amir Masoud, et al.
Veröffentlicht: (2024)
von: Rahmani, Amir Masoud, et al.
Veröffentlicht: (2024)
JenGAN: Stacked Shifted Filters in GAN-Based Speech Synthesis
von: Cho, Hyunjae, et al.
Veröffentlicht: (2024)
von: Cho, Hyunjae, et al.
Veröffentlicht: (2024)
Mix2Morph: Learning Sound Morphing from Noisy Mixes
von: Chu, Annie, et al.
Veröffentlicht: (2026)
von: Chu, Annie, et al.
Veröffentlicht: (2026)
ESDD 2026: Environmental Sound Deepfake Detection Challenge Evaluation Plan
von: Yin, Han, et al.
Veröffentlicht: (2025)
von: Yin, Han, et al.
Veröffentlicht: (2025)
Audiocards: Structured Metadata Improves Audio Language Models For Sound Design
von: Sridhar, Sripathi, et al.
Veröffentlicht: (2026)
von: Sridhar, Sripathi, et al.
Veröffentlicht: (2026)
Ultra-Lightweight Network for Ship-Radiated Sound Classification on Embedded Deployment
von: Park, Sangwon, et al.
Veröffentlicht: (2026)
von: Park, Sangwon, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Vehicle Detection Performance in Nordic Region
von: Mokayed, Hamam, et al.
Veröffentlicht: (2024) -
Findings of MEGA: Maths Explanation with LLMs using the Socratic Method for Active Learning
von: Adewumi, Tosin, et al.
Veröffentlicht: (2025) -
Counterargument for Critical Thinking as Judged by AI and Humans
von: Adewumi, Tosin, et al.
Veröffentlicht: (2026) -
CMGAN: Conformer-based Metric GAN for Speech Enhancement
von: Cao, Ruizhe, et al.
Veröffentlicht: (2022) -
CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement
von: Abdulatif, Sherif, et al.
Veröffentlicht: (2022)