Deep Learning for Tuberculosis Screening in a High-burden Setting using Cough Analysis and Speech Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Ning, Mirheidari, Bahman, Brown, Guy J., Sanjase, Nsala, Maimbolwa, Minyoi M., Chifwamba, Solomon, Muzazu, Seke, Muyoyeta, Monde, Kagujje, Mary |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Gender Disparities in Automatic Speech Recognition Technology
by: ElGhazaly, Hend, et al.
Published: (2025)
by: ElGhazaly, Hend, et al.
Published: (2025)
Tuberculosis Screening from Cough Audio: Baseline Models, Clinical Variables, and Uncertainty Quantification
by: Kafentzis, George P., et al.
Published: (2026)
by: Kafentzis, George P., et al.
Published: (2026)
HeAR -- Health Acoustic Representations
by: Baur, Sebastien, et al.
Published: (2024)
by: Baur, Sebastien, et al.
Published: (2024)
Predicting Tuberculosis from Real-World Cough Audio Recordings and Metadata
by: Kafentzis, George P., et al.
Published: (2023)
by: Kafentzis, George P., et al.
Published: (2023)
Estimating Respiratory Effort from Nocturnal Breathing Sounds for Obstructive Sleep Apnoea Screening
by: Xu, Xiaolei, et al.
Published: (2025)
by: Xu, Xiaolei, et al.
Published: (2025)
Early Dementia Detection Using Multiple Spontaneous Speech Prompts: The PROCESS Challenge
by: Tao, Fuxiang, et al.
Published: (2024)
by: Tao, Fuxiang, et al.
Published: (2024)
CoughViT: A Self-Supervised Vision Transformer for Cough Audio Representation Learning
by: Luong, Justin, et al.
Published: (2025)
by: Luong, Justin, et al.
Published: (2025)
CognoSpeak: an automatic, remote assessment of early cognitive decline in real-world conversational speech
by: Pahar, Madhurananda, et al.
Published: (2025)
by: Pahar, Madhurananda, et al.
Published: (2025)
Cough activity detection for automatic tuberculosis screening
by: van Vüren, Joshua Jansen, et al.
Published: (2026)
by: van Vüren, Joshua Jansen, et al.
Published: (2026)
COVID-19 Diagnosis from Cough Acoustics using ConvNets and Data Augmentation
by: Mahanta, Saranga Kingkor, et al.
Published: (2021)
by: Mahanta, Saranga Kingkor, et al.
Published: (2021)
Adapting Speech Foundation Models for Unified Multimodal Speech Recognition with Large Language Models
by: Zhang, Jing-Xuan, et al.
Published: (2025)
by: Zhang, Jing-Xuan, et al.
Published: (2025)
Cough-E: A multimodal, privacy-preserving cough detection algorithm for the edge
by: Albini, Stefano, et al.
Published: (2024)
by: Albini, Stefano, et al.
Published: (2024)
Activation Steering for Accent Adaptation in Speech Foundation Models
by: Sun, Jinuo, et al.
Published: (2026)
by: Sun, Jinuo, et al.
Published: (2026)
Using Speech Foundational Models in Loss Functions for Hearing Aid Speech Enhancement
by: Sutherland, Robert, et al.
Published: (2024)
by: Sutherland, Robert, et al.
Published: (2024)
Generative Speech Foundation Model Pretraining for High-Quality Speech Extraction and Restoration
by: Ku, Pin-Jui, et al.
Published: (2024)
by: Ku, Pin-Jui, et al.
Published: (2024)
Context-Driven Dynamic Pruning for Large Speech Foundation Models
by: Someki, Masao, et al.
Published: (2025)
by: Someki, Masao, et al.
Published: (2025)
Exploring Speech Foundation Models for Speaker Diarization Across Lifespan
by: Xu, Anfeng, et al.
Published: (2026)
by: Xu, Anfeng, et al.
Published: (2026)
Comparison of Classification Algorithms for COVID19 Detection using Cough Acoustic Signals
by: Erdoğan, Yunus Emre, et al.
Published: (2022)
by: Erdoğan, Yunus Emre, et al.
Published: (2022)
Predicting Cognitive Decline: A Multimodal AI Approach to Dementia Screening from Speech
by: Chi, Lei, et al.
Published: (2025)
by: Chi, Lei, et al.
Published: (2025)
Exploiting Foundation Models and Speech Enhancement for Parkinson's Disease Detection from Speech in Real-World Operative Conditions
by: La Quatra, Moreno, et al.
Published: (2024)
by: La Quatra, Moreno, et al.
Published: (2024)
Advancing Zero-Shot Open-Set Speech Deepfake Source Tracing
by: Chhibber, Manasi, et al.
Published: (2025)
by: Chhibber, Manasi, et al.
Published: (2025)
Prosody Labeling with Phoneme-BERT and Speech Foundation Models
by: Koriyama, Tomoki
Published: (2025)
by: Koriyama, Tomoki
Published: (2025)
Harnessing Smartwatch Microphone Sensors for Cough Detection and Classification
by: Jaiswal, Pranay, et al.
Published: (2024)
by: Jaiswal, Pranay, et al.
Published: (2024)
Vox-Profile: A Speech Foundation Model Benchmark for Characterizing Diverse Speaker and Speech Traits
by: Feng, Tiantian, et al.
Published: (2025)
by: Feng, Tiantian, et al.
Published: (2025)
Robust Speech and Natural Language Processing Models for Depression Screening
by: Lu, Y., et al.
Published: (2024)
by: Lu, Y., et al.
Published: (2024)
Speech-to-See: End-to-End Speech-Driven Open-Set Object Detection
by: Lu, Wenhuan, et al.
Published: (2025)
by: Lu, Wenhuan, et al.
Published: (2025)
Speech Recognition Rescoring with Large Speech-Text Foundation Models
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2024)
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2024)
What Do Speech Foundation Models Not Learn About Speech?
by: Waheed, Abdul, et al.
Published: (2024)
by: Waheed, Abdul, et al.
Published: (2024)
Self-Supervised Speech Quality Assessment (S3QA): Leveraging Speech Foundation Models for a Scalable Speech Quality Metric
by: Ogg, Mattson, et al.
Published: (2025)
by: Ogg, Mattson, et al.
Published: (2025)
FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
by: Guo, Hao-Han, et al.
Published: (2024)
by: Guo, Hao-Han, et al.
Published: (2024)
Speech-based Clinical Depression Screening: An Empirical Study
by: Chen, Yangbin, et al.
Published: (2024)
by: Chen, Yangbin, et al.
Published: (2024)
Hallucination Benchmark for Speech Foundation Models
by: Koudounas, Alkis, et al.
Published: (2025)
by: Koudounas, Alkis, et al.
Published: (2025)
XAI-Driven Spectral Analysis of Cough Sounds for Respiratory Disease Characterization
by: Amado-Caballero, Patricia, et al.
Published: (2025)
by: Amado-Caballero, Patricia, et al.
Published: (2025)
Transfer Learning for Paediatric Sleep Apnoea Detection Using Physiology-Guided Acoustic Models
by: Niu, Chaoyue, et al.
Published: (2025)
by: Niu, Chaoyue, et al.
Published: (2025)
Robust Nasality Representation Learning for Cleft Palate-Related Velopharyngeal Dysfunction Screening in Real-World Settings
by: Liu, Weixin, et al.
Published: (2026)
by: Liu, Weixin, et al.
Published: (2026)
On-the-fly Routing for Zero-shot MoE Speaker Adaptation of Speech Foundation Models for Dysarthric Speech Recognition
by: HU, Shujie, et al.
Published: (2025)
by: HU, Shujie, et al.
Published: (2025)
Learning Expressive Disentangled Speech Representations with Soft Speech Units and Adversarial Style Augmentation
by: Deng, Yimin, et al.
Published: (2024)
by: Deng, Yimin, et al.
Published: (2024)
Machine Unlearning in Speech Emotion Recognition via Forget Set Alone
by: Ren, Zhao, et al.
Published: (2025)
by: Ren, Zhao, et al.
Published: (2025)
Exploring Prediction Targets in Masked Pre-Training for Speech Foundation Models
by: Chen, Li-Wei, et al.
Published: (2024)
by: Chen, Li-Wei, et al.
Published: (2024)
Resource-Efficient Adaptation of Speech Foundation Models for Multi-Speaker ASR
by: Wang, Weiqing, et al.
Published: (2024)
by: Wang, Weiqing, et al.
Published: (2024)
Similar Items
-
Exploring Gender Disparities in Automatic Speech Recognition Technology
by: ElGhazaly, Hend, et al.
Published: (2025) -
Tuberculosis Screening from Cough Audio: Baseline Models, Clinical Variables, and Uncertainty Quantification
by: Kafentzis, George P., et al.
Published: (2026) -
HeAR -- Health Acoustic Representations
by: Baur, Sebastien, et al.
Published: (2024) -
Predicting Tuberculosis from Real-World Cough Audio Recordings and Metadata
by: Kafentzis, George P., et al.
Published: (2023) -
Estimating Respiratory Effort from Nocturnal Breathing Sounds for Obstructive Sleep Apnoea Screening
by: Xu, Xiaolei, et al.
Published: (2025)