Unified Acoustic Representations for Screening Neurological and Respiratory Pathologies from Voice
Fuente:
arXiv
Guardado en:
| Autores principales: | Piao, Ran, Lu, Yuan, Kemps, Hareld, Xia, Tong, Saeed, Aaqib |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
RespLLM: Unifying Audio and Text with Multimodal LLMs for Generalized Respiratory Health Prediction
por: Zhang, Yuwei, et al.
Publicado: (2024)
por: Zhang, Yuwei, et al.
Publicado: (2024)
Towards Open Respiratory Acoustic Foundation Models: Pretraining and Benchmarking
por: Zhang, Yuwei, et al.
Publicado: (2024)
por: Zhang, Yuwei, et al.
Publicado: (2024)
StethoLM: Audio Language Model for Cardiopulmonary Analysis Across Clinical Tasks
por: Wang, Yishan, et al.
Publicado: (2026)
por: Wang, Yishan, et al.
Publicado: (2026)
Towards Robust Assessment of Pathological Voices via Combined Low-Level Descriptors and Foundation Model Representations
por: Ariyanti, Whenty, et al.
Publicado: (2025)
por: Ariyanti, Whenty, et al.
Publicado: (2025)
AI-Driven Acoustic Voice Biomarker-Based Hierarchical Classification of Benign Laryngeal Voice Disorders from Sustained Vowels
por: Annabestani, Mohsen, et al.
Publicado: (2025)
por: Annabestani, Mohsen, et al.
Publicado: (2025)
Lightweight and Generalizable Acoustic Scene Representations via Contrastive Fine-Tuning and Distillation
por: Yuan, Kuang, et al.
Publicado: (2025)
por: Yuan, Kuang, et al.
Publicado: (2025)
A Multimodal Framework for Dementia Detection via Linguistic and Acoustic Representation Learning
por: Ilias, Loukas, et al.
Publicado: (2026)
por: Ilias, Loukas, et al.
Publicado: (2026)
RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity
por: Bertolino, Gaia A., et al.
Publicado: (2026)
por: Bertolino, Gaia A., et al.
Publicado: (2026)
Underwater Acoustic Target Recognition based on Smoothness-inducing Regularization and Spectrogram-based Data Augmentation
por: Xu, Ji, et al.
Publicado: (2023)
por: Xu, Ji, et al.
Publicado: (2023)
Benchmarking Representations for Speech, Music, and Acoustic Events
por: La Quatra, Moreno, et al.
Publicado: (2024)
por: La Quatra, Moreno, et al.
Publicado: (2024)
UniPACT: A Multimodal Framework for Prognostic Question Answering on Raw ECG and Structured EHR
por: Tang, Jialu, et al.
Publicado: (2026)
por: Tang, Jialu, et al.
Publicado: (2026)
Improving Deep Learning-based Respiratory Sound Analysis with Frequency Selection and Attention Mechanism
por: Fraihi, Nouhaila, et al.
Publicado: (2025)
por: Fraihi, Nouhaila, et al.
Publicado: (2025)
Generative Multi-modal Feedback for Singing Voice Synthesis Evaluation
por: Li, Xueyan, et al.
Publicado: (2025)
por: Li, Xueyan, et al.
Publicado: (2025)
Automated Dysphagia Screening Using Noninvasive Neck Acoustic Sensing
por: Chng, Jade, et al.
Publicado: (2026)
por: Chng, Jade, et al.
Publicado: (2026)
SVSNet+: Enhancing Speaker Voice Similarity Assessment Models with Representations from Speech Foundation Models
por: Yin, Chun, et al.
Publicado: (2024)
por: Yin, Chun, et al.
Publicado: (2024)
Reproducible Machine Learning-based Voice Pathology Detection: Introducing the Pitch Difference Feature
por: Vrba, Jan, et al.
Publicado: (2024)
por: Vrba, Jan, et al.
Publicado: (2024)
An AI-enabled Bias-Free Respiratory Disease Diagnosis Model using Cough Audio: A Case Study for COVID-19
por: Saeed, Tabish, et al.
Publicado: (2024)
por: Saeed, Tabish, et al.
Publicado: (2024)
Acoustic evaluation of a neural network dedicated to the detection of animal vocalisations
por: Rouch, Jérémy, et al.
Publicado: (2025)
por: Rouch, Jérémy, et al.
Publicado: (2025)
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
por: Chen, Wei, et al.
Publicado: (2025)
por: Chen, Wei, et al.
Publicado: (2025)
Single Microphone Own Voice Detection based on Simulated Transfer Functions for Hearing Aids
por: Mayuravaani, Mathuranathan, et al.
Publicado: (2026)
por: Mayuravaani, Mathuranathan, et al.
Publicado: (2026)
LibriVAD: A Scalable Open Dataset with Deep Learning Benchmarks for Voice Activity Detection
por: Stylianou, Ioannis, et al.
Publicado: (2025)
por: Stylianou, Ioannis, et al.
Publicado: (2025)
OpenVoice: Versatile Instant Voice Cloning
por: Qin, Zengyi, et al.
Publicado: (2023)
por: Qin, Zengyi, et al.
Publicado: (2023)
Hankel-FNO: Fast Underwater Acoustic Charting Via Physics-Encoded Fourier Neural Operator
por: Sun, Yifan, et al.
Publicado: (2025)
por: Sun, Yifan, et al.
Publicado: (2025)
Parameter-efficient Dual-encoder Architecture with Differentiable Choquet Integral Fusion for Underwater Acoustic Classification
por: Mohammadi, Amirmohammad, et al.
Publicado: (2026)
por: Mohammadi, Amirmohammad, et al.
Publicado: (2026)
Distributed Acoustic Sensing for Urban Traffic Monitoring: Spatio-Temporal Attention in Recurrent Neural Networks
por: Fakhruzi, Izhan, et al.
Publicado: (2026)
por: Fakhruzi, Izhan, et al.
Publicado: (2026)
MambaVoiceCloning: Efficient and Expressive Text-to-Speech via State-Space Modeling and Diffusion Control
por: Kumar, Sahil, et al.
Publicado: (2026)
por: Kumar, Sahil, et al.
Publicado: (2026)
Respiratory Disease Classification and Biometric Analysis Using Biosignals from Digital Stethoscopes
por: Casado, Constantino Álvarez, et al.
Publicado: (2023)
por: Casado, Constantino Álvarez, et al.
Publicado: (2023)
Voice Biomarkers for Depression and Anxiety
por: Abramenko, Oleksii, et al.
Publicado: (2026)
por: Abramenko, Oleksii, et al.
Publicado: (2026)
Disentangling Textual and Acoustic Features of Neural Speech Representations
por: Mohebbi, Hosein, et al.
Publicado: (2024)
por: Mohebbi, Hosein, et al.
Publicado: (2024)
Low-Resource Cross-Domain Singing Voice Synthesis via Reduced Self-Supervised Speech Representations
por: Kakoulidis, Panos, et al.
Publicado: (2024)
por: Kakoulidis, Panos, et al.
Publicado: (2024)
Temporal Convolution-based Hybrid Model Approach with Representation Learning for Real-Time Acoustic Anomaly Detection
por: Dissanayaka, Sahan, et al.
Publicado: (2024)
por: Dissanayaka, Sahan, et al.
Publicado: (2024)
Investigation for Relative Voice Impression Estimation
por: Fujita, Kenichi, et al.
Publicado: (2026)
por: Fujita, Kenichi, et al.
Publicado: (2026)
Electrocardiogram Report Generation and Question Answering via Retrieval-Augmented Self-Supervised Modeling
por: Tang, Jialu, et al.
Publicado: (2024)
por: Tang, Jialu, et al.
Publicado: (2024)
Electrocardiogram-Language Model for Few-Shot Question Answering with Meta Learning
por: Tang, Jialu, et al.
Publicado: (2024)
por: Tang, Jialu, et al.
Publicado: (2024)
Underwater-Art: Expanding Information Perspectives With Text Templates For Underwater Acoustic Target Recognition
por: Xie, Yuan, et al.
Publicado: (2023)
por: Xie, Yuan, et al.
Publicado: (2023)
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
por: Jin, Weifei, et al.
Publicado: (2025)
por: Jin, Weifei, et al.
Publicado: (2025)
Unsupervised Acoustic Scene Mapping Based on Acoustic Features and Dimensionality Reduction
por: Cohen, Idan, et al.
Publicado: (2023)
por: Cohen, Idan, et al.
Publicado: (2023)
Explainable Multi-Modal Deep Learning for Automatic Detection of Lung Diseases from Respiratory Audio Signals
por: Saky, S M Asiful Islam, et al.
Publicado: (2025)
por: Saky, S M Asiful Islam, et al.
Publicado: (2025)
Tri-MTL: A Triple Multitask Learning Approach for Respiratory Disease Diagnosis
por: Kim, June-Woo, et al.
Publicado: (2025)
por: Kim, June-Woo, et al.
Publicado: (2025)
Adaptive Test-Time Scaling for Zero-Shot Respiratory Audio Classification
por: Wang, Tsai-Ning, et al.
Publicado: (2026)
por: Wang, Tsai-Ning, et al.
Publicado: (2026)
Ejemplares similares
-
RespLLM: Unifying Audio and Text with Multimodal LLMs for Generalized Respiratory Health Prediction
por: Zhang, Yuwei, et al.
Publicado: (2024) -
Towards Open Respiratory Acoustic Foundation Models: Pretraining and Benchmarking
por: Zhang, Yuwei, et al.
Publicado: (2024) -
StethoLM: Audio Language Model for Cardiopulmonary Analysis Across Clinical Tasks
por: Wang, Yishan, et al.
Publicado: (2026) -
Towards Robust Assessment of Pathological Voices via Combined Low-Level Descriptors and Foundation Model Representations
por: Ariyanti, Whenty, et al.
Publicado: (2025) -
AI-Driven Acoustic Voice Biomarker-Based Hierarchical Classification of Benign Laryngeal Voice Disorders from Sustained Vowels
por: Annabestani, Mohsen, et al.
Publicado: (2025)