Detecting Throat Cancer from Speech Signals using Machine Learning: A Scoping Literature Review
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Paterson, Mary, Moor, James, Cutillo, Luisa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
von: Paterson, Mary, et al.
Veröffentlicht: (2024)
von: Paterson, Mary, et al.
Veröffentlicht: (2024)
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
von: Kim, Yunsik, et al.
Veröffentlicht: (2025)
von: Kim, Yunsik, et al.
Veröffentlicht: (2025)
Acoustic and Machine Learning Methods for Speech-Based Suicide Risk Assessment: A Systematic Review
von: Marie, Ambre, et al.
Veröffentlicht: (2025)
von: Marie, Ambre, et al.
Veröffentlicht: (2025)
BenSParX: A Robust Explainable Machine Learning Framework for Parkinson's Disease Detection from Bengali Conversational Speech
von: Hossain, Riad, et al.
Veröffentlicht: (2025)
von: Hossain, Riad, et al.
Veröffentlicht: (2025)
Hearing Your Blood Sugar: Non-Invasive Glucose Measurement Through Simple Vocal Signals, Transforming any Speech into a Sensor with Machine Learning
von: Ahmadli, Nihat, et al.
Veröffentlicht: (2024)
von: Ahmadli, Nihat, et al.
Veröffentlicht: (2024)
Voice Signal Processing for Machine Learning. The Case of Speaker Isolation
von: Ganchev, Radan
Veröffentlicht: (2024)
von: Ganchev, Radan
Veröffentlicht: (2024)
Multi-Channel Replay Speech Detection using Acoustic Maps
von: Neri, Michael, et al.
Veröffentlicht: (2026)
von: Neri, Michael, et al.
Veröffentlicht: (2026)
Meta-Learning in Audio and Speech Processing: An End to End Comprehensive Review
von: Raimon, Athul, et al.
Veröffentlicht: (2024)
von: Raimon, Athul, et al.
Veröffentlicht: (2024)
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
Active Restoration of Lost Audio Signals Using Machine Learning and Latent Information
von: Cheddad, Zohra Adila, et al.
Veröffentlicht: (2021)
von: Cheddad, Zohra Adila, et al.
Veröffentlicht: (2021)
From Diet to Free Lunch: Estimating Auxiliary Signal Properties using Dynamic Pruning Masks in Speech Enhancement Networks
von: Miccini, Riccardo, et al.
Veröffentlicht: (2026)
von: Miccini, Riccardo, et al.
Veröffentlicht: (2026)
Comparative Analysis of Mel-Frequency Cepstral Coefficients and Wavelet Based Audio Signal Processing for Emotion Detection and Mental Health Assessment in Spoken Speech
von: Agbo, Idoko, et al.
Veröffentlicht: (2024)
von: Agbo, Idoko, et al.
Veröffentlicht: (2024)
The 2025 PNPL Competition: Speech Detection and Phoneme Classification in the LibriBrain Dataset
von: Landau, Gilad, et al.
Veröffentlicht: (2025)
von: Landau, Gilad, et al.
Veröffentlicht: (2025)
Learning Disentangled Speech Representations
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
Speech Emotion Detection Based on MFCC and CNN-LSTM Architecture
von: Ouyang, Qianhe
Veröffentlicht: (2025)
von: Ouyang, Qianhe
Veröffentlicht: (2025)
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
Recurrence-Based Nonlinear Vocal Dynamics as Digital Biomarkers for Depression Detection from Conversational Speech
von: Samanta, Himadri S
Veröffentlicht: (2026)
von: Samanta, Himadri S
Veröffentlicht: (2026)
Exploring WavLM Back-ends for Speech Spoofing and Deepfake Detection
von: Stourbe, Theophile, et al.
Veröffentlicht: (2024)
von: Stourbe, Theophile, et al.
Veröffentlicht: (2024)
Noise-aware Speech Enhancement using Diffusion Probabilistic Model
von: Hu, Yuchen, et al.
Veröffentlicht: (2023)
von: Hu, Yuchen, et al.
Veröffentlicht: (2023)
RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching
von: Yang, Jinhyeok, et al.
Veröffentlicht: (2026)
von: Yang, Jinhyeok, et al.
Veröffentlicht: (2026)
More Similar than Dissimilar: Modeling Annotators for Cross-Corpus Speech Emotion Recognition
von: Tavernor, James, et al.
Veröffentlicht: (2025)
von: Tavernor, James, et al.
Veröffentlicht: (2025)
Speech to Speech Synthesis for Voice Impersonation
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
RevRIR: Joint Reverberant Speech and Room Impulse Response Embedding using Contrastive Learning with Application to Room Shape Classification
von: Bitterman, Jacob, et al.
Veröffentlicht: (2024)
von: Bitterman, Jacob, et al.
Veröffentlicht: (2024)
ParaMETA: Towards Learning Disentangled Paralinguistic Speaking Styles Representations from Speech
von: Lou, Haowei, et al.
Veröffentlicht: (2026)
von: Lou, Haowei, et al.
Veröffentlicht: (2026)
Multiple Choice Learning for Efficient Speech Separation with Many Speakers
von: Perera, David, et al.
Veröffentlicht: (2024)
von: Perera, David, et al.
Veröffentlicht: (2024)
Non-Intrusive Speech Intelligibility Prediction for Hearing Aids using Whisper and Metadata
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
Enhancing Speech Emotion Recognition using Dynamic Spectral Features and Kalman Smoothing
von: Hizabri, Marouane El, et al.
Veröffentlicht: (2026)
von: Hizabri, Marouane El, et al.
Veröffentlicht: (2026)
Diffusion Synthesizer for Efficient Multilingual Speech to Speech Translation
von: Hirschkind, Nameer, et al.
Veröffentlicht: (2024)
von: Hirschkind, Nameer, et al.
Veröffentlicht: (2024)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
von: Vaessen, Nik, et al.
Veröffentlicht: (2024)
von: Vaessen, Nik, et al.
Veröffentlicht: (2024)
TextrolSpeech: A Text Style Control Speech Corpus With Codec Language Text-to-Speech Models
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
Investigating the Effects of Diffusion-based Conditional Generative Speech Models Used for Speech Enhancement on Dysarthric Speech
von: Reszka, Joanna, et al.
Veröffentlicht: (2024)
von: Reszka, Joanna, et al.
Veröffentlicht: (2024)
Detection of Electric Motor Damage Through Analysis of Sound Signals Using Bayesian Neural Networks
von: Bauer, Waldemar, et al.
Veröffentlicht: (2024)
von: Bauer, Waldemar, et al.
Veröffentlicht: (2024)
RepCodec: A Speech Representation Codec for Speech Tokenization
von: Huang, Zhichao, et al.
Veröffentlicht: (2023)
von: Huang, Zhichao, et al.
Veröffentlicht: (2023)
End-to-End Integration of Speech Separation and Voice Activity Detection for Low-Latency Diarization of Telephone Conversations
von: Morrone, Giovanni, et al.
Veröffentlicht: (2023)
von: Morrone, Giovanni, et al.
Veröffentlicht: (2023)
Towards Robust Overlapping Speech Detection: A Speaker-Aware Progressive Approach Using WavLM
von: Sun, Zhaokai, et al.
Veröffentlicht: (2025)
von: Sun, Zhaokai, et al.
Veröffentlicht: (2025)
Understanding Self-Supervised Learning of Speech Representation via Invariance and Redundancy Reduction
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation
von: Richter, Julius, et al.
Veröffentlicht: (2024)
von: Richter, Julius, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
von: Paterson, Mary, et al.
Veröffentlicht: (2024) -
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
von: Kim, Yunsik, et al.
Veröffentlicht: (2025) -
Acoustic and Machine Learning Methods for Speech-Based Suicide Risk Assessment: A Systematic Review
von: Marie, Ambre, et al.
Veröffentlicht: (2025) -
BenSParX: A Robust Explainable Machine Learning Framework for Parkinson's Disease Detection from Bengali Conversational Speech
von: Hossain, Riad, et al.
Veröffentlicht: (2025) -
Hearing Your Blood Sugar: Non-Invasive Glucose Measurement Through Simple Vocal Signals, Transforming any Speech into a Sensor with Machine Learning
von: Ahmadli, Nihat, et al.
Veröffentlicht: (2024)