A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
Fuente:
arXiv
Salvato in:
| Autori principali: | Paterson, Mary, Moor, James, Cutillo, Luisa |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Detecting Throat Cancer from Speech Signals using Machine Learning: A Scoping Literature Review
di: Paterson, Mary, et al.
Pubblicazione: (2023)
di: Paterson, Mary, et al.
Pubblicazione: (2023)
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
di: Ballas, Aristotelis, et al.
Pubblicazione: (2023)
di: Ballas, Aristotelis, et al.
Pubblicazione: (2023)
Automated Measurement of Geniohyoid Muscle Thickness During Speech Using Deep Learning and Ultrasound
di: Myrgyyassov, Alisher, et al.
Pubblicazione: (2026)
di: Myrgyyassov, Alisher, et al.
Pubblicazione: (2026)
Screening method for early dementia using sound objects as voice biomarkers
di: Pluta, Adam, et al.
Pubblicazione: (2024)
di: Pluta, Adam, et al.
Pubblicazione: (2024)
Foundation Models for Bioacoustics -- a Comparative Review
di: Schwinger, Raphael, et al.
Pubblicazione: (2025)
di: Schwinger, Raphael, et al.
Pubblicazione: (2025)
Learning to rumble: Automated elephant call classification, detection and endpointing using deep architectures
di: Geldenhuys, Christiaan M., et al.
Pubblicazione: (2024)
di: Geldenhuys, Christiaan M., et al.
Pubblicazione: (2024)
From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings
di: Geldenhuys, Christiaan M., et al.
Pubblicazione: (2026)
di: Geldenhuys, Christiaan M., et al.
Pubblicazione: (2026)
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
di: Navine, Amanda K., et al.
Pubblicazione: (2024)
di: Navine, Amanda K., et al.
Pubblicazione: (2024)
WhaleVAD-BPN: Improving Baleen Whale Call Detection with Boundary Proposal Networks and Post-processing Optimisation
di: Geldenhuys, Christiaan M., et al.
Pubblicazione: (2025)
di: Geldenhuys, Christiaan M., et al.
Pubblicazione: (2025)
Learning to detect an animal sound from five examples
di: Nolasco, Inês, et al.
Pubblicazione: (2023)
di: Nolasco, Inês, et al.
Pubblicazione: (2023)
Fish Tracking, Counting, and Behaviour Analysis in Digital Aquaculture: A Comprehensive Survey
di: Cui, Meng, et al.
Pubblicazione: (2024)
di: Cui, Meng, et al.
Pubblicazione: (2024)
Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
di: Fonollosa, Laia Garrobé, et al.
Pubblicazione: (2024)
di: Fonollosa, Laia Garrobé, et al.
Pubblicazione: (2024)
Computational bioacoustics with deep learning: a review and roadmap
di: Stowell, Dan
Pubblicazione: (2021)
di: Stowell, Dan
Pubblicazione: (2021)
Adaptive Representations of Sound for Automatic Insect Recognition
di: Faiß, Marius, et al.
Pubblicazione: (2023)
di: Faiß, Marius, et al.
Pubblicazione: (2023)
Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example
di: Ghosh, Suhita, et al.
Pubblicazione: (2024)
di: Ghosh, Suhita, et al.
Pubblicazione: (2024)
Cochlear Wave Propagation and Dynamics in the Human Base and Apex: Model-Based Estimates from Noninvasive Measurements
di: Alkhairy, Samiya A
Pubblicazione: (2024)
di: Alkhairy, Samiya A
Pubblicazione: (2024)
Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases
di: Zhang, Pengfei, et al.
Pubblicazione: (2024)
di: Zhang, Pengfei, et al.
Pubblicazione: (2024)
Automatic detection of Mild Cognitive Impairment using high-dimensional acoustic features in spontaneous speech
di: Zhang, Cong, et al.
Pubblicazione: (2024)
di: Zhang, Cong, et al.
Pubblicazione: (2024)
Prosody of speech production in latent post-stroke aphasia
di: Zhang, Cong, et al.
Pubblicazione: (2024)
di: Zhang, Cong, et al.
Pubblicazione: (2024)
Multi Modal Information Fusion of Acoustic and Linguistic Data for Decoding Dairy Cow Vocalizations in Animal Welfare Assessment
di: Jobarteh, Bubacarr, et al.
Pubblicazione: (2024)
di: Jobarteh, Bubacarr, et al.
Pubblicazione: (2024)
A multimodal LLM for the non-invasive decoding of spoken text from brain recordings
di: Hmamouche, Youssef, et al.
Pubblicazione: (2024)
di: Hmamouche, Youssef, et al.
Pubblicazione: (2024)
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
di: Schäfer-Zimmermann, Julian C., et al.
Pubblicazione: (2024)
di: Schäfer-Zimmermann, Julian C., et al.
Pubblicazione: (2024)
Atrial Fibrillation Detection System via Acoustic Sensing for Mobile Phones
di: Liu, Xuanyu, et al.
Pubblicazione: (2024)
di: Liu, Xuanyu, et al.
Pubblicazione: (2024)
A Concept-based approach to Voice Disorder Detection
di: Ghia, Davide, et al.
Pubblicazione: (2025)
di: Ghia, Davide, et al.
Pubblicazione: (2025)
Improving Generalization for AI-Synthesized Voice Detection
di: Ren, Hainan, et al.
Pubblicazione: (2024)
di: Ren, Hainan, et al.
Pubblicazione: (2024)
OpenVoice: Versatile Instant Voice Cloning
di: Qin, Zengyi, et al.
Pubblicazione: (2023)
di: Qin, Zengyi, et al.
Pubblicazione: (2023)
Exploiting Longitudinal Speech Sessions via Voice Assistant Systems for Early Detection of Cognitive Decline
di: Qi, Kristin, et al.
Pubblicazione: (2024)
di: Qi, Kristin, et al.
Pubblicazione: (2024)
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
di: Du, Zongyang, et al.
Pubblicazione: (2025)
di: Du, Zongyang, et al.
Pubblicazione: (2025)
Voice-Driven Mortality Prediction in Hospitalized Heart Failure Patients: A Machine Learning Approach Enhanced with Diagnostic Biomarkers
di: Ahmadli, Nihat, et al.
Pubblicazione: (2024)
di: Ahmadli, Nihat, et al.
Pubblicazione: (2024)
The 2025 PNPL Competition: Speech Detection and Phoneme Classification in the LibriBrain Dataset
di: Landau, Gilad, et al.
Pubblicazione: (2025)
di: Landau, Gilad, et al.
Pubblicazione: (2025)
End-to-End Integration of Speech Separation and Voice Activity Detection for Low-Latency Diarization of Telephone Conversations
di: Morrone, Giovanni, et al.
Pubblicazione: (2023)
di: Morrone, Giovanni, et al.
Pubblicazione: (2023)
VANPY: Voice Analysis Framework
di: Koushnir, Gregory, et al.
Pubblicazione: (2025)
di: Koushnir, Gregory, et al.
Pubblicazione: (2025)
Tessellated Linear Model for Age Prediction from Voice
di: Alharthi, Dareen, et al.
Pubblicazione: (2025)
di: Alharthi, Dareen, et al.
Pubblicazione: (2025)
Leveraging cough sounds to optimize chest x-ray usage in low-resource settings
di: Philip, Alexander, et al.
Pubblicazione: (2024)
di: Philip, Alexander, et al.
Pubblicazione: (2024)
Compact Neural TTS Voices for Accessibility
di: Jain, Kunal, et al.
Pubblicazione: (2025)
di: Jain, Kunal, et al.
Pubblicazione: (2025)
Speech to Speech Synthesis for Voice Impersonation
di: Johnson, Bjorn, et al.
Pubblicazione: (2026)
di: Johnson, Bjorn, et al.
Pubblicazione: (2026)
Discrete Optimal Transport and Voice Conversion
di: Selitskiy, Anton, et al.
Pubblicazione: (2025)
di: Selitskiy, Anton, et al.
Pubblicazione: (2025)
Edge Intelligence for Wildlife Conservation: Real-Time Hornbill Call Classification Using TinyML
di: Hing, Kong Ka, et al.
Pubblicazione: (2025)
di: Hing, Kong Ka, et al.
Pubblicazione: (2025)
BiSinger: Bilingual Singing Voice Synthesis
di: Zhou, Huali, et al.
Pubblicazione: (2023)
di: Zhou, Huali, et al.
Pubblicazione: (2023)
Zero-shot Voice Conversion with Diffusion Transformers
di: Liu, Songting
Pubblicazione: (2024)
di: Liu, Songting
Pubblicazione: (2024)
Documenti analoghi
-
Detecting Throat Cancer from Speech Signals using Machine Learning: A Scoping Literature Review
di: Paterson, Mary, et al.
Pubblicazione: (2023) -
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
di: Ballas, Aristotelis, et al.
Pubblicazione: (2023) -
Automated Measurement of Geniohyoid Muscle Thickness During Speech Using Deep Learning and Ultrasound
di: Myrgyyassov, Alisher, et al.
Pubblicazione: (2026) -
Screening method for early dementia using sound objects as voice biomarkers
di: Pluta, Adam, et al.
Pubblicazione: (2024) -
Foundation Models for Bioacoustics -- a Comparative Review
di: Schwinger, Raphael, et al.
Pubblicazione: (2025)