Selfsupervised learning for pathological speech detection
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Sheikh, Shakeel Ahmad |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Impact of Speech Mode in Automatic Pathological Speech Detection
par: Sheikh, Shakeel A., et autres
Publié: (2024)
par: Sheikh, Shakeel A., et autres
Publié: (2024)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
par: Kaloga, Yacouba, et autres
Publié: (2024)
par: Kaloga, Yacouba, et autres
Publié: (2024)
Generalizable speech deepfake detection via meta-learned LoRA
par: Laakkonen, Janne, et autres
Publié: (2025)
par: Laakkonen, Janne, et autres
Publié: (2025)
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
par: Kim, Yunsik, et autres
Publié: (2025)
par: Kim, Yunsik, et autres
Publié: (2025)
Context-aware child-directed speech detection from long-form recordings
par: Charlot, Théo, et autres
Publié: (2026)
par: Charlot, Théo, et autres
Publié: (2026)
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
par: Maiti, Soumi, et autres
Publié: (2023)
par: Maiti, Soumi, et autres
Publié: (2023)
Towards the Synthesis of Non-speech Vocalizations
par: Hoq, Enjamamul, et autres
Publié: (2024)
par: Hoq, Enjamamul, et autres
Publié: (2024)
A contrastive-learning approach for auditory attention detection
par: Bajestan, Seyed Ali Alavi, et autres
Publié: (2024)
par: Bajestan, Seyed Ali Alavi, et autres
Publié: (2024)
A multimodal dynamical variational autoencoder for audiovisual speech representation learning
par: Sadok, Samir, et autres
Publié: (2023)
par: Sadok, Samir, et autres
Publié: (2023)
CR-CTC: Consistency regularization on CTC for improved speech recognition
par: Yao, Zengwei, et autres
Publié: (2024)
par: Yao, Zengwei, et autres
Publié: (2024)
Single-channel speech enhancement using learnable loss mixup
par: Chang, Oscar, et autres
Publié: (2023)
par: Chang, Oscar, et autres
Publié: (2023)
Zipformer: A faster and better encoder for automatic speech recognition
par: Yao, Zengwei, et autres
Publié: (2023)
par: Yao, Zengwei, et autres
Publié: (2023)
Robustifying automatic speech recognition by extracting slowly varying features
par: Pizarro, Matías, et autres
Publié: (2021)
par: Pizarro, Matías, et autres
Publié: (2021)
Late fusion ensembles for speech recognition on diverse input audio representations
par: Jezidžić, Marin, et autres
Publié: (2024)
par: Jezidžić, Marin, et autres
Publié: (2024)
Boosting keyword spotting through on-device learnable user speech characteristics
par: Cioflan, Cristian, et autres
Publié: (2024)
par: Cioflan, Cristian, et autres
Publié: (2024)
Acoustic characterization of speech rhythm: going beyond metrics with recurrent neural networks
par: Deloche, François, et autres
Publié: (2024)
par: Deloche, François, et autres
Publié: (2024)
Dementia classification from spontaneous speech using wrapper-based feature selection
par: Niemelä, Marko, et autres
Publié: (2025)
par: Niemelä, Marko, et autres
Publié: (2025)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
par: Pepino, Leonardo, et autres
Publié: (2024)
par: Pepino, Leonardo, et autres
Publié: (2024)
An Attention Long Short-Term Memory based system for automatic classification of speech intelligibility
par: Fernández-Díaz, Miguel, et autres
Publié: (2024)
par: Fernández-Díaz, Miguel, et autres
Publié: (2024)
SeMaScore : a new evaluation metric for automatic speech recognition tasks
par: Sasindran, Zitha, et autres
Publié: (2024)
par: Sasindran, Zitha, et autres
Publié: (2024)
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
par: Maharana, Sarthak Kumar, et autres
Publié: (2023)
par: Maharana, Sarthak Kumar, et autres
Publié: (2023)
CognoSpeak: an automatic, remote assessment of early cognitive decline in real-world conversational speech
par: Pahar, Madhurananda, et autres
Publié: (2025)
par: Pahar, Madhurananda, et autres
Publié: (2025)
Self-supervised learning of speech representations with Dutch archival data
par: Vaessen, Nik, et autres
Publié: (2025)
par: Vaessen, Nik, et autres
Publié: (2025)
Towards objective and interpretable speech disorder assessment: a comparative analysis of CNN and transformer-based models
par: Maisonneuve, Malo, et autres
Publié: (2024)
par: Maisonneuve, Malo, et autres
Publié: (2024)
Objective and subjective evaluation of speech enhancement methods in the UDASE task of the 7th CHiME challenge
par: Leglaive, Simon, et autres
Publié: (2024)
par: Leglaive, Simon, et autres
Publié: (2024)
U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation
par: Dang, Shaoxiang, et autres
Publié: (2024)
par: Dang, Shaoxiang, et autres
Publié: (2024)
Online speaker diarization of meetings guided by speech separation
par: Gruttadauria, Elio, et autres
Publié: (2024)
par: Gruttadauria, Elio, et autres
Publié: (2024)
A low latency attention module for streaming self-supervised speech representation learning
par: Ma, Jianbo, et autres
Publié: (2023)
par: Ma, Jianbo, et autres
Publié: (2023)
Robust detection of overlapping bioacoustic sound events
par: Mahon, Louis, et autres
Publié: (2025)
par: Mahon, Louis, et autres
Publié: (2025)
Cough activity detection for automatic tuberculosis screening
par: van Vüren, Joshua Jansen, et autres
Publié: (2026)
par: van Vüren, Joshua Jansen, et autres
Publié: (2026)
Denoising by neural network for muzzle blast detection
par: Pujol, Hadrien, et autres
Publié: (2025)
par: Pujol, Hadrien, et autres
Publié: (2025)
A vector quantized masked autoencoder for audiovisual speech emotion recognition
par: Sadok, Samir, et autres
Publié: (2023)
par: Sadok, Samir, et autres
Publié: (2023)
Introduction to speech recognition
par: Dauphin, Gabriel
Publié: (2024)
par: Dauphin, Gabriel
Publié: (2024)
Optimising MFCC parameters for the automatic detection of respiratory diseases
par: Yan, Yuyang, et autres
Publié: (2024)
par: Yan, Yuyang, et autres
Publié: (2024)
Towards generalizing deep-audio fake detection networks
par: Gasenzer, Konstantin, et autres
Publié: (2023)
par: Gasenzer, Konstantin, et autres
Publié: (2023)
Unsupervised outlier detection to improve bird audio dataset labels
par: Collins, Bruce
Publié: (2025)
par: Collins, Bruce
Publié: (2025)
In-context learning capabilities of Large Language Models to detect suicide risk among adolescents from speech transcripts
par: Roquefort, Filomene, et autres
Publié: (2025)
par: Roquefort, Filomene, et autres
Publié: (2025)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
par: Tabatabaee, Saba, et autres
Publié: (2026)
par: Tabatabaee, Saba, et autres
Publié: (2026)
The impact of non-target events in synthetic soundscapes for sound event detection
par: Ronchini, Francesca, et autres
Publié: (2021)
par: Ronchini, Francesca, et autres
Publié: (2021)
Synthetic data enables context-aware bioacoustic sound event detection
par: Hoffman, Benjamin, et autres
Publié: (2025)
par: Hoffman, Benjamin, et autres
Publié: (2025)
Documents similaires
-
Impact of Speech Mode in Automatic Pathological Speech Detection
par: Sheikh, Shakeel A., et autres
Publié: (2024) -
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
par: Kaloga, Yacouba, et autres
Publié: (2024) -
Generalizable speech deepfake detection via meta-learned LoRA
par: Laakkonen, Janne, et autres
Publié: (2025) -
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
par: Kim, Yunsik, et autres
Publié: (2025) -
Context-aware child-directed speech detection from long-form recordings
par: Charlot, Théo, et autres
Publié: (2026)