Recurrence-Based Nonlinear Vocal Dynamics as Digital Biomarkers for Depression Detection from Conversational Speech
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Samanta, Himadri S |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vocal Melody Construction for Persian Lyrics Using LSTM Recurrent Neural Networks
von: Jafari, Farshad, et al.
Veröffentlicht: (2024)
von: Jafari, Farshad, et al.
Veröffentlicht: (2024)
Dynamic Gated Recurrent Neural Network for Compute-efficient Speech Enhancement
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
Extract and Diffuse: Latent Integration for Improved Diffusion-based Speech and Vocal Enhancement
von: Yang, Yudong, et al.
Veröffentlicht: (2024)
von: Yang, Yudong, et al.
Veröffentlicht: (2024)
Switchboard-Affect: Emotion Perception Labels from Conversational Speech
von: Romana, Amrit, et al.
Veröffentlicht: (2025)
von: Romana, Amrit, et al.
Veröffentlicht: (2025)
End-to-End Integration of Speech Separation and Voice Activity Detection for Low-Latency Diarization of Telephone Conversations
von: Morrone, Giovanni, et al.
Veröffentlicht: (2023)
von: Morrone, Giovanni, et al.
Veröffentlicht: (2023)
Test-Time Training for Depression Detection
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
Towards the Synthesis of Non-speech Vocalizations
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
BenSParX: A Robust Explainable Machine Learning Framework for Parkinson's Disease Detection from Bengali Conversational Speech
von: Hossain, Riad, et al.
Veröffentlicht: (2025)
von: Hossain, Riad, et al.
Veröffentlicht: (2025)
Speech Emotion Detection Based on MFCC and CNN-LSTM Architecture
von: Ouyang, Qianhe
Veröffentlicht: (2025)
von: Ouyang, Qianhe
Veröffentlicht: (2025)
Feature Representations for Automatic Meerkat Vocalization Classification
von: Mahmoud, Imen Ben, et al.
Veröffentlicht: (2024)
von: Mahmoud, Imen Ben, et al.
Veröffentlicht: (2024)
Hearing Your Blood Sugar: Non-Invasive Glucose Measurement Through Simple Vocal Signals, Transforming any Speech into a Sensor with Machine Learning
von: Ahmadli, Nihat, et al.
Veröffentlicht: (2024)
von: Ahmadli, Nihat, et al.
Veröffentlicht: (2024)
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
Self-Supervised Embeddings for Detecting Individual Symptoms of Depression
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
Evaluation of Speech Foundation Models for ASR on Child-Adult Conversations in Autism Diagnostic Sessions
von: Ashvin, Aditya, et al.
Veröffentlicht: (2024)
von: Ashvin, Aditya, et al.
Veröffentlicht: (2024)
JaCappella Corpus: A Japanese a Cappella Vocal Ensemble Corpus
von: Nakamura, Tomohiko, et al.
Veröffentlicht: (2022)
von: Nakamura, Tomohiko, et al.
Veröffentlicht: (2022)
Machine Learning Approaches to Vocal Register Classification in Contemporary Male Pop Music
von: Kim, Alexander, et al.
Veröffentlicht: (2025)
von: Kim, Alexander, et al.
Veröffentlicht: (2025)
Scalable Speech Enhancement with Dynamic Channel Pruning
von: Miccini, Riccardo, et al.
Veröffentlicht: (2024)
von: Miccini, Riccardo, et al.
Veröffentlicht: (2024)
Improving Inference-Time Optimisation for Vocal Effects Style Transfer with a Gaussian Prior
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
Dynamic-SUPERB: Towards A Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech
von: Huang, Chien-yu, et al.
Veröffentlicht: (2023)
von: Huang, Chien-yu, et al.
Veröffentlicht: (2023)
Multi-Channel Replay Speech Detection using Acoustic Maps
von: Neri, Michael, et al.
Veröffentlicht: (2026)
von: Neri, Michael, et al.
Veröffentlicht: (2026)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
A Real-Time Lyrics Alignment System Using Chroma And Phonetic Features For Classical Vocal Performance
von: Park, Jiyun, et al.
Veröffentlicht: (2024)
von: Park, Jiyun, et al.
Veröffentlicht: (2024)
Exploring WavLM Back-ends for Speech Spoofing and Deepfake Detection
von: Stourbe, Theophile, et al.
Veröffentlicht: (2024)
von: Stourbe, Theophile, et al.
Veröffentlicht: (2024)
Speech as a Biomarker for Disease Detection
von: Botelho, Catarina, et al.
Veröffentlicht: (2024)
von: Botelho, Catarina, et al.
Veröffentlicht: (2024)
Robust and Explainable Depression Identification from Speech Using Vowel-Based Ensemble Learning Approaches
von: Feng, Kexin, et al.
Veröffentlicht: (2024)
von: Feng, Kexin, et al.
Veröffentlicht: (2024)
Detecting Throat Cancer from Speech Signals using Machine Learning: A Scoping Literature Review
von: Paterson, Mary, et al.
Veröffentlicht: (2023)
von: Paterson, Mary, et al.
Veröffentlicht: (2023)
Enhancing Speech Emotion Recognition using Dynamic Spectral Features and Kalman Smoothing
von: Hizabri, Marouane El, et al.
Veröffentlicht: (2026)
von: Hizabri, Marouane El, et al.
Veröffentlicht: (2026)
The 2025 PNPL Competition: Speech Detection and Phoneme Classification in the LibriBrain Dataset
von: Landau, Gilad, et al.
Veröffentlicht: (2025)
von: Landau, Gilad, et al.
Veröffentlicht: (2025)
Posterior Transition Modeling for Unsupervised Diffusion-Based Speech Enhancement
von: Sadeghi, Mostafa, et al.
Veröffentlicht: (2025)
von: Sadeghi, Mostafa, et al.
Veröffentlicht: (2025)
Comparative Analysis of Mel-Frequency Cepstral Coefficients and Wavelet Based Audio Signal Processing for Emotion Detection and Mental Health Assessment in Spoken Speech
von: Agbo, Idoko, et al.
Veröffentlicht: (2024)
von: Agbo, Idoko, et al.
Veröffentlicht: (2024)
Speech to Speech Synthesis for Voice Impersonation
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
Improving Speaker-independent Speech Emotion Recognition Using Dynamic Joint Distribution Adaptation
von: Lu, Cheng, et al.
Veröffentlicht: (2024)
von: Lu, Cheng, et al.
Veröffentlicht: (2024)
Diffusion-Based Speech Enhancement in Matched and Mismatched Conditions Using a Heun-Based Sampler
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
Predicting Individual Depression Symptoms from Acoustic Features During Speech
von: Rodriguez, Sebastian, et al.
Veröffentlicht: (2024)
von: Rodriguez, Sebastian, et al.
Veröffentlicht: (2024)
Diffusion Synthesizer for Efficient Multilingual Speech to Speech Translation
von: Hirschkind, Nameer, et al.
Veröffentlicht: (2024)
von: Hirschkind, Nameer, et al.
Veröffentlicht: (2024)
Analysis of Self-Supervised Speech Models on Children's Speech and Infant Vocalizations
von: Li, Jialu, et al.
Veröffentlicht: (2024)
von: Li, Jialu, et al.
Veröffentlicht: (2024)
Mouth Articulation-Based Anchoring for Improved Cross-Corpus Speech Emotion Recognition
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
TextrolSpeech: A Text Style Control Speech Corpus With Codec Language Text-to-Speech Models
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Vocal Melody Construction for Persian Lyrics Using LSTM Recurrent Neural Networks
von: Jafari, Farshad, et al.
Veröffentlicht: (2024) -
Dynamic Gated Recurrent Neural Network for Compute-efficient Speech Enhancement
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024) -
Extract and Diffuse: Latent Integration for Improved Diffusion-based Speech and Vocal Enhancement
von: Yang, Yudong, et al.
Veröffentlicht: (2024) -
Switchboard-Affect: Emotion Perception Labels from Conversational Speech
von: Romana, Amrit, et al.
Veröffentlicht: (2025) -
End-to-End Integration of Speech Separation and Voice Activity Detection for Low-Latency Diarization of Telephone Conversations
von: Morrone, Giovanni, et al.
Veröffentlicht: (2023)