Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms
Fuente:
arXiv
Guardado en:
| Autores principales: | Premananth, Gowtham, Espy-Wilson, Carol |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Quantifying Articulatory Coordination as a Biomarker for Schizophrenia
por: Premananth, Gowtham, et al.
Publicado: (2025)
por: Premananth, Gowtham, et al.
Publicado: (2025)
A Computational Approach to Analyzing Disrupted Language in Schizophrenia: Integrating Surprisal and Coherence Measures
por: Premananth, Gowtham, et al.
Publicado: (2025)
por: Premananth, Gowtham, et al.
Publicado: (2025)
Speech-Based Prioritization for Schizophrenia Intervention
por: Premananth, Gowtham, et al.
Publicado: (2025)
por: Premananth, Gowtham, et al.
Publicado: (2025)
Speech-Based Estimation of Schizophrenia Severity Using Feature Fusion
por: Premananth, Gowtham, et al.
Publicado: (2024)
por: Premananth, Gowtham, et al.
Publicado: (2024)
Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives
por: Premananth, Gowtham, et al.
Publicado: (2025)
por: Premananth, Gowtham, et al.
Publicado: (2025)
Multimodal Biomarkers for Schizophrenia: Towards Individual Symptom Severity Estimation
por: Premananth, Gowtham, et al.
Publicado: (2025)
por: Premananth, Gowtham, et al.
Publicado: (2025)
A multi-modal approach for identifying schizophrenia using cross-modal attention
por: Premananth, Gowtham, et al.
Publicado: (2023)
por: Premananth, Gowtham, et al.
Publicado: (2023)
A Multimodal Framework for the Assessment of the Schizophrenia Spectrum
por: Premananth, Gowtham, et al.
Publicado: (2024)
por: Premananth, Gowtham, et al.
Publicado: (2024)
Reverse Attention for Lightweight Speech Enhancement on Edge Devices
por: Ojha, Shuubham, et al.
Publicado: (2025)
por: Ojha, Shuubham, et al.
Publicado: (2025)
Generic Speech Enhancement with Self-Supervised Representation Space Loss
por: Sato, Hiroshi, et al.
Publicado: (2025)
por: Sato, Hiroshi, et al.
Publicado: (2025)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
por: Tabatabaee, Saba, et al.
Publicado: (2026)
por: Tabatabaee, Saba, et al.
Publicado: (2026)
A Study on Speech Assessment with Visual Cues
por: Ahmed, Shafique, et al.
Publicado: (2025)
por: Ahmed, Shafique, et al.
Publicado: (2025)
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
por: Maghsoudi, Maryam, et al.
Publicado: (2026)
por: Maghsoudi, Maryam, et al.
Publicado: (2026)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
por: Wang, Bo, et al.
Publicado: (2024)
por: Wang, Bo, et al.
Publicado: (2024)
Improving Speech Inversion Through Self-Supervised Embeddings and Enhanced Tract Variables
por: Attia, Ahmed Adel, et al.
Publicado: (2023)
por: Attia, Ahmed Adel, et al.
Publicado: (2023)
Confidence-Based Self-Training for EMG-to-Speech: Leveraging Synthetic EMG for Robust Modeling
por: Chen, Xiaodan, et al.
Publicado: (2025)
por: Chen, Xiaodan, et al.
Publicado: (2025)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
por: Shi, Runwu, et al.
Publicado: (2024)
por: Shi, Runwu, et al.
Publicado: (2024)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
por: Qi, Tianhua, et al.
Publicado: (2026)
por: Qi, Tianhua, et al.
Publicado: (2026)
Speech Self-Supervised Representations Benchmarking: a Case for Larger Probing Heads
por: Zaiem, Salah, et al.
Publicado: (2023)
por: Zaiem, Salah, et al.
Publicado: (2023)
Semantic Communications for Speech Recognition
por: Weng, Zhenzi, et al.
Publicado: (2021)
por: Weng, Zhenzi, et al.
Publicado: (2021)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
por: Wang, Kuan-Chen, et al.
Publicado: (2024)
por: Wang, Kuan-Chen, et al.
Publicado: (2024)
RealClass: A Framework for Classroom Speech Simulation with Public Datasets and Game Engines
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
Speech dereverberation constrained on room impulse response characteristics
por: Bahrman, Louis, et al.
Publicado: (2024)
por: Bahrman, Louis, et al.
Publicado: (2024)
Toward Universal Speech Enhancement for Diverse Input Conditions
por: Zhang, Wangyou, et al.
Publicado: (2023)
por: Zhang, Wangyou, et al.
Publicado: (2023)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
por: Joysingh, S. Johanan, et al.
Publicado: (2024)
por: Joysingh, S. Johanan, et al.
Publicado: (2024)
On Improving Error Resilience of Neural End-to-End Speech Coders
por: Gupta, Kishan, et al.
Publicado: (2024)
por: Gupta, Kishan, et al.
Publicado: (2024)
Speech-Declipping Transformer with Complex Spectrogram and Learnerble Temporal Features
por: Kwon, Younghoo, et al.
Publicado: (2024)
por: Kwon, Younghoo, et al.
Publicado: (2024)
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution
por: Yu, Chin-Yun, et al.
Publicado: (2022)
por: Yu, Chin-Yun, et al.
Publicado: (2022)
Lessons Learned from the URGENT 2024 Speech Enhancement Challenge
por: Zhang, Wangyou, et al.
Publicado: (2025)
por: Zhang, Wangyou, et al.
Publicado: (2025)
Reading to Listen at the Cocktail Party: Multi-Modal Speech Separation
por: Rahimi, Akam, et al.
Publicado: (2025)
por: Rahimi, Akam, et al.
Publicado: (2025)
Contrastive Knowledge Distillation for Embedding Refinement in Personalized Speech Enhancement
por: Serre, Thomas, et al.
Publicado: (2026)
por: Serre, Thomas, et al.
Publicado: (2026)
Continuous Speech Tokens Makes LLMs Robust Multi-Modality Learners
por: Yuan, Ze, et al.
Publicado: (2024)
por: Yuan, Ze, et al.
Publicado: (2024)
FlexIO: Flexible Single- and Multi-Channel Speech Separation and Enhancement
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
Speech Watermarking with Discrete Intermediate Representations
por: Ji, Shengpeng, et al.
Publicado: (2024)
por: Ji, Shengpeng, et al.
Publicado: (2024)
Phase-Based Signal Representations for Scattering
por: Haider, Daniel, et al.
Publicado: (2022)
por: Haider, Daniel, et al.
Publicado: (2022)
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization
por: Gao, Xiaoxue, et al.
Publicado: (2024)
por: Gao, Xiaoxue, et al.
Publicado: (2024)
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
por: Wang, Ziqian, et al.
Publicado: (2025)
por: Wang, Ziqian, et al.
Publicado: (2025)
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
por: Lu, Ye-Xin, et al.
Publicado: (2024)
por: Lu, Ye-Xin, et al.
Publicado: (2024)
Ejemplares similares
-
Quantifying Articulatory Coordination as a Biomarker for Schizophrenia
por: Premananth, Gowtham, et al.
Publicado: (2025) -
A Computational Approach to Analyzing Disrupted Language in Schizophrenia: Integrating Surprisal and Coherence Measures
por: Premananth, Gowtham, et al.
Publicado: (2025) -
Speech-Based Prioritization for Schizophrenia Intervention
por: Premananth, Gowtham, et al.
Publicado: (2025) -
Speech-Based Estimation of Schizophrenia Severity Using Feature Fusion
por: Premananth, Gowtham, et al.
Publicado: (2024) -
Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives
por: Premananth, Gowtham, et al.
Publicado: (2025)