Data Augmentation for Pathological Speech Enhancement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hou, Mingchi, Hermann, Enno, Kodrasi, Ina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Variational Autoencoder for Personalized Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Influence of Clean Speech Characteristics on Speech Enhancement Performance
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Generalizability of Predictive and Generative Speech Enhancement Models to Pathological Speakers
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2024)
von: Amiri, Mahdi, et al.
Veröffentlicht: (2024)
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2025)
von: Amiri, Mahdi, et al.
Veröffentlicht: (2025)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
Overview of Automatic Speech Analysis and Technologies for Neurodegenerative Disorders: Diagnosis and Assistive Applications
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2025)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2025)
Towards interpretable emotion recognition: Identifying key features with machine learning
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
CLAP-Based Automatic Word Naming Recognition in Post-Stroke Aphasia
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2026)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2026)
Graph Neural Networks for Parkinsons Disease Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
A Differentiable Alignment Framework for Sequence-to-Sequence Modeling via Optimal Transport
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
Generative Data Augmentation Challenge: Zero-Shot Speech Synthesis for Personalized Speech Enhancement
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2025)
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2025)
Weakly Supervised Phonological Features for Pathological Speech Analysis
von: Thienpondt, Jenthe, et al.
Veröffentlicht: (2025)
von: Thienpondt, Jenthe, et al.
Veröffentlicht: (2025)
A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition
von: de Groot, Dimme, et al.
Veröffentlicht: (2026)
von: de Groot, Dimme, et al.
Veröffentlicht: (2026)
Assessing the Impact of Noise and Speech Enhancement on the Intelligibility of Speech Codecs
von: Behringer, Lyonel, et al.
Veröffentlicht: (2026)
von: Behringer, Lyonel, et al.
Veröffentlicht: (2026)
Schrödinger Bridge for Generative Speech Enhancement
von: Jukić, Ante, et al.
Veröffentlicht: (2024)
von: Jukić, Ante, et al.
Veröffentlicht: (2024)
Interspeech 2025 URGENT Speech Enhancement Challenge
von: Saijo, Kohei, et al.
Veröffentlicht: (2025)
von: Saijo, Kohei, et al.
Veröffentlicht: (2025)
ProSE: Diffusion Priors for Speech Enhancement
von: Kumar, Sonal, et al.
Veröffentlicht: (2025)
von: Kumar, Sonal, et al.
Veröffentlicht: (2025)
Enhancement of Dysarthric Speech Reconstruction by Contrastive Learning
von: Fatemeh, Keshvari, et al.
Veröffentlicht: (2024)
von: Fatemeh, Keshvari, et al.
Veröffentlicht: (2024)
Latent Filling: Latent Space Data Augmentation for Zero-shot Speech Synthesis
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2023)
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2023)
On Speech Pre-emphasis as a Simple and Inexpensive Method to Boost Speech Enhancement
von: López-Espejo, Iván, et al.
Veröffentlicht: (2024)
von: López-Espejo, Iván, et al.
Veröffentlicht: (2024)
Less is More: Data Curation Matters in Scaling Speech Enhancement
von: Li, Chenda, et al.
Veröffentlicht: (2025)
von: Li, Chenda, et al.
Veröffentlicht: (2025)
SNR-Progressive Model with Harmonic Compensation for Low-SNR Speech Enhancement
von: Hou, Zhongshu, et al.
Veröffentlicht: (2024)
von: Hou, Zhongshu, et al.
Veröffentlicht: (2024)
Investigation of Speech and Noise Latent Representations in Single-channel VAE-based Speech Enhancement
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
TripleC Learning and Lightweight Speech Enhancement for Multi-Condition Target Speech Extraction
von: Huang, Ziling
Veröffentlicht: (2025)
von: Huang, Ziling
Veröffentlicht: (2025)
Rethinking Flow and Diffusion Bridge Models for Speech Enhancement
von: Wang, Dahan, et al.
Veröffentlicht: (2026)
von: Wang, Dahan, et al.
Veröffentlicht: (2026)
Speech-dependent Data Augmentation for Own Voice Reconstruction with Hearable Microphones in Noisy Environments
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2024)
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2024)
Flexible Multichannel Speech Enhancement for Noise-Robust Frontend
von: Jukić, Ante, et al.
Veröffentlicht: (2024)
von: Jukić, Ante, et al.
Veröffentlicht: (2024)
Exploring Length Generalization For Transformer-based Speech Enhancement
von: Zhang, Qiquan, et al.
Veröffentlicht: (2025)
von: Zhang, Qiquan, et al.
Veröffentlicht: (2025)
An Exploration of Length Generalization in Transformer-Based Speech Enhancement
von: Zhang, Qiquan, et al.
Veröffentlicht: (2024)
von: Zhang, Qiquan, et al.
Veröffentlicht: (2024)
Advances in Microphone Array Processing and Multichannel Speech Enhancement
von: Huang, Gongping, et al.
Veröffentlicht: (2025)
von: Huang, Gongping, et al.
Veröffentlicht: (2025)
Complex Recurrent Variational Autoencoder with Application to Speech Enhancement
von: Xie, Yuying, et al.
Veröffentlicht: (2022)
von: Xie, Yuying, et al.
Veröffentlicht: (2022)
Selective State Space Model for Monaural Speech Enhancement
von: Chen, Moran, et al.
Veröffentlicht: (2024)
von: Chen, Moran, et al.
Veröffentlicht: (2024)
LLMs and Speech: Integration vs. Combination
von: Schmitt, Robin, et al.
Veröffentlicht: (2026)
von: Schmitt, Robin, et al.
Veröffentlicht: (2026)
Too Good to Be True: A Study on Modern Automatic Speech Recognition for the Evaluation of Speech Enhancement
von: de Oliveira, Danilo, et al.
Veröffentlicht: (2026)
von: de Oliveira, Danilo, et al.
Veröffentlicht: (2026)
Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
Plugin Speech Enhancement: A Universal Speech Enhancement Framework Inspired by Dynamic Neural Network
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
Unified Diffusion Refinement for Multi-Channel Speech Enhancement and Separation
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026)
Test-Time Adaptation For Speech Enhancement Via Mask Polarization
von: Raichle, Tobias, et al.
Veröffentlicht: (2026)
von: Raichle, Tobias, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Variational Autoencoder for Personalized Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2025) -
Influence of Clean Speech Characteristics on Speech Enhancement Performance
von: Hou, Mingchi, et al.
Veröffentlicht: (2025) -
Generalizability of Predictive and Generative Speech Enhancement Models to Pathological Speakers
von: Hou, Mingchi, et al.
Veröffentlicht: (2025) -
Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2024) -
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)