Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Amiri, Mahdi, Kodrasi, Ina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2025)
von: Amiri, Mahdi, et al.
Veröffentlicht: (2025)
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
Variational Autoencoder for Personalized Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
CLAP-Based Automatic Word Naming Recognition in Post-Stroke Aphasia
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2026)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2026)
Towards interpretable emotion recognition: Identifying key features with machine learning
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
Data Augmentation for Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2026)
von: Hou, Mingchi, et al.
Veröffentlicht: (2026)
Generalizability of Predictive and Generative Speech Enhancement Models to Pathological Speakers
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Reducing Geographic Disparities in Automatic Speech Recognition via Elastic Weight Consolidation
von: Trinh, Viet Anh, et al.
Veröffentlicht: (2022)
von: Trinh, Viet Anh, et al.
Veröffentlicht: (2022)
Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
A Differentiable Alignment Framework for Sequence-to-Sequence Modeling via Optimal Transport
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
VC-ENHANCE: Speech Restoration with Integrated Noise Suppression and Voice Conversion
von: Byun, Kyungguen, et al.
Veröffentlicht: (2024)
von: Byun, Kyungguen, et al.
Veröffentlicht: (2024)
Influence of Clean Speech Characteristics on Speech Enhancement Performance
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Exploring Gender Disparities in Automatic Speech Recognition Technology
von: ElGhazaly, Hend, et al.
Veröffentlicht: (2025)
von: ElGhazaly, Hend, et al.
Veröffentlicht: (2025)
Overview of Automatic Speech Analysis and Technologies for Neurodegenerative Disorders: Diagnosis and Assistive Applications
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2025)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2025)
Noisy Disentanglement with Tri-stage Training for Noise-Robust Speech Recognition
von: Chen, Shuangyuan, et al.
Veröffentlicht: (2025)
von: Chen, Shuangyuan, et al.
Veröffentlicht: (2025)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection
von: Fan, Cunhang, et al.
Veröffentlicht: (2023)
von: Fan, Cunhang, et al.
Veröffentlicht: (2023)
Findings of the 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
NAST: Noise Aware Speech Tokenization for Speech Language Models
von: Messica, Shoval, et al.
Veröffentlicht: (2024)
von: Messica, Shoval, et al.
Veröffentlicht: (2024)
A Neural Speech Codec for Noise Robust Speech Coding
von: Huang, Jiayi, et al.
Veröffentlicht: (2023)
von: Huang, Jiayi, et al.
Veröffentlicht: (2023)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
When Voice Matters: Evidence of Gender Disparity in Positional Bias of SpeechLLMs
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2025)
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2025)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2024)
von: Shi, Hao, et al.
Veröffentlicht: (2024)
On the Relevance of Clinical Assessment Tasks for the Automatic Detection of Parkinson's Disease Medication State from Speech
von: Gimeno-Gómez, David, et al.
Veröffentlicht: (2025)
von: Gimeno-Gómez, David, et al.
Veröffentlicht: (2025)
Noise-Aware Speech Separation with Contrastive Learning
von: Zhang, Zizheng, et al.
Veröffentlicht: (2023)
von: Zhang, Zizheng, et al.
Veröffentlicht: (2023)
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
von: Nespoli, Francesco, et al.
Veröffentlicht: (2024)
von: Nespoli, Francesco, et al.
Veröffentlicht: (2024)
Joint Training of Speaker Embedding Extractor, Speech and Overlap Detection for Diarization
von: Pálka, Petr, et al.
Veröffentlicht: (2024)
von: Pálka, Petr, et al.
Veröffentlicht: (2024)
Streaming Decoder-Only Automatic Speech Recognition with Discrete Speech Units: A Pilot Study
von: Chen, Peikun, et al.
Veröffentlicht: (2024)
von: Chen, Peikun, et al.
Veröffentlicht: (2024)
SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics
von: Saeki, Takaaki, et al.
Veröffentlicht: (2024)
von: Saeki, Takaaki, et al.
Veröffentlicht: (2024)
Emotion-Coherent Speech Data Augmentation and Self-Supervised Contrastive Style Training for Enhancing Kids's Story Speech Synthesis
von: Chung, Raymond
Veröffentlicht: (2026)
von: Chung, Raymond
Veröffentlicht: (2026)
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2023)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2023)
Improved Remixing Process for Domain Adaptation-Based Speech Enhancement by Mitigating Data Imbalance in Signal-to-Noise Ratio
von: Li, Li, et al.
Veröffentlicht: (2024)
von: Li, Li, et al.
Veröffentlicht: (2024)
Transient Noise Removal via Diffusion-based Speech Inpainting
von: Moradi, Mordehay, et al.
Veröffentlicht: (2025)
von: Moradi, Mordehay, et al.
Veröffentlicht: (2025)
Fine-Tuning Automatic Speech Recognition for People with Parkinson's: An Effective Strategy for Enhancing Speech Technology Accessibility
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2024)
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2024)
SMRU: Split-and-Merge Recurrent-based UNet for Acoustic Echo Cancellation and Noise Suppression
von: Sun, Zhihang, et al.
Veröffentlicht: (2024)
von: Sun, Zhihang, et al.
Veröffentlicht: (2024)
Towards Automatic Assessment of Self-Supervised Speech Models using Rank
von: Aldeneh, Zakaria, et al.
Veröffentlicht: (2024)
von: Aldeneh, Zakaria, et al.
Veröffentlicht: (2024)
Latent-Level Enhancement with Flow Matching for Robust Automatic Speech Recognition
von: Yang, Da-Hee, et al.
Veröffentlicht: (2026)
von: Yang, Da-Hee, et al.
Veröffentlicht: (2026)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
von: Poncelet, Jakob, et al.
Veröffentlicht: (2025)
von: Poncelet, Jakob, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2025) -
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024) -
Variational Autoencoder for Personalized Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2025) -
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024) -
CLAP-Based Automatic Word Naming Recognition in Post-Stroke Aphasia
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2026)