Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Amiri, Mahdi, Shahreza, Hatef Otroshi, Kodrasi, Ina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2024)
von: Amiri, Mahdi, et al.
Veröffentlicht: (2024)
Variational Autoencoder for Personalized Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
Towards interpretable emotion recognition: Identifying key features with machine learning
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
CLAP-Based Automatic Word Naming Recognition in Post-Stroke Aphasia
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2026)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2026)
Data Augmentation for Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2026)
von: Hou, Mingchi, et al.
Veröffentlicht: (2026)
Generalizability of Predictive and Generative Speech Enhancement Models to Pathological Speakers
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Exploring the Capability of Mamba in Speech Applications
von: Miyazaki, Koichi, et al.
Veröffentlicht: (2024)
von: Miyazaki, Koichi, et al.
Veröffentlicht: (2024)
A Differentiable Alignment Framework for Sequence-to-Sequence Modeling via Optimal Transport
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)
Influence of Clean Speech Characteristics on Speech Enhancement Performance
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
von: Zhuo, Le, et al.
Veröffentlicht: (2023)
von: Zhuo, Le, et al.
Veröffentlicht: (2023)
Efficient Long-Form Speech Recognition for General Speech In-Context Learning
von: Yen, Hao, et al.
Veröffentlicht: (2024)
von: Yen, Hao, et al.
Veröffentlicht: (2024)
Long-Context Speech Synthesis with Context-Aware Memory
von: Li, Zhipeng, et al.
Veröffentlicht: (2025)
von: Li, Zhipeng, et al.
Veröffentlicht: (2025)
Conditional Latent Diffusion-Based Speech Enhancement Via Dual Context Learning
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
CosyEdit: Unlocking End-to-End Speech Editing Capability from Zero-Shot Text-to-Speech Models
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
Speech-Mamba: Long-Context Speech Recognition with Selective State Spaces Models
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
Freeze and Learn: Continual Learning with Selective Freezing for Speech Deepfake Detection
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
Seeing the Context: Rich Visual Context-Aware Speech Recognition via Multimodal Reasoning
von: Tian, Wenjie, et al.
Veröffentlicht: (2026)
von: Tian, Wenjie, et al.
Veröffentlicht: (2026)
Naturalness-Aware Curriculum Learning with Dynamic Temperature for Speech Deepfake Detection
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
Parallel GPT: Harmonizing the Independence and Interdependence of Acoustic and Semantic Information for Zero-Shot Text-to-Speech
von: Xing, Jingyuan, et al.
Veröffentlicht: (2025)
von: Xing, Jingyuan, et al.
Veröffentlicht: (2025)
Distance Sampling-based Paraphraser Leveraging ChatGPT for Text Data Manipulation
von: Oh, Yoori, et al.
Veröffentlicht: (2024)
von: Oh, Yoori, et al.
Veröffentlicht: (2024)
Overview of Automatic Speech Analysis and Technologies for Neurodegenerative Disorders: Diagnosis and Assistive Applications
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2025)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2025)
Beyond the Utterance: An Empirical Study of Very Long Context Speech Recognition
von: Flynn, Robert, et al.
Veröffentlicht: (2026)
von: Flynn, Robert, et al.
Veröffentlicht: (2026)
Speech as a Biomarker for Disease Detection
von: Botelho, Catarina, et al.
Veröffentlicht: (2024)
von: Botelho, Catarina, et al.
Veröffentlicht: (2024)
Exploring Efficient Directional and Distance Cues for Regional Speech Separation
von: Jiang, Yiheng, et al.
Veröffentlicht: (2025)
von: Jiang, Yiheng, et al.
Veröffentlicht: (2025)
Context-Aware Two-Step Training Scheme for Domain Invariant Speech Separation
von: Wang, Wupeng, et al.
Veröffentlicht: (2025)
von: Wang, Wupeng, et al.
Veröffentlicht: (2025)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
Exploring Prediction Targets in Masked Pre-Training for Speech Foundation Models
von: Chen, Li-Wei, et al.
Veröffentlicht: (2024)
von: Chen, Li-Wei, et al.
Veröffentlicht: (2024)
Probing Whisper for Dysarthric Speech in Detection and Assessment
von: Yue, Zhengjun, et al.
Veröffentlicht: (2025)
von: Yue, Zhengjun, et al.
Veröffentlicht: (2025)
Two-pass Endpoint Detection for Speech Recognition
von: Raju, Anirudh, et al.
Veröffentlicht: (2024)
von: Raju, Anirudh, et al.
Veröffentlicht: (2024)
A Multilingual Framework for Dysarthria: Detection, Severity Classification, Speech-to-Text, and Clean Speech Generation
von: Raghu, Ananya, et al.
Veröffentlicht: (2025)
von: Raghu, Ananya, et al.
Veröffentlicht: (2025)
STSR: High-Fidelity Speech Super-Resolution via Spectral-Transient Context Modeling
von: Yuan, Jiajun, et al.
Veröffentlicht: (2025)
von: Yuan, Jiajun, et al.
Veröffentlicht: (2025)
WavCaps: A ChatGPT-Assisted Weakly-Labelled Audio Captioning Dataset for Audio-Language Multimodal Research
von: Mei, Xinhao, et al.
Veröffentlicht: (2023)
von: Mei, Xinhao, et al.
Veröffentlicht: (2023)
Advancing Electrolaryngeal Speech Enhancement Through Speech-Text Representation Learning
von: Ma, Ding, et al.
Veröffentlicht: (2026)
von: Ma, Ding, et al.
Veröffentlicht: (2026)
EELE: Exploring Efficient and Extensible LoRA Integration in Emotional Text-to-Speech
von: Qi, Xin, et al.
Veröffentlicht: (2024)
von: Qi, Xin, et al.
Veröffentlicht: (2024)
Speaker Anonymisation for Speech-based Suicide Risk Detection
von: Cui, Ziyun, et al.
Veröffentlicht: (2025)
von: Cui, Ziyun, et al.
Veröffentlicht: (2025)
From Sharpness to Better Generalization for Speech Deepfake Detection
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
Comparative Analysis of ASR Methods for Speech Deepfake Detection
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2024) -
Variational Autoencoder for Personalized Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2025) -
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024) -
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024) -
Towards interpretable emotion recognition: Identifying key features with machine learning
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2025)