Utilizing TTS Synthesized Data for Efficient Development of Keyword Spotting Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Hyun Jin, Agarwal, Dhruuv, Chen, Neng, Sun, Rentao, Partridge, Kurt, Chen, Justin, Zhang, Harry, Zhu, Pai, Bartel, Jacob, Kastner, Kyle, Wang, Gary, Rosenberg, Andrew, Wang, Quan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
von: Park, Hyun Jin, et al.
Veröffentlicht: (2024)
von: Park, Hyun Jin, et al.
Veröffentlicht: (2024)
Synth4Kws: Synthesized Speech for User Defined Keyword Spotting in Low Resource Environments
von: Zhu, Pai, et al.
Veröffentlicht: (2024)
von: Zhu, Pai, et al.
Veröffentlicht: (2024)
GraphemeAug: A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples
von: Zhang, Harry, et al.
Veröffentlicht: (2025)
von: Zhang, Harry, et al.
Veröffentlicht: (2025)
GE2E-KWS: Generalized End-to-End Training and Evaluation for Zero-shot Keyword Spotting
von: Zhu, Pai, et al.
Veröffentlicht: (2024)
von: Zhu, Pai, et al.
Veröffentlicht: (2024)
LLM-Synth4KWS: Scalable Automatic Generation and Synthesis of Confusable Data for Custom Keyword Spotting
von: Zhu, Pai, et al.
Veröffentlicht: (2025)
von: Zhu, Pai, et al.
Veröffentlicht: (2025)
The Universal Personalizer: Few-Shot Dysarthric Speech Recognition via Meta-Learning
von: Agarwal, Dhruuv, et al.
Veröffentlicht: (2025)
von: Agarwal, Dhruuv, et al.
Veröffentlicht: (2025)
Zero-shot Cross-lingual Voice Transfer for TTS
von: Biadsy, Fadi, et al.
Veröffentlicht: (2024)
von: Biadsy, Fadi, et al.
Veröffentlicht: (2024)
Dark Experience for Incremental Keyword Spotting
von: Peng, Tianyi, et al.
Veröffentlicht: (2024)
von: Peng, Tianyi, et al.
Veröffentlicht: (2024)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
Effective Integration of KAN for Keyword Spotting
von: Xu, Anfeng, et al.
Veröffentlicht: (2024)
von: Xu, Anfeng, et al.
Veröffentlicht: (2024)
Multichannel Keyword Spotting for Noisy Conditions
von: Saladukha, Dzmitry, et al.
Veröffentlicht: (2025)
von: Saladukha, Dzmitry, et al.
Veröffentlicht: (2025)
Maximum-Entropy Adversarial Audio Augmentation for Keyword Spotting
von: Ye, Zuzhao, et al.
Veröffentlicht: (2024)
von: Ye, Zuzhao, et al.
Veröffentlicht: (2024)
End-to-End Direction-Aware Keyword Spotting with Spatial Priors in Noisy Environments
von: Wang, Rui, et al.
Veröffentlicht: (2026)
von: Wang, Rui, et al.
Veröffentlicht: (2026)
Sparse Binarization for Fast Keyword Spotting
von: Svirsky, Jonathan, et al.
Veröffentlicht: (2024)
von: Svirsky, Jonathan, et al.
Veröffentlicht: (2024)
Text-aware Speech Separation for Multi-talker Keyword Spotting
von: Li, Haoyu, et al.
Veröffentlicht: (2024)
von: Li, Haoyu, et al.
Veröffentlicht: (2024)
Traceable TTS: Toward Watermark-Free TTS with Strong Traceability
von: Zhao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Zhao, Yuxiang, et al.
Veröffentlicht: (2025)
MM-KWS: Multi-modal Prompts for Multilingual User-defined Keyword Spotting
von: Ai, Zhiqi, et al.
Veröffentlicht: (2024)
von: Ai, Zhiqi, et al.
Veröffentlicht: (2024)
Contrastive Augmentation: An Unsupervised Learning Approach for Keyword Spotting in Speech Technology
von: Dai, Weinan, et al.
Veröffentlicht: (2024)
von: Dai, Weinan, et al.
Veröffentlicht: (2024)
ImKWS: Test-Time Adaptation for Keyword Spotting with Class Imbalance
von: Ding, Hanyu, et al.
Veröffentlicht: (2026)
von: Ding, Hanyu, et al.
Veröffentlicht: (2026)
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
NTC-KWS: Noise-aware CTC for Robust Keyword Spotting
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
MALEFA: Multi-grAnularity Learning and Effective False Alarm Suppression for Zero-shot Keyword Spotting
von: Li, Lo-Ya, et al.
Veröffentlicht: (2026)
von: Li, Lo-Ya, et al.
Veröffentlicht: (2026)
EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting
von: Buyuksolak, Oguzhan, et al.
Veröffentlicht: (2026)
von: Buyuksolak, Oguzhan, et al.
Veröffentlicht: (2026)
Disentangled Training with Adversarial Examples For Robust Small-footprint Keyword Spotting
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
Noise-Robust Keyword Spotting through Self-supervised Pretraining
von: Mørk, Jacob, et al.
Veröffentlicht: (2024)
von: Mørk, Jacob, et al.
Veröffentlicht: (2024)
AdaKWS: Towards Robust Keyword Spotting with Test-Time Adaptation
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
Indonesian-English Code-Switching Speech Synthesizer Utilizing Multilingual STEN-TTS and Bert LID
von: Handoyo, Ahmad Alfani, et al.
Veröffentlicht: (2024)
von: Handoyo, Ahmad Alfani, et al.
Veröffentlicht: (2024)
Text-Aware Adapter for Few-Shot Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
MFA-KWS: Effective Keyword Spotting with Multi-head Frame-asynchronous Decoding
von: Xi, Yu, et al.
Veröffentlicht: (2025)
von: Xi, Yu, et al.
Veröffentlicht: (2025)
TDT-KWS: Fast And Accurate Keyword Spotting Using Token-and-duration Transducer
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
Phoneme-Level Contrastive Learning for User-Defined Keyword Spotting with Flexible Enrollment
von: Kewei, Li, et al.
Veröffentlicht: (2024)
von: Kewei, Li, et al.
Veröffentlicht: (2024)
Frequency & Channel Attention Network for Small Footprint Noisy Spoken Keyword Spotting
von: Lin, Yuanxi, et al.
Veröffentlicht: (2024)
von: Lin, Yuanxi, et al.
Veröffentlicht: (2024)
Masked Self-distilled Transducer-based Keyword Spotting with Semi-autoregressive Decoding
von: Xi, Yu, et al.
Veröffentlicht: (2025)
von: Xi, Yu, et al.
Veröffentlicht: (2025)
Keyword Spotting with Hyper-Matched Filters for Small Footprint Devices
von: Segal-Feldman, Yael, et al.
Veröffentlicht: (2025)
von: Segal-Feldman, Yael, et al.
Veröffentlicht: (2025)
PatchDSU: Uncertainty Modeling for Out of Distribution Generalization in Keyword Spotting
von: Chernyak, Bronya Roni, et al.
Veröffentlicht: (2025)
von: Chernyak, Bronya Roni, et al.
Veröffentlicht: (2025)
MATE: Matryoshka Audio-Text Embeddings for Open-Vocabulary Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
Bridging the Gap between Audio and Text using Parallel-attention for User-defined Keyword Spotting
von: Kim, Youkyum, et al.
Veröffentlicht: (2024)
von: Kim, Youkyum, et al.
Veröffentlicht: (2024)
Advances in Small-Footprint Keyword Spotting: A Comprehensive Review of Efficient Models and Algorithms
von: Garai, Soumen, et al.
Veröffentlicht: (2025)
von: Garai, Soumen, et al.
Veröffentlicht: (2025)
Global-Local Convolution with Spiking Neural Networks for Energy-efficient Keyword Spotting
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
von: Park, Hyun Jin, et al.
Veröffentlicht: (2024) -
Synth4Kws: Synthesized Speech for User Defined Keyword Spotting in Low Resource Environments
von: Zhu, Pai, et al.
Veröffentlicht: (2024) -
GraphemeAug: A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples
von: Zhang, Harry, et al.
Veröffentlicht: (2025) -
GE2E-KWS: Generalized End-to-End Training and Evaluation for Zero-shot Keyword Spotting
von: Zhu, Pai, et al.
Veröffentlicht: (2024) -
LLM-Synth4KWS: Scalable Automatic Generation and Synthesis of Confusable Data for Custom Keyword Spotting
von: Zhu, Pai, et al.
Veröffentlicht: (2025)