Contrastive Augmentation: An Unsupervised Learning Approach for Keyword Spotting in Speech Technology
Fuente:
arXiv
Guardado en:
| Autores principales: | Dai, Weinan, Jiang, Yifeng, Liu, Yuanjing, Chen, Jinkun, Sun, Xin, Tao, Jinglei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
por: Xi, Yu, et al.
Publicado: (2024)
por: Xi, Yu, et al.
Publicado: (2024)
Phoneme-Level Contrastive Learning for User-Defined Keyword Spotting with Flexible Enrollment
por: Kewei, Li, et al.
Publicado: (2024)
por: Kewei, Li, et al.
Publicado: (2024)
Effective Integration of KAN for Keyword Spotting
por: Xu, Anfeng, et al.
Publicado: (2024)
por: Xu, Anfeng, et al.
Publicado: (2024)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
por: Ding, Hanyu, et al.
Publicado: (2025)
por: Ding, Hanyu, et al.
Publicado: (2025)
Multichannel Keyword Spotting for Noisy Conditions
por: Saladukha, Dzmitry, et al.
Publicado: (2025)
por: Saladukha, Dzmitry, et al.
Publicado: (2025)
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
por: Xi, Yu, et al.
Publicado: (2024)
por: Xi, Yu, et al.
Publicado: (2024)
Does Single-channel Speech Enhancement Improve Keyword Spotting Accuracy? A Case Study
por: Brueggeman, Avamarie, et al.
Publicado: (2023)
por: Brueggeman, Avamarie, et al.
Publicado: (2023)
MFA-KWS: Effective Keyword Spotting with Multi-head Frame-asynchronous Decoding
por: Xi, Yu, et al.
Publicado: (2025)
por: Xi, Yu, et al.
Publicado: (2025)
Large-scale Contrastive Language-Audio Pretraining with Feature Fusion and Keyword-to-Caption Augmentation
por: Wu, Yusong, et al.
Publicado: (2022)
por: Wu, Yusong, et al.
Publicado: (2022)
NTC-KWS: Noise-aware CTC for Robust Keyword Spotting
por: Xi, Yu, et al.
Publicado: (2024)
por: Xi, Yu, et al.
Publicado: (2024)
AdaKWS: Towards Robust Keyword Spotting with Test-Time Adaptation
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
TDT-KWS: Fast And Accurate Keyword Spotting Using Token-and-duration Transducer
por: Xi, Yu, et al.
Publicado: (2024)
por: Xi, Yu, et al.
Publicado: (2024)
Frequency & Channel Attention Network for Small Footprint Noisy Spoken Keyword Spotting
por: Lin, Yuanxi, et al.
Publicado: (2024)
por: Lin, Yuanxi, et al.
Publicado: (2024)
Masked Self-distilled Transducer-based Keyword Spotting with Semi-autoregressive Decoding
por: Xi, Yu, et al.
Publicado: (2025)
por: Xi, Yu, et al.
Publicado: (2025)
Noise-Aware Speech Separation with Contrastive Learning
por: Zhang, Zizheng, et al.
Publicado: (2023)
por: Zhang, Zizheng, et al.
Publicado: (2023)
Advances in Small-Footprint Keyword Spotting: A Comprehensive Review of Efficient Models and Algorithms
por: Garai, Soumen, et al.
Publicado: (2025)
por: Garai, Soumen, et al.
Publicado: (2025)
Disentangled Training with Adversarial Examples For Robust Small-footprint Keyword Spotting
por: Wang, Zhenyu, et al.
Publicado: (2024)
por: Wang, Zhenyu, et al.
Publicado: (2024)
Sparse Binarization for Fast Keyword Spotting
por: Svirsky, Jonathan, et al.
Publicado: (2024)
por: Svirsky, Jonathan, et al.
Publicado: (2024)
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
por: Park, Hyun Jin, et al.
Publicado: (2024)
por: Park, Hyun Jin, et al.
Publicado: (2024)
GraphemeAug: A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples
por: Zhang, Harry, et al.
Publicado: (2025)
por: Zhang, Harry, et al.
Publicado: (2025)
Effective User-defined Keyword Spotting with Dual-stage Matching, Multi-modal Enrollment, and Continual Adaptation
por: Ai, Zhiqi, et al.
Publicado: (2026)
por: Ai, Zhiqi, et al.
Publicado: (2026)
LLM-Synth4KWS: Scalable Automatic Generation and Synthesis of Confusable Data for Custom Keyword Spotting
por: Zhu, Pai, et al.
Publicado: (2025)
por: Zhu, Pai, et al.
Publicado: (2025)
Quantization-Based Score Calibration for Few-Shot Keyword Spotting with Dynamic Time Warping in Noisy Environments
por: Wilkinghoff, Kevin, et al.
Publicado: (2025)
por: Wilkinghoff, Kevin, et al.
Publicado: (2025)
Utilizing TTS Synthesized Data for Efficient Development of Keyword Spotting Model
por: Park, Hyun Jin, et al.
Publicado: (2024)
por: Park, Hyun Jin, et al.
Publicado: (2024)
Adaptive Noise Resilient Keyword Spotting Using One-Shot Learning
por: Martinez-Rau, Luciano Sebastian, et al.
Publicado: (2025)
por: Martinez-Rau, Luciano Sebastian, et al.
Publicado: (2025)
EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting
por: Buyuksolak, Oguzhan, et al.
Publicado: (2026)
por: Buyuksolak, Oguzhan, et al.
Publicado: (2026)
Enhancing Few-shot Keyword Spotting Performance through Pre-Trained Self-supervised Speech Models
por: Gok, Alican, et al.
Publicado: (2025)
por: Gok, Alican, et al.
Publicado: (2025)
Prototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification
por: Huang, Wen, et al.
Publicado: (2024)
por: Huang, Wen, et al.
Publicado: (2024)
Robust Dual-Modal Speech Keyword Spotting for XR Headsets
por: Cai, Zhuojiang, et al.
Publicado: (2024)
por: Cai, Zhuojiang, et al.
Publicado: (2024)
Muyan-TTS: A Trainable Text-to-Speech Model Optimized for Podcast Scenarios with a $50K Budget
por: Li, Xin, et al.
Publicado: (2025)
por: Li, Xin, et al.
Publicado: (2025)
Self-Learning for Personalized Keyword Spotting on Ultra-Low-Power Audio Sensors
por: Rusci, Manuele, et al.
Publicado: (2024)
por: Rusci, Manuele, et al.
Publicado: (2024)
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval
por: Sun, Haoqin, et al.
Publicado: (2025)
por: Sun, Haoqin, et al.
Publicado: (2025)
Boosting Multi-Speaker Expressive Speech Synthesis with Semi-supervised Contrastive Learning
por: Zhu, Xinfa, et al.
Publicado: (2023)
por: Zhu, Xinfa, et al.
Publicado: (2023)
MM-KWS: Multi-modal Prompts for Multilingual User-defined Keyword Spotting
por: Ai, Zhiqi, et al.
Publicado: (2024)
por: Ai, Zhiqi, et al.
Publicado: (2024)
Emotion-Coherent Speech Data Augmentation and Self-Supervised Contrastive Style Training for Enhancing Kids's Story Speech Synthesis
por: Chung, Raymond
Publicado: (2026)
por: Chung, Raymond
Publicado: (2026)
Keyword Spotting with Hyper-Matched Filters for Small Footprint Devices
por: Segal-Feldman, Yael, et al.
Publicado: (2025)
por: Segal-Feldman, Yael, et al.
Publicado: (2025)
On-Device Domain Learning for Keyword Spotting on Low-Power Extreme Edge Embedded Systems
por: Cioflan, Cristian, et al.
Publicado: (2024)
por: Cioflan, Cristian, et al.
Publicado: (2024)
ToneUnit: A Speech Discretization Approach for Tonal Language Speech Synthesis
por: Tao, Dehua, et al.
Publicado: (2024)
por: Tao, Dehua, et al.
Publicado: (2024)
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
Unsupervised Multi-channel Speech Dereverberation via Diffusion
por: Wu, Yulun, et al.
Publicado: (2025)
por: Wu, Yulun, et al.
Publicado: (2025)
Ejemplares similares
-
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
por: Xi, Yu, et al.
Publicado: (2024) -
Phoneme-Level Contrastive Learning for User-Defined Keyword Spotting with Flexible Enrollment
por: Kewei, Li, et al.
Publicado: (2024) -
Effective Integration of KAN for Keyword Spotting
por: Xu, Anfeng, et al.
Publicado: (2024) -
Keyword Mamba: Spoken Keyword Spotting with State Space Models
por: Ding, Hanyu, et al.
Publicado: (2025) -
Multichannel Keyword Spotting for Noisy Conditions
por: Saladukha, Dzmitry, et al.
Publicado: (2025)