TDT-KWS: Fast And Accurate Keyword Spotting Using Token-and-duration Transducer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xi, Yu, Li, Hao, Yang, Baochen, Li, Haoyu, Xu, Hainan, Yu, Kai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
NTC-KWS: Noise-aware CTC for Robust Keyword Spotting
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
MFA-KWS: Effective Keyword Spotting with Multi-head Frame-asynchronous Decoding
von: Xi, Yu, et al.
Veröffentlicht: (2025)
von: Xi, Yu, et al.
Veröffentlicht: (2025)
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
Masked Self-distilled Transducer-based Keyword Spotting with Semi-autoregressive Decoding
von: Xi, Yu, et al.
Veröffentlicht: (2025)
von: Xi, Yu, et al.
Veröffentlicht: (2025)
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
AdaKWS: Towards Robust Keyword Spotting with Test-Time Adaptation
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Text-aware Speech Separation for Multi-talker Keyword Spotting
von: Li, Haoyu, et al.
Veröffentlicht: (2024)
von: Li, Haoyu, et al.
Veröffentlicht: (2024)
MM-KWS: Multi-modal Prompts for Multilingual User-defined Keyword Spotting
von: Ai, Zhiqi, et al.
Veröffentlicht: (2024)
von: Ai, Zhiqi, et al.
Veröffentlicht: (2024)
LLM-Synth4KWS: Scalable Automatic Generation and Synthesis of Confusable Data for Custom Keyword Spotting
von: Zhu, Pai, et al.
Veröffentlicht: (2025)
von: Zhu, Pai, et al.
Veröffentlicht: (2025)
ED-sKWS: Early-Decision Spiking Neural Networks for Rapid,and Energy-Efficient Keyword Spotting
von: Song, Zeyang, et al.
Veröffentlicht: (2024)
von: Song, Zeyang, et al.
Veröffentlicht: (2024)
AnalyticKWS: Towards Exemplar-Free Analytic Class Incremental Learning for Small-footprint Keyword Spotting
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Effective Integration of KAN for Keyword Spotting
von: Xu, Anfeng, et al.
Veröffentlicht: (2024)
von: Xu, Anfeng, et al.
Veröffentlicht: (2024)
ProKWS: Personalized Keyword Spotting via Collaborative Learning of Phonemes and Prosody
von: Pan, Jianan, et al.
Veröffentlicht: (2026)
von: Pan, Jianan, et al.
Veröffentlicht: (2026)
Phoneme-Level Contrastive Learning for User-Defined Keyword Spotting with Flexible Enrollment
von: Kewei, Li, et al.
Veröffentlicht: (2024)
von: Kewei, Li, et al.
Veröffentlicht: (2024)
PCOV-KWS: Multi-task Learning for Personalized Customizable Open Vocabulary Keyword Spotting
von: Pan, Jianan, et al.
Veröffentlicht: (2026)
von: Pan, Jianan, et al.
Veröffentlicht: (2026)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
Sparse Binarization for Fast Keyword Spotting
von: Svirsky, Jonathan, et al.
Veröffentlicht: (2024)
von: Svirsky, Jonathan, et al.
Veröffentlicht: (2024)
Multichannel Keyword Spotting for Noisy Conditions
von: Saladukha, Dzmitry, et al.
Veröffentlicht: (2025)
von: Saladukha, Dzmitry, et al.
Veröffentlicht: (2025)
Effective User-defined Keyword Spotting with Dual-stage Matching, Multi-modal Enrollment, and Continual Adaptation
von: Ai, Zhiqi, et al.
Veröffentlicht: (2026)
von: Ai, Zhiqi, et al.
Veröffentlicht: (2026)
Multi-blank Transducers for Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
Frequency & Channel Attention Network for Small Footprint Noisy Spoken Keyword Spotting
von: Lin, Yuanxi, et al.
Veröffentlicht: (2024)
von: Lin, Yuanxi, et al.
Veröffentlicht: (2024)
Neural Directed Speech Enhancement with Dual Microphone Array in High Noise Scenario
von: Wen, Wen, et al.
Veröffentlicht: (2024)
von: Wen, Wen, et al.
Veröffentlicht: (2024)
ImKWS: Test-Time Adaptation for Keyword Spotting with Class Imbalance
von: Ding, Hanyu, et al.
Veröffentlicht: (2026)
von: Ding, Hanyu, et al.
Veröffentlicht: (2026)
CUSIDE-T: Chunking, Simulating Future and Decoding for Transducer based Streaming ASR
von: Zhao, Wenbo, et al.
Veröffentlicht: (2024)
von: Zhao, Wenbo, et al.
Veröffentlicht: (2024)
Advances in Small-Footprint Keyword Spotting: A Comprehensive Review of Efficient Models and Algorithms
von: Garai, Soumen, et al.
Veröffentlicht: (2025)
von: Garai, Soumen, et al.
Veröffentlicht: (2025)
Disentangled Training with Adversarial Examples For Robust Small-footprint Keyword Spotting
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
Does Single-channel Speech Enhancement Improve Keyword Spotting Accuracy? A Case Study
von: Brueggeman, Avamarie, et al.
Veröffentlicht: (2023)
von: Brueggeman, Avamarie, et al.
Veröffentlicht: (2023)
VALL-T: Decoder-Only Generative Transducer for Robust and Decoding-Controllable Text-to-Speech
von: Du, Chenpeng, et al.
Veröffentlicht: (2024)
von: Du, Chenpeng, et al.
Veröffentlicht: (2024)
Adaptive Noise Resilient Keyword Spotting Using One-Shot Learning
von: Martinez-Rau, Luciano Sebastian, et al.
Veröffentlicht: (2025)
von: Martinez-Rau, Luciano Sebastian, et al.
Veröffentlicht: (2025)
Advanced Long-Content Speech Recognition With Factorized Neural Transducer
von: Gong, Xun, et al.
Veröffentlicht: (2024)
von: Gong, Xun, et al.
Veröffentlicht: (2024)
Quantization-Based Score Calibration for Few-Shot Keyword Spotting with Dynamic Time Warping in Noisy Environments
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2025)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2025)
EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting
von: Buyuksolak, Oguzhan, et al.
Veröffentlicht: (2026)
von: Buyuksolak, Oguzhan, et al.
Veröffentlicht: (2026)
Transducers with Pronunciation-aware Embeddings for Automatic Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2024)
von: Xu, Hainan, et al.
Veröffentlicht: (2024)
Emotion Neural Transducer for Fine-Grained Speech Emotion Recognition
von: Shen, Siyuan, et al.
Veröffentlicht: (2024)
von: Shen, Siyuan, et al.
Veröffentlicht: (2024)
WCTC-Biasing: Retraining-free Contextual Biasing ASR with Wildcard CTC-based Keyword Spotting and Inter-layer Biasing
von: Nakagome, Yu, et al.
Veröffentlicht: (2025)
von: Nakagome, Yu, et al.
Veröffentlicht: (2025)
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
von: Park, Hyun Jin, et al.
Veröffentlicht: (2024)
von: Park, Hyun Jin, et al.
Veröffentlicht: (2024)
Keyword Spotting with Hyper-Matched Filters for Small Footprint Devices
von: Segal-Feldman, Yael, et al.
Veröffentlicht: (2025)
von: Segal-Feldman, Yael, et al.
Veröffentlicht: (2025)
Utilizing TTS Synthesized Data for Efficient Development of Keyword Spotting Model
von: Park, Hyun Jin, et al.
Veröffentlicht: (2024)
von: Park, Hyun Jin, et al.
Veröffentlicht: (2024)
Multi-Sample Dynamic Time Warping for Few-Shot Keyword Spotting
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2024)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2024)
Contrastive Augmentation: An Unsupervised Learning Approach for Keyword Spotting in Speech Technology
von: Dai, Weinan, et al.
Veröffentlicht: (2024)
von: Dai, Weinan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
NTC-KWS: Noise-aware CTC for Robust Keyword Spotting
von: Xi, Yu, et al.
Veröffentlicht: (2024) -
MFA-KWS: Effective Keyword Spotting with Multi-head Frame-asynchronous Decoding
von: Xi, Yu, et al.
Veröffentlicht: (2025) -
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
von: Xi, Yu, et al.
Veröffentlicht: (2024) -
Masked Self-distilled Transducer-based Keyword Spotting with Semi-autoregressive Decoding
von: Xi, Yu, et al.
Veröffentlicht: (2025) -
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
von: Xi, Yu, et al.
Veröffentlicht: (2024)