Query-by-Example Keyword Spotting Using Spectral-Temporal Graph Attentive Pooling and Multi-Task Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Zhenyu, Kong, Shuyu, Wan, Li, Zhang, Biqiao, Huang, Yiteng, Jin, Mumin, Sun, Ming, Lei, Xin, Yang, Zhaojun |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Effective Integration of KAN for Keyword Spotting
par: Xu, Anfeng, et autres
Publié: (2024)
par: Xu, Anfeng, et autres
Publié: (2024)
Disentangled Training with Adversarial Examples For Robust Small-footprint Keyword Spotting
par: Wang, Zhenyu, et autres
Publié: (2024)
par: Wang, Zhenyu, et autres
Publié: (2024)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
par: Ding, Hanyu, et autres
Publié: (2025)
par: Ding, Hanyu, et autres
Publié: (2025)
Multichannel Keyword Spotting for Noisy Conditions
par: Saladukha, Dzmitry, et autres
Publié: (2025)
par: Saladukha, Dzmitry, et autres
Publié: (2025)
GraphemeAug: A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples
par: Zhang, Harry, et autres
Publié: (2025)
par: Zhang, Harry, et autres
Publié: (2025)
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
par: Xi, Yu, et autres
Publié: (2024)
par: Xi, Yu, et autres
Publié: (2024)
NTC-KWS: Noise-aware CTC for Robust Keyword Spotting
par: Xi, Yu, et autres
Publié: (2024)
par: Xi, Yu, et autres
Publié: (2024)
Recursive Attentive Pooling for Extracting Speaker Embeddings from Multi-Speaker Recordings
par: Horiguchi, Shota, et autres
Publié: (2024)
par: Horiguchi, Shota, et autres
Publié: (2024)
AdaKWS: Towards Robust Keyword Spotting with Test-Time Adaptation
par: Xiao, Yang, et autres
Publié: (2025)
par: Xiao, Yang, et autres
Publié: (2025)
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
par: Xi, Yu, et autres
Publié: (2024)
par: Xi, Yu, et autres
Publié: (2024)
FADI-AEC: Fast Score Based Diffusion Model Guided by Far-end Signal for Acoustic Echo Cancellation
par: Liu, Yang, et autres
Publié: (2024)
par: Liu, Yang, et autres
Publié: (2024)
MFA-KWS: Effective Keyword Spotting with Multi-head Frame-asynchronous Decoding
par: Xi, Yu, et autres
Publié: (2025)
par: Xi, Yu, et autres
Publié: (2025)
TDT-KWS: Fast And Accurate Keyword Spotting Using Token-and-duration Transducer
par: Xi, Yu, et autres
Publié: (2024)
par: Xi, Yu, et autres
Publié: (2024)
Phoneme-Level Contrastive Learning for User-Defined Keyword Spotting with Flexible Enrollment
par: Kewei, Li, et autres
Publié: (2024)
par: Kewei, Li, et autres
Publié: (2024)
Frequency & Channel Attention Network for Small Footprint Noisy Spoken Keyword Spotting
par: Lin, Yuanxi, et autres
Publié: (2024)
par: Lin, Yuanxi, et autres
Publié: (2024)
Masked Self-distilled Transducer-based Keyword Spotting with Semi-autoregressive Decoding
par: Xi, Yu, et autres
Publié: (2025)
par: Xi, Yu, et autres
Publié: (2025)
Sparse Binarization for Fast Keyword Spotting
par: Svirsky, Jonathan, et autres
Publié: (2024)
par: Svirsky, Jonathan, et autres
Publié: (2024)
Advances in Small-Footprint Keyword Spotting: A Comprehensive Review of Efficient Models and Algorithms
par: Garai, Soumen, et autres
Publié: (2025)
par: Garai, Soumen, et autres
Publié: (2025)
Does Single-channel Speech Enhancement Improve Keyword Spotting Accuracy? A Case Study
par: Brueggeman, Avamarie, et autres
Publié: (2023)
par: Brueggeman, Avamarie, et autres
Publié: (2023)
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification
par: Peng, Junyi, et autres
Publié: (2024)
par: Peng, Junyi, et autres
Publié: (2024)
Contrastive Augmentation: An Unsupervised Learning Approach for Keyword Spotting in Speech Technology
par: Dai, Weinan, et autres
Publié: (2024)
par: Dai, Weinan, et autres
Publié: (2024)
MASV: Speaker Verification with Global and Local Context Mamba
par: Liu, Yang, et autres
Publié: (2024)
par: Liu, Yang, et autres
Publié: (2024)
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
par: Park, Hyun Jin, et autres
Publié: (2024)
par: Park, Hyun Jin, et autres
Publié: (2024)
EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting
par: Buyuksolak, Oguzhan, et autres
Publié: (2026)
par: Buyuksolak, Oguzhan, et autres
Publié: (2026)
Effective User-defined Keyword Spotting with Dual-stage Matching, Multi-modal Enrollment, and Continual Adaptation
par: Ai, Zhiqi, et autres
Publié: (2026)
par: Ai, Zhiqi, et autres
Publié: (2026)
LLM-Synth4KWS: Scalable Automatic Generation and Synthesis of Confusable Data for Custom Keyword Spotting
par: Zhu, Pai, et autres
Publié: (2025)
par: Zhu, Pai, et autres
Publié: (2025)
Quantization-Based Score Calibration for Few-Shot Keyword Spotting with Dynamic Time Warping in Noisy Environments
par: Wilkinghoff, Kevin, et autres
Publié: (2025)
par: Wilkinghoff, Kevin, et autres
Publié: (2025)
Multi-Channel Differential ASR for Robust Wearer Speech Recognition on Smart Glasses
par: Yang, Yufeng, et autres
Publié: (2025)
par: Yang, Yufeng, et autres
Publié: (2025)
Attention-Based Audio Embeddings for Query-by-Example
par: Singh, Anup, et autres
Publié: (2022)
par: Singh, Anup, et autres
Publié: (2022)
CTC-aligned Audio-Text Embedding for Streaming Open-vocabulary Keyword Spotting
par: Jin, Sichen, et autres
Publié: (2024)
par: Jin, Sichen, et autres
Publié: (2024)
Utilizing TTS Synthesized Data for Efficient Development of Keyword Spotting Model
par: Park, Hyun Jin, et autres
Publié: (2024)
par: Park, Hyun Jin, et autres
Publié: (2024)
Keyword Spotting with Hyper-Matched Filters for Small Footprint Devices
par: Segal-Feldman, Yael, et autres
Publié: (2025)
par: Segal-Feldman, Yael, et autres
Publié: (2025)
Multi-Sample Dynamic Time Warping for Few-Shot Keyword Spotting
par: Wilkinghoff, Kevin, et autres
Publié: (2024)
par: Wilkinghoff, Kevin, et autres
Publié: (2024)
Adaptive Noise Resilient Keyword Spotting Using One-Shot Learning
par: Martinez-Rau, Luciano Sebastian, et autres
Publié: (2025)
par: Martinez-Rau, Luciano Sebastian, et autres
Publié: (2025)
OnDA: On-device Channel Pruning for Efficient Personalized Keyword Spotting
par: Risso, Matteo, et autres
Publié: (2026)
par: Risso, Matteo, et autres
Publié: (2026)
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
par: Nam, Hyeonuk, et autres
Publié: (2025)
par: Nam, Hyeonuk, et autres
Publié: (2025)
End-to-End User-Defined Keyword Spotting using Shifted Delta Coefficients
par: V, Kesavaraj, et autres
Publié: (2024)
par: V, Kesavaraj, et autres
Publié: (2024)
Self-Learning for Personalized Keyword Spotting on Ultra-Low-Power Audio Sensors
par: Rusci, Manuele, et autres
Publié: (2024)
par: Rusci, Manuele, et autres
Publié: (2024)
MM-KWS: Multi-modal Prompts for Multilingual User-defined Keyword Spotting
par: Ai, Zhiqi, et autres
Publié: (2024)
par: Ai, Zhiqi, et autres
Publié: (2024)
On-Device Domain Learning for Keyword Spotting on Low-Power Extreme Edge Embedded Systems
par: Cioflan, Cristian, et autres
Publié: (2024)
par: Cioflan, Cristian, et autres
Publié: (2024)
Documents similaires
-
Effective Integration of KAN for Keyword Spotting
par: Xu, Anfeng, et autres
Publié: (2024) -
Disentangled Training with Adversarial Examples For Robust Small-footprint Keyword Spotting
par: Wang, Zhenyu, et autres
Publié: (2024) -
Keyword Mamba: Spoken Keyword Spotting with State Space Models
par: Ding, Hanyu, et autres
Publié: (2025) -
Multichannel Keyword Spotting for Noisy Conditions
par: Saladukha, Dzmitry, et autres
Publié: (2025) -
GraphemeAug: A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples
par: Zhang, Harry, et autres
Publié: (2025)