Boosting keyword spotting through on-device learnable user speech characteristics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cioflan, Cristian, Cavigelli, Lukas, Benini, Luca |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On-Device Domain Learning for Keyword Spotting on Low-Power Extreme Edge Embedded Systems
von: Cioflan, Cristian, et al.
Veröffentlicht: (2024)
von: Cioflan, Cristian, et al.
Veröffentlicht: (2024)
Single-channel speech enhancement using learnable loss mixup
von: Chang, Oscar, et al.
Veröffentlicht: (2023)
von: Chang, Oscar, et al.
Veröffentlicht: (2023)
Open vocabulary keyword spotting through transfer learning from speech synthesis
von: V, Kesavaraj, et al.
Veröffentlicht: (2024)
von: V, Kesavaraj, et al.
Veröffentlicht: (2024)
Adaptive ship-radiated noise recognition with learnable fine-grained wavelet transform
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
Toward noise-robust whisper keyword spotting on headphones with in-earcup microphone and curriculum learning
von: Yang, Qiaoyu
Veröffentlicht: (2025)
von: Yang, Qiaoyu
Veröffentlicht: (2025)
Hardware-accelerated graph neural networks: an alternative approach for neuromorphic event-based audio classification and keyword spotting on SoC FPGA
von: Jeziorek, Kamil, et al.
Veröffentlicht: (2026)
von: Jeziorek, Kamil, et al.
Veröffentlicht: (2026)
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
von: Kim, Yunsik, et al.
Veröffentlicht: (2025)
von: Kim, Yunsik, et al.
Veröffentlicht: (2025)
Improving vision-inspired keyword spotting using dynamic module skipping in streaming conformer encoder
von: Bittar, Alexandre, et al.
Veröffentlicht: (2023)
von: Bittar, Alexandre, et al.
Veröffentlicht: (2023)
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
Selfsupervised learning for pathological speech detection
von: Sheikh, Shakeel Ahmad
Veröffentlicht: (2024)
von: Sheikh, Shakeel Ahmad
Veröffentlicht: (2024)
Towards the Synthesis of Non-speech Vocalizations
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
Multitaper mel-spectrograms for keyword spotting
von: de Souza, Douglas Baptista, et al.
Veröffentlicht: (2024)
von: de Souza, Douglas Baptista, et al.
Veröffentlicht: (2024)
CR-CTC: Consistency regularization on CTC for improved speech recognition
von: Yao, Zengwei, et al.
Veröffentlicht: (2024)
von: Yao, Zengwei, et al.
Veröffentlicht: (2024)
Zipformer: A faster and better encoder for automatic speech recognition
von: Yao, Zengwei, et al.
Veröffentlicht: (2023)
von: Yao, Zengwei, et al.
Veröffentlicht: (2023)
Robustifying automatic speech recognition by extracting slowly varying features
von: Pizarro, Matías, et al.
Veröffentlicht: (2021)
von: Pizarro, Matías, et al.
Veröffentlicht: (2021)
Late fusion ensembles for speech recognition on diverse input audio representations
von: Jezidžić, Marin, et al.
Veröffentlicht: (2024)
von: Jezidžić, Marin, et al.
Veröffentlicht: (2024)
Generalizable speech deepfake detection via meta-learned LoRA
von: Laakkonen, Janne, et al.
Veröffentlicht: (2025)
von: Laakkonen, Janne, et al.
Veröffentlicht: (2025)
Acoustic characterization of speech rhythm: going beyond metrics with recurrent neural networks
von: Deloche, François, et al.
Veröffentlicht: (2024)
von: Deloche, François, et al.
Veröffentlicht: (2024)
Context-aware child-directed speech detection from long-form recordings
von: Charlot, Théo, et al.
Veröffentlicht: (2026)
von: Charlot, Théo, et al.
Veröffentlicht: (2026)
Dementia classification from spontaneous speech using wrapper-based feature selection
von: Niemelä, Marko, et al.
Veröffentlicht: (2025)
von: Niemelä, Marko, et al.
Veröffentlicht: (2025)
On-device Streaming Discrete Speech Units
von: Choi, Kwanghee, et al.
Veröffentlicht: (2025)
von: Choi, Kwanghee, et al.
Veröffentlicht: (2025)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
An Attention Long Short-Term Memory based system for automatic classification of speech intelligibility
von: Fernández-Díaz, Miguel, et al.
Veröffentlicht: (2024)
von: Fernández-Díaz, Miguel, et al.
Veröffentlicht: (2024)
SeMaScore : a new evaluation metric for automatic speech recognition tasks
von: Sasindran, Zitha, et al.
Veröffentlicht: (2024)
von: Sasindran, Zitha, et al.
Veröffentlicht: (2024)
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2023)
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2023)
CognoSpeak: an automatic, remote assessment of early cognitive decline in real-world conversational speech
von: Pahar, Madhurananda, et al.
Veröffentlicht: (2025)
von: Pahar, Madhurananda, et al.
Veröffentlicht: (2025)
Towards objective and interpretable speech disorder assessment: a comparative analysis of CNN and transformer-based models
von: Maisonneuve, Malo, et al.
Veröffentlicht: (2024)
von: Maisonneuve, Malo, et al.
Veröffentlicht: (2024)
Objective and subjective evaluation of speech enhancement methods in the UDASE task of the 7th CHiME challenge
von: Leglaive, Simon, et al.
Veröffentlicht: (2024)
von: Leglaive, Simon, et al.
Veröffentlicht: (2024)
TinySV: Speaker Verification in TinyML with On-device Learning
von: Pavan, Massimo, et al.
Veröffentlicht: (2024)
von: Pavan, Massimo, et al.
Veröffentlicht: (2024)
The taste of IPA: Towards open-vocabulary keyword spotting and forced alignment in any language
von: Zhu, Jian, et al.
Veröffentlicht: (2023)
von: Zhu, Jian, et al.
Veröffentlicht: (2023)
OnDA: On-device Channel Pruning for Efficient Personalized Keyword Spotting
von: Risso, Matteo, et al.
Veröffentlicht: (2026)
von: Risso, Matteo, et al.
Veröffentlicht: (2026)
U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation
von: Dang, Shaoxiang, et al.
Veröffentlicht: (2024)
von: Dang, Shaoxiang, et al.
Veröffentlicht: (2024)
Online speaker diarization of meetings guided by speech separation
von: Gruttadauria, Elio, et al.
Veröffentlicht: (2024)
von: Gruttadauria, Elio, et al.
Veröffentlicht: (2024)
A multimodal dynamical variational autoencoder for audiovisual speech representation learning
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
A vector quantized masked autoencoder for audiovisual speech emotion recognition
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
Introduction to speech recognition
von: Dauphin, Gabriel
Veröffentlicht: (2024)
von: Dauphin, Gabriel
Veröffentlicht: (2024)
Generalization in birdsong classification: impact of transfer learning methods and dataset characteristics
von: Ghani, Burooj, et al.
Veröffentlicht: (2024)
von: Ghani, Burooj, et al.
Veröffentlicht: (2024)
Quantifying the Corpus Bias Problem in Automatic Music Transcription Systems
von: Marták, Lukáš Samuel, et al.
Veröffentlicht: (2024)
von: Marták, Lukáš Samuel, et al.
Veröffentlicht: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
EuleroDec: A Complex-Valued RVQ-VAE for Efficient and Robust Audio Coding
von: Cerovaz, Luca, et al.
Veröffentlicht: (2026)
von: Cerovaz, Luca, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
On-Device Domain Learning for Keyword Spotting on Low-Power Extreme Edge Embedded Systems
von: Cioflan, Cristian, et al.
Veröffentlicht: (2024) -
Single-channel speech enhancement using learnable loss mixup
von: Chang, Oscar, et al.
Veröffentlicht: (2023) -
Open vocabulary keyword spotting through transfer learning from speech synthesis
von: V, Kesavaraj, et al.
Veröffentlicht: (2024) -
Adaptive ship-radiated noise recognition with learnable fine-grained wavelet transform
von: Xie, Yuan, et al.
Veröffentlicht: (2023) -
Toward noise-robust whisper keyword spotting on headphones with in-earcup microphone and curriculum learning
von: Yang, Qiaoyu
Veröffentlicht: (2025)