Speech Understanding on Tiny Devices with A Learning Cache
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Benazir, Afsara, Xu, Zhiming, Lin, Felix Xiaozhu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Safeguarding Privacy in Edge Speech Understanding with Tiny Foundation Models
von: Benazir, Afsara, et al.
Veröffentlicht: (2025)
von: Benazir, Afsara, et al.
Veröffentlicht: (2025)
Turbocharge Speech Understanding with Pilot Inference
von: Wang, Rongxiang, et al.
Veröffentlicht: (2023)
von: Wang, Rongxiang, et al.
Veröffentlicht: (2023)
WhisperFlow: speech foundation models in real time
von: Wang, Rongxiang, et al.
Veröffentlicht: (2024)
von: Wang, Rongxiang, et al.
Veröffentlicht: (2024)
TF-MLPNet: Tiny Real-Time Neural Speech Separation
von: Itani, Malek, et al.
Veröffentlicht: (2025)
von: Itani, Malek, et al.
Veröffentlicht: (2025)
TinySV: Speaker Verification in TinyML with On-device Learning
von: Pavan, Massimo, et al.
Veröffentlicht: (2024)
von: Pavan, Massimo, et al.
Veröffentlicht: (2024)
Comparison of Tiny Machine Learning Techniques for Embedded Acoustic Emission Analysis
von: Muthumala, Uditha, et al.
Veröffentlicht: (2024)
von: Muthumala, Uditha, et al.
Veröffentlicht: (2024)
Understanding Self-Supervised Learning of Speech Representation via Invariance and Redundancy Reduction
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
Universal Semantic Disentangled Privacy-preserving Speech Representation Learning
von: Vecino, Biel Tura, et al.
Veröffentlicht: (2025)
von: Vecino, Biel Tura, et al.
Veröffentlicht: (2025)
A Closer Look at Wav2Vec2 Embeddings for On-Device Single-Channel Speech Enhancement
von: Shankar, Ravi, et al.
Veröffentlicht: (2024)
von: Shankar, Ravi, et al.
Veröffentlicht: (2024)
SpeechOp: Inference-Time Task Composition for Generative Speech Processing
von: Lovelace, Justin, et al.
Veröffentlicht: (2025)
von: Lovelace, Justin, et al.
Veröffentlicht: (2025)
Learning Disentangled Speech Representations
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
SimulTron: On-Device Simultaneous Speech to Speech Translation
von: Agranovich, Alex, et al.
Veröffentlicht: (2024)
von: Agranovich, Alex, et al.
Veröffentlicht: (2024)
A Multimodal Approach to Device-Directed Speech Detection with Large Language Models
von: Wagner, Dominik, et al.
Veröffentlicht: (2024)
von: Wagner, Dominik, et al.
Veröffentlicht: (2024)
Plug-in Losses for Evidential Deep Learning: A Simplified Framework for Uncertainty Estimation that Includes the Softmax Classifier
von: Hayta, Berk, et al.
Veröffentlicht: (2026)
von: Hayta, Berk, et al.
Veröffentlicht: (2026)
Comparing Self-Supervised Learning Models Pre-Trained on Human Speech and Animal Vocalizations for Bioacoustics Processing
von: Sarkar, Eklavya, et al.
Veröffentlicht: (2025)
von: Sarkar, Eklavya, et al.
Veröffentlicht: (2025)
Hold Me Tight: Stable Encoder-Decoder Design for Speech Enhancement
von: Haider, Daniel, et al.
Veröffentlicht: (2024)
von: Haider, Daniel, et al.
Veröffentlicht: (2024)
TinyChirp: Bird Song Recognition Using TinyML Models on Low-power Wireless Acoustic Sensors
von: Huang, Zhaolan, et al.
Veröffentlicht: (2024)
von: Huang, Zhaolan, et al.
Veröffentlicht: (2024)
Simultaneous or Sequential Training? How Speech Representations Cooperate in a Multi-Task Self-Supervised Learning System
von: Khorrami, Khazar, et al.
Veröffentlicht: (2023)
von: Khorrami, Khazar, et al.
Veröffentlicht: (2023)
Edge Intelligence for Wildlife Conservation: Real-Time Hornbill Call Classification Using TinyML
von: Hing, Kong Ka, et al.
Veröffentlicht: (2025)
von: Hing, Kong Ka, et al.
Veröffentlicht: (2025)
Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis
von: Lin, Weiwei, et al.
Veröffentlicht: (2025)
von: Lin, Weiwei, et al.
Veröffentlicht: (2025)
SA-SSL-MOS: Self-supervised Learning MOS Prediction with Spectral Augmentation for Generalized Multi-Rate Speech Assessment
von: Cao, Fengyuan, et al.
Veröffentlicht: (2026)
von: Cao, Fengyuan, et al.
Veröffentlicht: (2026)
Vibravox: A Dataset of French Speech Captured with Body-conduction Audio Sensors
von: Hauret, Julien, et al.
Veröffentlicht: (2024)
von: Hauret, Julien, et al.
Veröffentlicht: (2024)
A Context-Based Numerical Format Prediction for a Text-To-Speech System
von: Darwesh, Yaser, et al.
Veröffentlicht: (2024)
von: Darwesh, Yaser, et al.
Veröffentlicht: (2024)
AV-CrossNet: an Audiovisual Complex Spectral Mapping Network for Speech Separation By Leveraging Narrow- and Cross-Band Modeling
von: Kalkhorani, Vahid Ahmadi, et al.
Veröffentlicht: (2024)
von: Kalkhorani, Vahid Ahmadi, et al.
Veröffentlicht: (2024)
Discrete-Time Diffusion-Like Models for Speech Synthesis
von: Tan, Xiaozhou, et al.
Veröffentlicht: (2025)
von: Tan, Xiaozhou, et al.
Veröffentlicht: (2025)
Principled Coarse-Grained Acceptance for Speculative Decoding in Speech
von: Yanuka, Moran, et al.
Veröffentlicht: (2025)
von: Yanuka, Moran, et al.
Veröffentlicht: (2025)
Predictive-Generative Drift Decomposition for Speech Enhancement and Separation
von: Richter, Julius, et al.
Veröffentlicht: (2026)
von: Richter, Julius, et al.
Veröffentlicht: (2026)
RepCodec: A Speech Representation Codec for Speech Tokenization
von: Huang, Zhichao, et al.
Veröffentlicht: (2023)
von: Huang, Zhichao, et al.
Veröffentlicht: (2023)
TextrolSpeech: A Text Style Control Speech Corpus With Codec Language Text-to-Speech Models
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
Multi-blank Transducers for Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching
von: Yang, Jinhyeok, et al.
Veröffentlicht: (2026)
von: Yang, Jinhyeok, et al.
Veröffentlicht: (2026)
SPIRIT: Patching Speech Language Models against Jailbreak Attacks
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
MaskCycleGAN-based Whisper to Normal Speech Conversion
von: Gupta, K. Rohith, et al.
Veröffentlicht: (2024)
von: Gupta, K. Rohith, et al.
Veröffentlicht: (2024)
Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Speech Processing
von: Fu, Yonggan, et al.
Veröffentlicht: (2022)
von: Fu, Yonggan, et al.
Veröffentlicht: (2022)
Speech to Speech Synthesis for Voice Impersonation
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
Multiple Choice Learning for Efficient Speech Separation with Many Speakers
von: Perera, David, et al.
Veröffentlicht: (2024)
von: Perera, David, et al.
Veröffentlicht: (2024)
Gradient Norm-based Fine-Tuning for Backdoor Defense in Automatic Speech Recognition
von: Zhou, Nanjun, et al.
Veröffentlicht: (2025)
von: Zhou, Nanjun, et al.
Veröffentlicht: (2025)
Who Said What? An Automated Approach to Analyzing Speech in Preschool Classrooms
von: Sun, Anchen, et al.
Veröffentlicht: (2024)
von: Sun, Anchen, et al.
Veröffentlicht: (2024)
EmoSLLM: Parameter-Efficient Adaptation of LLMs for Speech Emotion Recognition
von: Thimonier, Hugo, et al.
Veröffentlicht: (2025)
von: Thimonier, Hugo, et al.
Veröffentlicht: (2025)
Benchmarking Automatic Speech Recognition coupled LLM Modules for Medical Diagnostics
von: Kumar, Kabir
Veröffentlicht: (2025)
von: Kumar, Kabir
Veröffentlicht: (2025)
Ähnliche Einträge
-
Safeguarding Privacy in Edge Speech Understanding with Tiny Foundation Models
von: Benazir, Afsara, et al.
Veröffentlicht: (2025) -
Turbocharge Speech Understanding with Pilot Inference
von: Wang, Rongxiang, et al.
Veröffentlicht: (2023) -
WhisperFlow: speech foundation models in real time
von: Wang, Rongxiang, et al.
Veröffentlicht: (2024) -
TF-MLPNet: Tiny Real-Time Neural Speech Separation
von: Itani, Malek, et al.
Veröffentlicht: (2025) -
TinySV: Speaker Verification in TinyML with On-device Learning
von: Pavan, Massimo, et al.
Veröffentlicht: (2024)