Ultra-low power on-chip learning of speech commands with phase-change memories
Fuente:
arXiv
Guardado en:
| Autores principales: | Miriyala, Venkata Pavan Kumar, Ishii, Masatoshi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2020
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Real-Time Piano Note Frequency Detection Using FPGA and FFT Core
por: Anik, Shafayet M., et al.
Publicado: (2025)
por: Anik, Shafayet M., et al.
Publicado: (2025)
Mass-Spring Models for Passive Keyword Spotting: A Springtronics Approach
por: Bohte, Finn, et al.
Publicado: (2025)
por: Bohte, Finn, et al.
Publicado: (2025)
Acoustic Local Positioning With Encoded Emission Beacons
por: Urena, Jesus, et al.
Publicado: (2024)
por: Urena, Jesus, et al.
Publicado: (2024)
Folding Attention: Memory and Power Optimization for On-Device Transformer-based Streaming Speech Recognition
por: Li, Yang, et al.
Publicado: (2023)
por: Li, Yang, et al.
Publicado: (2023)
Memristive Nanowire Network for Energy Efficient Audio Classification: Pre-Processing-Free Reservoir Computing with Reduced Latency
por: Rajesh, Akshaya, et al.
Publicado: (2024)
por: Rajesh, Akshaya, et al.
Publicado: (2024)
Acoustic neural networks: Identifying design principles and exploring physical feasibility
por: Kalthoff, Ivan, et al.
Publicado: (2025)
por: Kalthoff, Ivan, et al.
Publicado: (2025)
Prototype: A Keyword Spotting-Based Intelligent Audio SoC for IoT
por: Liang, Huihong, et al.
Publicado: (2025)
por: Liang, Huihong, et al.
Publicado: (2025)
DeltaKWS: A 65nm 36nJ/Decision Bio-inspired Temporal-Sparsity-Aware Digital Keyword Spotting IC with 0.6V Near-Threshold SRAM
por: Chen, Qinyu, et al.
Publicado: (2024)
por: Chen, Qinyu, et al.
Publicado: (2024)
ASAP-FE: Energy-Efficient Feature Extraction Enabling Multi-Channel Keyword Spotting on Edge Processors
por: Choi, Jongin, et al.
Publicado: (2025)
por: Choi, Jongin, et al.
Publicado: (2025)
TsetlinKWS: A 65nm 16.58uW, 0.63mm2 State-Driven Convolutional Tsetlin Machine-Based Accelerator For Keyword Spotting
por: Lin, Baizhou, et al.
Publicado: (2025)
por: Lin, Baizhou, et al.
Publicado: (2025)
DHFP-PE: Dual-Precision Hybrid Floating Point Processing Element for AI Acceleration
por: Kumar, Shubham, et al.
Publicado: (2026)
por: Kumar, Shubham, et al.
Publicado: (2026)
Low-power SNN-based audio source localisation using a Hilbert Transform spike encoding scheme
por: Haghighatshoar, Saeid, et al.
Publicado: (2024)
por: Haghighatshoar, Saeid, et al.
Publicado: (2024)
A 71.2-$μ$W Speech Recognition Accelerator with Recurrent Spiking Neural Network
por: Yang, Chih-Chyau, et al.
Publicado: (2025)
por: Yang, Chih-Chyau, et al.
Publicado: (2025)
A 14uJ/Decision Keyword Spotting Accelerator with In-SRAM-Computing and On Chip Learning for Customization
por: Chiang, Yu-Hsiang, et al.
Publicado: (2022)
por: Chiang, Yu-Hsiang, et al.
Publicado: (2022)
Bhasha-Rupantarika: Algorithm-Hardware Co-design approach for Multilingual Neural Machine Translation
por: Lokhande, Mukul, et al.
Publicado: (2025)
por: Lokhande, Mukul, et al.
Publicado: (2025)
Language model integration based on memory control for sequence to sequence speech recognition
por: Cho, Jaejin, et al.
Publicado: (2018)
por: Cho, Jaejin, et al.
Publicado: (2018)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
por: Tabatabaee, Saba, et al.
Publicado: (2026)
por: Tabatabaee, Saba, et al.
Publicado: (2026)
A Low-Power Streaming Speech Enhancement Accelerator For Edge Devices
por: Wu, Ci-Hao, et al.
Publicado: (2025)
por: Wu, Ci-Hao, et al.
Publicado: (2025)
Uncontrolled learning: co-design of neuromorphic hardware topology for neuromorphic algorithms
por: Barrows, Frank, et al.
Publicado: (2024)
por: Barrows, Frank, et al.
Publicado: (2024)
Grammatical Structure and Grammatical Variations in Non-Metric Iranian Classical Music
por: Kanani, Maziar, et al.
Publicado: (2025)
por: Kanani, Maziar, et al.
Publicado: (2025)
Generative Voice Bursts during Phone Call
por: Ranjan, Paritosh, et al.
Publicado: (2025)
por: Ranjan, Paritosh, et al.
Publicado: (2025)
Spiking-LEAF: A Learnable Auditory front-end for Spiking Neural Networks
por: Song, Zeyang, et al.
Publicado: (2023)
por: Song, Zeyang, et al.
Publicado: (2023)
Neurobench: DCASE 2020 Acoustic Scene Classification benchmark on XyloAudio 2
por: Ke, Weijie, et al.
Publicado: (2024)
por: Ke, Weijie, et al.
Publicado: (2024)
Resource-Efficient Speech Quality Prediction through Quantization Aware Training and Binary Activation Maps
por: Nilsson, Mattias, et al.
Publicado: (2024)
por: Nilsson, Mattias, et al.
Publicado: (2024)
DeepSpeech models show Human-like Performance and Processing of Cochlear Implant Inputs
por: Steinhardt, Cynthia R., et al.
Publicado: (2024)
por: Steinhardt, Cynthia R., et al.
Publicado: (2024)
A Novel Transfer Learning Approach for Mental Stability Classification from Voice Signal
por: Islam, Rafiul, et al.
Publicado: (2026)
por: Islam, Rafiul, et al.
Publicado: (2026)
sVAD: A Robust, Low-Power, and Light-Weight Voice Activity Detection with Spiking Neural Networks
por: Yang, Qu, et al.
Publicado: (2024)
por: Yang, Qu, et al.
Publicado: (2024)
Controlling quantum chaos via Parrondo strategies on noisy intermediate-scale quantum hardware
por: Rath, Aditi, et al.
Publicado: (2025)
por: Rath, Aditi, et al.
Publicado: (2025)
Monaural speech enhancement on drone via Adapter based transfer learning
por: Chen, Xingyu, et al.
Publicado: (2024)
por: Chen, Xingyu, et al.
Publicado: (2024)
Deep low-latency joint speech transmission and enhancement over a gaussian channel
por: Bokaei, Mohammad, et al.
Publicado: (2024)
por: Bokaei, Mohammad, et al.
Publicado: (2024)
A two-step approach for speech enhancement in low-SNR scenarios using cyclostationary beamforming and DNNs
por: Bologni, Giovanni, et al.
Publicado: (2026)
por: Bologni, Giovanni, et al.
Publicado: (2026)
Using RLHF to align speech enhancement approaches to mean-opinion quality scores
por: Kumar, Anurag, et al.
Publicado: (2024)
por: Kumar, Anurag, et al.
Publicado: (2024)
Thinking in cocktail party: Chain-of-Thought and reinforcement learning for target speaker automatic speech recognition
por: Zhang, Yiru, et al.
Publicado: (2025)
por: Zhang, Yiru, et al.
Publicado: (2025)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
Spoken Conversational Agents with Large Language Models
por: Yang, Chao-Han Huck, et al.
Publicado: (2025)
por: Yang, Chao-Han Huck, et al.
Publicado: (2025)
Infrastructure-free, Deep Learned Urban Noise Monitoring at $\sim$100mW
por: Yun, Jihoon, et al.
Publicado: (2022)
por: Yun, Jihoon, et al.
Publicado: (2022)
Edge Computing in Distributed Acoustic Sensing: An Application in Traffic Monitoring
por: Truong, Khanh, et al.
Publicado: (2024)
por: Truong, Khanh, et al.
Publicado: (2024)
On the relationship between speech and hearing
por: Umesh, Srinivasan, et al.
Publicado: (2024)
por: Umesh, Srinivasan, et al.
Publicado: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
In-context learning capabilities of Large Language Models to detect suicide risk among adolescents from speech transcripts
por: Roquefort, Filomene, et al.
Publicado: (2025)
por: Roquefort, Filomene, et al.
Publicado: (2025)
Ejemplares similares
-
Real-Time Piano Note Frequency Detection Using FPGA and FFT Core
por: Anik, Shafayet M., et al.
Publicado: (2025) -
Mass-Spring Models for Passive Keyword Spotting: A Springtronics Approach
por: Bohte, Finn, et al.
Publicado: (2025) -
Acoustic Local Positioning With Encoded Emission Beacons
por: Urena, Jesus, et al.
Publicado: (2024) -
Folding Attention: Memory and Power Optimization for On-Device Transformer-based Streaming Speech Recognition
por: Li, Yang, et al.
Publicado: (2023) -
Memristive Nanowire Network for Energy Efficient Audio Classification: Pre-Processing-Free Reservoir Computing with Reduced Latency
por: Rajesh, Akshaya, et al.
Publicado: (2024)