Acoustic neural networks: Identifying design principles and exploring physical feasibility
Fuente:
arXiv
Saved in:
| Main Authors: | Kalthoff, Ivan, Rey, Marcel, Wittkowski, Raphael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Memristive Nanowire Network for Energy Efficient Audio Classification: Pre-Processing-Free Reservoir Computing with Reduced Latency
by: Rajesh, Akshaya, et al.
Published: (2024)
by: Rajesh, Akshaya, et al.
Published: (2024)
Mass-Spring Models for Passive Keyword Spotting: A Springtronics Approach
by: Bohte, Finn, et al.
Published: (2025)
by: Bohte, Finn, et al.
Published: (2025)
Neurobench: DCASE 2020 Acoustic Scene Classification benchmark on XyloAudio 2
by: Ke, Weijie, et al.
Published: (2024)
by: Ke, Weijie, et al.
Published: (2024)
Grammatical Structure and Grammatical Variations in Non-Metric Iranian Classical Music
by: Kanani, Maziar, et al.
Published: (2025)
by: Kanani, Maziar, et al.
Published: (2025)
Generative Voice Bursts during Phone Call
by: Ranjan, Paritosh, et al.
Published: (2025)
by: Ranjan, Paritosh, et al.
Published: (2025)
Spiking-LEAF: A Learnable Auditory front-end for Spiking Neural Networks
by: Song, Zeyang, et al.
Published: (2023)
by: Song, Zeyang, et al.
Published: (2023)
Resource-Efficient Speech Quality Prediction through Quantization Aware Training and Binary Activation Maps
by: Nilsson, Mattias, et al.
Published: (2024)
by: Nilsson, Mattias, et al.
Published: (2024)
Low-power SNN-based audio source localisation using a Hilbert Transform spike encoding scheme
by: Haghighatshoar, Saeid, et al.
Published: (2024)
by: Haghighatshoar, Saeid, et al.
Published: (2024)
DeepSpeech models show Human-like Performance and Processing of Cochlear Implant Inputs
by: Steinhardt, Cynthia R., et al.
Published: (2024)
by: Steinhardt, Cynthia R., et al.
Published: (2024)
A Novel Transfer Learning Approach for Mental Stability Classification from Voice Signal
by: Islam, Rafiul, et al.
Published: (2026)
by: Islam, Rafiul, et al.
Published: (2026)
sVAD: A Robust, Low-Power, and Light-Weight Voice Activity Detection with Spiking Neural Networks
by: Yang, Qu, et al.
Published: (2024)
by: Yang, Qu, et al.
Published: (2024)
Artificial Neural Networks Trained on Noisy Speech Exhibit the McGurk Effect
by: Grasse, Lukas, et al.
Published: (2024)
by: Grasse, Lukas, et al.
Published: (2024)
Deep Photonic Reservoir Computer for Speech Recognition
by: Picco, Enrico, et al.
Published: (2023)
by: Picco, Enrico, et al.
Published: (2023)
Ultra-low power on-chip learning of speech commands with phase-change memories
by: Miriyala, Venkata Pavan Kumar, et al.
Published: (2020)
by: Miriyala, Venkata Pavan Kumar, et al.
Published: (2020)
Ternary Spike-based Neuromorphic Signal Processing System
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
Spiketrum: An FPGA-based Implementation of a Neuromorphic Cochlea
by: Alsakkal, MHD Anas, et al.
Published: (2024)
by: Alsakkal, MHD Anas, et al.
Published: (2024)
Spoken Conversational Agents with Large Language Models
by: Yang, Chao-Han Huck, et al.
Published: (2025)
by: Yang, Chao-Han Huck, et al.
Published: (2025)
Parsing Musical Structure to Enable Meaningful Variations
by: Kanani, Maziar, et al.
Published: (2025)
by: Kanani, Maziar, et al.
Published: (2025)
How to Estimate Model Transferability of Pre-Trained Speech Models?
by: Chen, Zih-Ching, et al.
Published: (2023)
by: Chen, Zih-Ching, et al.
Published: (2023)
Parallel Stacked Aggregated Network for Voice Authentication in IoT-Enabled Smart Devices
by: Khan, Awais, et al.
Published: (2024)
by: Khan, Awais, et al.
Published: (2024)
DPSNN: Spiking Neural Network for Low-Latency Streaming Speech Enhancement
by: Sun, Tao, et al.
Published: (2024)
by: Sun, Tao, et al.
Published: (2024)
Biomimetic Frontend for Differentiable Audio Processing
by: Famularo, Ruolan Leslie, et al.
Published: (2024)
by: Famularo, Ruolan Leslie, et al.
Published: (2024)
Global-Local Convolution with Spiking Neural Networks for Energy-efficient Keyword Spotting
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
Robust online reconstruction of continuous-time signals from a lean spike train ensemble code
by: Chattopadhyay, Anik, et al.
Published: (2024)
by: Chattopadhyay, Anik, et al.
Published: (2024)
LVNS-RAVE: Diversified audio generation with RAVE and Latent Vector Novelty Search
by: Guo, Jinyue, et al.
Published: (2024)
by: Guo, Jinyue, et al.
Published: (2024)
LACTOSE: Linear Array of Conditions, TOpologies with Separated Error-backpropagation -- The Differentiable "IF" Conditional for Differentiable Digital Signal Processing
by: Clarke, Christopher Johann
Published: (2025)
by: Clarke, Christopher Johann
Published: (2025)
Spiking Music: Audio Compression with Event Based Auto-encoders
by: Lisboa, Martim, et al.
Published: (2024)
by: Lisboa, Martim, et al.
Published: (2024)
Polariton lattices as binarized neuromorphic networks
by: Sedov, Evgeny, et al.
Published: (2024)
by: Sedov, Evgeny, et al.
Published: (2024)
Deformable Audio Transformer for Audio Event Detection
by: Zhu, Wentao
Published: (2023)
by: Zhu, Wentao
Published: (2023)
Long-Form Text-to-Music Generation with Adaptive Prompts: A Case Study in Tabletop Role-Playing Games Soundtracks
by: Marra, Felipe, et al.
Published: (2024)
by: Marra, Felipe, et al.
Published: (2024)
Automatic Voice Identification after Speech Resynthesis using PPG
by: Gaudier, Thibault, et al.
Published: (2024)
by: Gaudier, Thibault, et al.
Published: (2024)
Structure of activity in multiregion recurrent neural networks
by: Clark, David G., et al.
Published: (2024)
by: Clark, David G., et al.
Published: (2024)
Connectivity structure and dynamics of nonlinear recurrent neural networks
by: Clark, David G., et al.
Published: (2024)
by: Clark, David G., et al.
Published: (2024)
Identifying the impact of local connectivity patterns on dynamics in excitatory-inhibitory networks
by: Shao, Yuxiu, et al.
Published: (2024)
by: Shao, Yuxiu, et al.
Published: (2024)
HyperSound: Generating Implicit Neural Representations of Audio Signals with Hypernetworks
by: Szatkowski, Filip, et al.
Published: (2022)
by: Szatkowski, Filip, et al.
Published: (2022)
LMUFormer: Low Complexity Yet Powerful Spiking Model With Legendre Memory Units
by: Liu, Zeyu, et al.
Published: (2024)
by: Liu, Zeyu, et al.
Published: (2024)
Pruning-induced phases in fully-connected neural networks: the eumentia, the dementia, and the amentia
by: Pan, Haining, et al.
Published: (2026)
by: Pan, Haining, et al.
Published: (2026)
Accurate Mapping of RNNs on Neuromorphic Hardware with Adaptive Spiking Neurons
by: Boeshertz, Gauthier, et al.
Published: (2024)
by: Boeshertz, Gauthier, et al.
Published: (2024)
A Comparison of Temporal Encoders for Neuromorphic Keyword Spotting with Few Neurons
by: Nilsson, Mattias, et al.
Published: (2023)
by: Nilsson, Mattias, et al.
Published: (2023)
Spiking mode-based neural networks
by: Lin, Zhanghan, et al.
Published: (2023)
by: Lin, Zhanghan, et al.
Published: (2023)
Similar Items
-
Memristive Nanowire Network for Energy Efficient Audio Classification: Pre-Processing-Free Reservoir Computing with Reduced Latency
by: Rajesh, Akshaya, et al.
Published: (2024) -
Mass-Spring Models for Passive Keyword Spotting: A Springtronics Approach
by: Bohte, Finn, et al.
Published: (2025) -
Neurobench: DCASE 2020 Acoustic Scene Classification benchmark on XyloAudio 2
by: Ke, Weijie, et al.
Published: (2024) -
Grammatical Structure and Grammatical Variations in Non-Metric Iranian Classical Music
by: Kanani, Maziar, et al.
Published: (2025) -
Generative Voice Bursts during Phone Call
by: Ranjan, Paritosh, et al.
Published: (2025)