Gespeichert in:
| Hauptverfasser: | Adeli, Behtom, Mclinden, John, Pandey, Pankaj, Shao, Ming, Shahriari, Yalda |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2506.00039 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeepASA: An Object-Oriented Multi-Purpose Network for Auditory Scene Analysis
von: Lee, Dongheon, et al.
Veröffentlicht: (2025)
von: Lee, Dongheon, et al.
Veröffentlicht: (2025)
Speech Denoising with Auditory Models
von: Saddler, Mark R., et al.
Veröffentlicht: (2020)
von: Saddler, Mark R., et al.
Veröffentlicht: (2020)
Low-Rank Adaptation of Deep Prior Neural Networks For Room Impulse Response Reconstruction
von: Pezzoli, Mirco, et al.
Veröffentlicht: (2025)
von: Pezzoli, Mirco, et al.
Veröffentlicht: (2025)
3D Room Geometry Inference from Multichannel Room Impulse Response using Deep Neural Network
von: Yeon, Inmo, et al.
Veröffentlicht: (2024)
von: Yeon, Inmo, et al.
Veröffentlicht: (2024)
Detecting gamma-band responses to the speech envelope for the ICASSP 2024 Auditory EEG Decoding Signal Processing Grand Challenge
von: Thornton, Mike, et al.
Veröffentlicht: (2024)
von: Thornton, Mike, et al.
Veröffentlicht: (2024)
Auditory Representation Effective for Estimating Vocal Tract Information
von: Irino, Toshio, et al.
Veröffentlicht: (2023)
von: Irino, Toshio, et al.
Veröffentlicht: (2023)
ListenNet: A Lightweight Spatio-Temporal Enhancement Nested Network for Auditory Attention Detection
von: Fan, Cunhang, et al.
Veröffentlicht: (2025)
von: Fan, Cunhang, et al.
Veröffentlicht: (2025)
Decoupled Spatial and Temporal Processing for Resource Efficient Multichannel Speech Enhancement
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
Spiking-LEAF: A Learnable Auditory front-end for Spiking Neural Networks
von: Song, Zeyang, et al.
Veröffentlicht: (2023)
von: Song, Zeyang, et al.
Veröffentlicht: (2023)
Audio-Visual Target Speaker Extraction with Reverse Selective Auditory Attention
von: Tao, Ruijie, et al.
Veröffentlicht: (2024)
von: Tao, Ruijie, et al.
Veröffentlicht: (2024)
StreamAAD: Decoding Spatial Auditory Attention with a Streaming Architecture
von: Qiu, Zelin, et al.
Veröffentlicht: (2024)
von: Qiu, Zelin, et al.
Veröffentlicht: (2024)
AADNet: An End-to-End Deep Learning Model for Auditory Attention Decoding
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
Training a Perceptual Model for Evaluating Auditory Similarity in Music Adversarial Attack
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
Harmonic Detection from Noisy Speech with Auditory Frame Gain for Intelligibility Enhancement
von: Queiroz, A., et al.
Veröffentlicht: (2024)
von: Queiroz, A., et al.
Veröffentlicht: (2024)
Réduire le bruit grâce à la réalité augmentée sonore -- Auditory Concealer
von: Boukhemia, Clara
Veröffentlicht: (2025)
von: Boukhemia, Clara
Veröffentlicht: (2025)
Noise-to-mask Ratio Loss for Deep Neural Network based Audio Watermarking
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
GMM-ResNet2: Ensemble of Group ResNet Networks for Synthetic Speech Detection
von: Lei, Zhenchun, et al.
Veröffentlicht: (2024)
von: Lei, Zhenchun, et al.
Veröffentlicht: (2024)
Enhanced Heart Sound Classification Using Mel Frequency Cepstral Coefficients and Comparative Analysis of Single vs. Ensemble Classifier Strategies
von: Rahmani, Amir Masoud, et al.
Veröffentlicht: (2024)
von: Rahmani, Amir Masoud, et al.
Veröffentlicht: (2024)
dCoNNear: An Artifact-Free Neural Network Architecture for Closed-loop Audio Signal Processing
von: Wen, Chuan, et al.
Veröffentlicht: (2025)
von: Wen, Chuan, et al.
Veröffentlicht: (2025)
All Neural Low-latency Directional Speech Extraction
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
Fitting Auditory Filterbanks with Multiresolution Neural Networks
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2023)
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2023)
Music Tagging with Classifier Group Chains
von: Hasumi, Takuya, et al.
Veröffentlicht: (2025)
von: Hasumi, Takuya, et al.
Veröffentlicht: (2025)
PrimeK-Net: Multi-scale Spectral Learning via Group Prime-Kernel Convolutional Neural Networks for Single Channel Speech Enhancement
von: Lin, Zizhen, et al.
Veröffentlicht: (2025)
von: Lin, Zizhen, et al.
Veröffentlicht: (2025)
Investigation of Deep Neural Network Acoustic Modelling Approaches for Low Resource Accented Mandarin Speech Recognition
von: Xie, Xurong, et al.
Veröffentlicht: (2022)
von: Xie, Xurong, et al.
Veröffentlicht: (2022)
On the Importance of Neural Wiener Filter for Resource Efficient Multichannel Speech Enhancement
von: Hsieh, Tsun-An, et al.
Veröffentlicht: (2024)
von: Hsieh, Tsun-An, et al.
Veröffentlicht: (2024)
Acoustic Volume Rendering for Neural Impulse Response Fields
von: Lan, Zitong, et al.
Veröffentlicht: (2024)
von: Lan, Zitong, et al.
Veröffentlicht: (2024)
Toward Improving fNIRS Classification: A Study on Activation Functions in Deep Neural Architectures
von: Adeli, Behtom, et al.
Veröffentlicht: (2025)
von: Adeli, Behtom, et al.
Veröffentlicht: (2025)
Auditory Intelligence: Understanding the World Through Sound
von: Nam, Hyeonuk
Veröffentlicht: (2025)
von: Nam, Hyeonuk
Veröffentlicht: (2025)
Moravec's Paradox: Towards an Auditory Turing Test
von: Noever, David, et al.
Veröffentlicht: (2025)
von: Noever, David, et al.
Veröffentlicht: (2025)
NoiseBandNet: Controllable Time-Varying Neural Synthesis of Sound Effects Using Filterbanks
von: Barahona-Ríos, Adrián, et al.
Veröffentlicht: (2023)
von: Barahona-Ríos, Adrián, et al.
Veröffentlicht: (2023)
RaD-Net: A Repairing and Denoising Network for Speech Signal Improvement
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
TBDM-Net: Bidirectional Dense Networks with Gender Information for Speech Emotion Recognition
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
Room Impulse Responses help attackers to evade Deep Fake Detection
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
MHANet: Multi-scale Hybrid Attention Network for Auditory Attention Detection
von: Li, Lu, et al.
Veröffentlicht: (2025)
von: Li, Lu, et al.
Veröffentlicht: (2025)
Vision Language Models Are Few-Shot Audio Spectrogram Classifiers
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
Bayesian Learning for Deep Neural Network Adaptation
von: Xie, Xurong, et al.
Veröffentlicht: (2020)
von: Xie, Xurong, et al.
Veröffentlicht: (2020)
A Novel Deep Learning Framework for Efficient Multichannel Acoustic Feedback Control
von: Wu, Yuan-Kuei, et al.
Veröffentlicht: (2025)
von: Wu, Yuan-Kuei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DeepASA: An Object-Oriented Multi-Purpose Network for Auditory Scene Analysis
von: Lee, Dongheon, et al.
Veröffentlicht: (2025) -
Speech Denoising with Auditory Models
von: Saddler, Mark R., et al.
Veröffentlicht: (2020) -
Low-Rank Adaptation of Deep Prior Neural Networks For Room Impulse Response Reconstruction
von: Pezzoli, Mirco, et al.
Veröffentlicht: (2025) -
3D Room Geometry Inference from Multichannel Room Impulse Response using Deep Neural Network
von: Yeon, Inmo, et al.
Veröffentlicht: (2024) -
Detecting gamma-band responses to the speech envelope for the ICASSP 2024 Auditory EEG Decoding Signal Processing Grand Challenge
von: Thornton, Mike, et al.
Veröffentlicht: (2024)