DAT-CFTNet: Speech Enhancement for Cochlear Implant Recipients using Attention-based Dual-Path Recurrent Neural Network
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mamun, Nursadul, Hansen, John H. L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cochleagram-based Noise Adapted Speaker Identification System for Distorted Speech
von: Ahmed, Sabbir, et al.
Veröffentlicht: (2025)
von: Ahmed, Sabbir, et al.
Veröffentlicht: (2025)
EmoTech: A Multi-modal Speech Emotion Recognition Using Multi-source Low-level Information with Hybrid Recurrent Network
von: Avro, Shamin Bin Habib, et al.
Veröffentlicht: (2025)
von: Avro, Shamin Bin Habib, et al.
Veröffentlicht: (2025)
EmoFormer: A Text-Independent Speech Emotion Recognition using a Hybrid Transformer-CNN model
von: Hasan, Rashedul, et al.
Veröffentlicht: (2025)
von: Hasan, Rashedul, et al.
Veröffentlicht: (2025)
Brain-Informed Speech Separation for Cochlear Implants
von: Gajecki, Tom, et al.
Veröffentlicht: (2026)
von: Gajecki, Tom, et al.
Veröffentlicht: (2026)
CIS-BWE: Chaos-Informed Speech Bandwidth Extension
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
Real-time Stereo Speech Enhancement with Spatial-Cue Preservation based on Dual-Path Structure
von: Togami, Masahito, et al.
Veröffentlicht: (2024)
von: Togami, Masahito, et al.
Veröffentlicht: (2024)
Leveraging Self-Supervised Audio-Visual Pretrained Models to Improve Vocoded Speech Intelligibility in Cochlear Implant Simulation
von: Lai, Richard Lee, et al.
Veröffentlicht: (2023)
von: Lai, Richard Lee, et al.
Veröffentlicht: (2023)
Binaural Speech Enhancement Using Complex Convolutional Recurrent Networks
von: Tokala, Vikas, et al.
Veröffentlicht: (2025)
von: Tokala, Vikas, et al.
Veröffentlicht: (2025)
Complex Recurrent Variational Autoencoder with Application to Speech Enhancement
von: Xie, Yuying, et al.
Veröffentlicht: (2022)
von: Xie, Yuying, et al.
Veröffentlicht: (2022)
ZipEnhancer: Dual-Path Down-Up Sampling-based Zipformer for Monaural Speech Enhancement
von: Wang, Haoxu, et al.
Veröffentlicht: (2025)
von: Wang, Haoxu, et al.
Veröffentlicht: (2025)
Activation Steering for Accent-Neutralized Zero-Shot Text-To-Speech
von: Yang, Mu, et al.
Veröffentlicht: (2026)
von: Yang, Mu, et al.
Veröffentlicht: (2026)
AMDM-SE: Attention-based Multichannel Diffusion Model for Speech Enhancement
von: Opochinsky, Renana, et al.
Veröffentlicht: (2026)
von: Opochinsky, Renana, et al.
Veröffentlicht: (2026)
Enhancing Cochlear Implant Signal Coding with Scaled Dot-Product Attention
von: Essaid, Billel, et al.
Veröffentlicht: (2025)
von: Essaid, Billel, et al.
Veröffentlicht: (2025)
Dynamic Gated Recurrent Neural Network for Compute-efficient Speech Enhancement
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
PLDNet: PLD-Guided Lightweight Deep Network Boosted by Efficient Attention for Handheld Dual-Microphone Speech Enhancement
von: Zhou, Nan, et al.
Veröffentlicht: (2024)
von: Zhou, Nan, et al.
Veröffentlicht: (2024)
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
Plugin Speech Enhancement: A Universal Speech Enhancement Framework Inspired by Dynamic Neural Network
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
Cortical Temporal Mismatch Compensation in Bimodal Cochlear Implant Users: Selective Attention Decoding and Pupillometry Study
von: Dolhopiatenko, Hanna, et al.
Veröffentlicht: (2025)
von: Dolhopiatenko, Hanna, et al.
Veröffentlicht: (2025)
Neural Directed Speech Enhancement with Dual Microphone Array in High Noise Scenario
von: Wen, Wen, et al.
Veröffentlicht: (2024)
von: Wen, Wen, et al.
Veröffentlicht: (2024)
Distributed Asynchronous Device Speech Enhancement via Windowed Cross-Attention
von: Yang, Gene-Ping, et al.
Veröffentlicht: (2025)
von: Yang, Gene-Ping, et al.
Veröffentlicht: (2025)
From Spikes to Speech: NeuroVoc -- A Biologically Plausible Vocoder Framework for Auditory Perception and Cochlear Implant Simulation
von: de Nobel, Jacob, et al.
Veröffentlicht: (2025)
von: de Nobel, Jacob, et al.
Veröffentlicht: (2025)
Magnitude-Phase Dual-Path Speech Enhancement Network based on Self-Supervised Embedding and Perceptual Contrast Stretch Boosting
von: Mattursun, Alimjan, et al.
Veröffentlicht: (2025)
von: Mattursun, Alimjan, et al.
Veröffentlicht: (2025)
Tracking Listener Attention: Gaze-Guided Audio-Visual Speech Enhancement Framework
von: Yang, Hsiang-Cheng, et al.
Veröffentlicht: (2026)
von: Yang, Hsiang-Cheng, et al.
Veröffentlicht: (2026)
FastEnhancer: Speed-Optimized Streaming Neural Speech Enhancement
von: Ahn, Sunghwan, et al.
Veröffentlicht: (2025)
von: Ahn, Sunghwan, et al.
Veröffentlicht: (2025)
Reverse Attention for Lightweight Speech Enhancement on Edge Devices
von: Ojha, Shuubham, et al.
Veröffentlicht: (2025)
von: Ojha, Shuubham, et al.
Veröffentlicht: (2025)
Attention-Based Beamformer For Multi-Channel Speech Enhancement
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
Leveraging Spatial Cues from Cochlear Implant Microphones to Efficiently Enhance Speech Separation in Real-World Listening Scenes
von: Olalere, Feyisayo, et al.
Veröffentlicht: (2025)
von: Olalere, Feyisayo, et al.
Veröffentlicht: (2025)
Hybrid Real- And Complex-Valued Neural Network Concept For Low-Complexity Phase-Aware Speech Enhancement
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2025)
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2025)
DeepSpeech models show Human-like Performance and Processing of Cochlear Implant Inputs
von: Steinhardt, Cynthia R., et al.
Veröffentlicht: (2024)
von: Steinhardt, Cynthia R., et al.
Veröffentlicht: (2024)
A Dual-Branch Parallel Network for Speech Enhancement and Restoration
von: Yang, Da-Hee, et al.
Veröffentlicht: (2024)
von: Yang, Da-Hee, et al.
Veröffentlicht: (2024)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
Heterogeneous Space Fusion and Dual-Dimension Attention: A New Paradigm for Speech Enhancement
von: Zheng, Tao, et al.
Veröffentlicht: (2024)
von: Zheng, Tao, et al.
Veröffentlicht: (2024)
Exploring Length Generalization For Transformer-based Speech Enhancement
von: Zhang, Qiquan, et al.
Veröffentlicht: (2025)
von: Zhang, Qiquan, et al.
Veröffentlicht: (2025)
Spatial-Filter-Bank-Based Neural Method for Multichannel Speech Enhancement
von: Zheng, Tianqin, et al.
Veröffentlicht: (2025)
von: Zheng, Tianqin, et al.
Veröffentlicht: (2025)
Bridging the Modality Gap: Softly Discretizing Audio Representation for LLM-based Automatic Speech Recognition
von: Yang, Mu, et al.
Veröffentlicht: (2025)
von: Yang, Mu, et al.
Veröffentlicht: (2025)
TokenSE: a Mamba-based discrete token speech enhancement framework for cochlear implants
von: Chiang, Hsin-Tien, et al.
Veröffentlicht: (2026)
von: Chiang, Hsin-Tien, et al.
Veröffentlicht: (2026)
Dynamically Slimmable Speech Enhancement Network with Metric-Guided Training
von: Zhao, Haixin, et al.
Veröffentlicht: (2025)
von: Zhao, Haixin, et al.
Veröffentlicht: (2025)
Investigation of Speech and Noise Latent Representations in Single-channel VAE-based Speech Enhancement
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
Pruning-aware Loss Functions for STOI-Optimized Pruned Recurrent Autoencoders for the Compression of the Stimulation Patterns of Cochlear Implants at Zero Delay
von: Hinrichs, Reemt, et al.
Veröffentlicht: (2025)
von: Hinrichs, Reemt, et al.
Veröffentlicht: (2025)
MixRep: Hidden Representation Mixup for Low-Resource Speech Recognition
von: Xie, Jiamin, et al.
Veröffentlicht: (2023)
von: Xie, Jiamin, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Cochleagram-based Noise Adapted Speaker Identification System for Distorted Speech
von: Ahmed, Sabbir, et al.
Veröffentlicht: (2025) -
EmoTech: A Multi-modal Speech Emotion Recognition Using Multi-source Low-level Information with Hybrid Recurrent Network
von: Avro, Shamin Bin Habib, et al.
Veröffentlicht: (2025) -
EmoFormer: A Text-Independent Speech Emotion Recognition using a Hybrid Transformer-CNN model
von: Hasan, Rashedul, et al.
Veröffentlicht: (2025) -
Brain-Informed Speech Separation for Cochlear Implants
von: Gajecki, Tom, et al.
Veröffentlicht: (2026) -
CIS-BWE: Chaos-Informed Speech Bandwidth Extension
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)