xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
Fuente:
arXiv
Salvato in:
| Autori principali: | Kühne, Nikolai Lund, Østergaard, Jan, Jensen, Jesper, Tan, Zheng-Hua |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MambAttention: Mamba with Multi-Head Attention for Generalizable Single-Channel Speech Enhancement
di: Kühne, Nikolai Lund, et al.
Pubblicazione: (2025)
di: Kühne, Nikolai Lund, et al.
Pubblicazione: (2025)
Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enhancement
di: Kühne, Nikolai Lund, et al.
Pubblicazione: (2025)
di: Kühne, Nikolai Lund, et al.
Pubblicazione: (2025)
Investigating the Design Space of Diffusion Models for Speech Enhancement
di: Gonzalez, Philippe, et al.
Pubblicazione: (2023)
di: Gonzalez, Philippe, et al.
Pubblicazione: (2023)
Joint Minimum Processing Beamforming and Near-end Listening Enhancement
di: Fuglsig, Andreas J., et al.
Pubblicazione: (2023)
di: Fuglsig, Andreas J., et al.
Pubblicazione: (2023)
Diffusion-Based Speech Enhancement in Matched and Mismatched Conditions Using a Heun-Based Sampler
di: Gonzalez, Philippe, et al.
Pubblicazione: (2023)
di: Gonzalez, Philippe, et al.
Pubblicazione: (2023)
Improvement and Implementation of a Speech Emotion Recognition Model Based on Dual-Layer LSTM
di: Yang, Xiaoran, et al.
Pubblicazione: (2024)
di: Yang, Xiaoran, et al.
Pubblicazione: (2024)
Speech Emotion Recognition Using MFCC Features and LSTM-Based Deep Learning Model
di: Oluwademilade, Adelekun, et al.
Pubblicazione: (2026)
di: Oluwademilade, Adelekun, et al.
Pubblicazione: (2026)
Enhanced Speech Emotion Recognition with Efficient Channel Attention Guided Deep CNN-BiLSTM Framework
di: Kundu, Niloy Kumar, et al.
Pubblicazione: (2024)
di: Kundu, Niloy Kumar, et al.
Pubblicazione: (2024)
A Study of the Scale Invariant Signal to Distortion Ratio in Speech Separation with Noisy References
di: Jepsen, Simon Dahl, et al.
Pubblicazione: (2025)
di: Jepsen, Simon Dahl, et al.
Pubblicazione: (2025)
Deep low-latency joint speech transmission and enhancement over a gaussian channel
di: Bokaei, Mohammad, et al.
Pubblicazione: (2024)
di: Bokaei, Mohammad, et al.
Pubblicazione: (2024)
A Novel Bi-LSTM And Transformer Architecture For Generating Tabla Music
di: Mayya, Roopa, et al.
Pubblicazione: (2024)
di: Mayya, Roopa, et al.
Pubblicazione: (2024)
A Mel Spectrogram Enhancement Paradigm Based on CWT in Speech Synthesis
di: Hu, Guoqiang, et al.
Pubblicazione: (2024)
di: Hu, Guoqiang, et al.
Pubblicazione: (2024)
EffiFusion-GAN: Efficient Fusion Generative Adversarial Network for Speech Enhancement
di: Wen, Bin, et al.
Pubblicazione: (2025)
di: Wen, Bin, et al.
Pubblicazione: (2025)
Unsupervised Speech Enhancement using Data-defined Priors
di: Klement, Dominik, et al.
Pubblicazione: (2025)
di: Klement, Dominik, et al.
Pubblicazione: (2025)
An Investigation of Incorporating Mamba for Speech Enhancement
di: Chao, Rong, et al.
Pubblicazione: (2024)
di: Chao, Rong, et al.
Pubblicazione: (2024)
Detection and Forecasting of Parkinson Disease Progression from Speech Signal Features Using MultiLayer Perceptron and LSTM
di: Ali, Majid, et al.
Pubblicazione: (2024)
di: Ali, Majid, et al.
Pubblicazione: (2024)
Classification of Heart Sounds Using Multi-Branch Deep Convolutional Network and LSTM-CNN
di: Latifi, Seyed Amir, et al.
Pubblicazione: (2024)
di: Latifi, Seyed Amir, et al.
Pubblicazione: (2024)
Online Single-Channel Audio-Based Sound Speed Estimation for Robust Multi-Channel Audio Control
di: Fuglsig, Andreas Jonas, et al.
Pubblicazione: (2026)
di: Fuglsig, Andreas Jonas, et al.
Pubblicazione: (2026)
DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2025)
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2025)
Audio Mamba: Selective State Spaces for Self-Supervised Audio Representations
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
di: Wang, Junyu, et al.
Pubblicazione: (2024)
di: Wang, Junyu, et al.
Pubblicazione: (2024)
Speech Intelligibility Assessment with Uncertainty-Aware Whisper Embeddings and sLSTM
di: Zezario, Ryandhimas E., et al.
Pubblicazione: (2025)
di: Zezario, Ryandhimas E., et al.
Pubblicazione: (2025)
LLM-Guided Reinforcement Learning for Audio-Visual Speech Enhancement
di: Chen, Chih-Ning, et al.
Pubblicazione: (2026)
di: Chen, Chih-Ning, et al.
Pubblicazione: (2026)
Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
di: Wang, Xinsheng, et al.
Pubblicazione: (2025)
di: Wang, Xinsheng, et al.
Pubblicazione: (2025)
Continuous Modeling of the Denoising Process for Speech Enhancement Based on Deep Learning
di: Guo, Zilu, et al.
Pubblicazione: (2023)
di: Guo, Zilu, et al.
Pubblicazione: (2023)
AxLSTMs: learning self-supervised audio representations with xLSTMs
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
Single and Few-step Diffusion for Generative Speech Enhancement
di: Lay, Bunlong, et al.
Pubblicazione: (2023)
di: Lay, Bunlong, et al.
Pubblicazione: (2023)
Study of Lightweight Transformer Architectures for Single-Channel Speech Enhancement
di: Zhao, Haixin, et al.
Pubblicazione: (2025)
di: Zhao, Haixin, et al.
Pubblicazione: (2025)
AudioMAE++: learning better masked audio representations with SwiGLU FFNs
di: Yadav, Sarthak, et al.
Pubblicazione: (2025)
di: Yadav, Sarthak, et al.
Pubblicazione: (2025)
An overview of neural architectures for self-supervised audio representation learning from masked spectrograms
di: Yadav, Sarthak, et al.
Pubblicazione: (2025)
di: Yadav, Sarthak, et al.
Pubblicazione: (2025)
Frequency-Weighted Training Losses for Phoneme-Level DNN-based Speech Enhancement
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2025)
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2025)
CleanMel: Mel-Spectrogram Enhancement for Improving Both Speech Quality and ASR
di: Shao, Nian, et al.
Pubblicazione: (2025)
di: Shao, Nian, et al.
Pubblicazione: (2025)
A Two-Stage Hierarchical Deep Filtering Framework for Real-Time Speech Enhancement
di: Lu, Shenghui, et al.
Pubblicazione: (2025)
di: Lu, Shenghui, et al.
Pubblicazione: (2025)
A Lightweight and Real-Time Binaural Speech Enhancement Model with Spatial Cues Preservation
di: Wang, Jingyuan, et al.
Pubblicazione: (2024)
di: Wang, Jingyuan, et al.
Pubblicazione: (2024)
xLSTM-ECG: Multi-label ECG Classification via Feature Fusion with xLSTM
di: Kang, Lei, et al.
Pubblicazione: (2025)
di: Kang, Lei, et al.
Pubblicazione: (2025)
Temporal-Channel Modeling in Multi-head Self-Attention for Synthetic Speech Detection
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2024)
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2024)
Speech Enhancement Based on Drifting Models
di: Xu, Liang, et al.
Pubblicazione: (2026)
di: Xu, Liang, et al.
Pubblicazione: (2026)
Stage-Wise and Prior-Aware Neural Speech Phase Prediction
di: Liu, Fei, et al.
Pubblicazione: (2024)
di: Liu, Fei, et al.
Pubblicazione: (2024)
WaveLLDM: Design and Development of a Lightweight Latent Diffusion Model for Speech Enhancement and Restoration
di: Santoso, Kevin Putra, et al.
Pubblicazione: (2025)
di: Santoso, Kevin Putra, et al.
Pubblicazione: (2025)
Speech Boosting: Low-Latency Live Speech Enhancement for TWS Earbuds
di: Bae, Hanbin, et al.
Pubblicazione: (2024)
di: Bae, Hanbin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MambAttention: Mamba with Multi-Head Attention for Generalizable Single-Channel Speech Enhancement
di: Kühne, Nikolai Lund, et al.
Pubblicazione: (2025) -
Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enhancement
di: Kühne, Nikolai Lund, et al.
Pubblicazione: (2025) -
Investigating the Design Space of Diffusion Models for Speech Enhancement
di: Gonzalez, Philippe, et al.
Pubblicazione: (2023) -
Joint Minimum Processing Beamforming and Near-end Listening Enhancement
di: Fuglsig, Andreas J., et al.
Pubblicazione: (2023) -
Diffusion-Based Speech Enhancement in Matched and Mismatched Conditions Using a Heun-Based Sampler
di: Gonzalez, Philippe, et al.
Pubblicazione: (2023)