U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dang, Shaoxiang, Matsumoto, Tetsuya, Takeuchi, Yoshinori, Kudo, Hiroaki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Omni-directional attention mechanism based on Mamba for speech separation
von: Xue, Ke, et al.
Veröffentlicht: (2026)
von: Xue, Ke, et al.
Veröffentlicht: (2026)
Developing vocal system impaired patient-aimed voice quality assessment approach using ASR representation-included multiple features
von: Dang, Shaoxiang, et al.
Veröffentlicht: (2024)
von: Dang, Shaoxiang, et al.
Veröffentlicht: (2024)
SepMamba: State-space models for speaker separation using Mamba
von: Avenstrup, Thor Højhus, et al.
Veröffentlicht: (2024)
von: Avenstrup, Thor Højhus, et al.
Veröffentlicht: (2024)
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
von: Wang, Junyu, et al.
Veröffentlicht: (2024)
von: Wang, Junyu, et al.
Veröffentlicht: (2024)
Mamba-based Segmentation Model for Speaker Diarization
von: Plaquet, Alexis, et al.
Veröffentlicht: (2024)
von: Plaquet, Alexis, et al.
Veröffentlicht: (2024)
Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enhancement
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
BiCrossMamba-ST: Speech Deepfake Detection with Bidirectional Mamba Spectro-Temporal Cross-Attention
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
Exploring the Capability of Mamba in Speech Applications
von: Miyazaki, Koichi, et al.
Veröffentlicht: (2024)
von: Miyazaki, Koichi, et al.
Veröffentlicht: (2024)
Towards robust paralinguistic assessment for real-world mobile health (mHealth) monitoring: an initial study of reverberation effects on speech
von: Dineley, Judith, et al.
Veröffentlicht: (2023)
von: Dineley, Judith, et al.
Veröffentlicht: (2023)
Moises-Light: Resource-efficient Band-split U-Net For Music Source Separation
von: Yun-Ning, et al.
Veröffentlicht: (2025)
von: Yun-Ning, et al.
Veröffentlicht: (2025)
Mamba in Speech: Towards an Alternative to Self-Attention
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
Universal Speech Enhancement with Regression and Generative Mamba
von: Chao, Rong, et al.
Veröffentlicht: (2025)
von: Chao, Rong, et al.
Veröffentlicht: (2025)
TAME: Temporal Audio-based Mamba for Enhanced Drone Trajectory Estimation and Classification
von: Xiao, Zhenyuan, et al.
Veröffentlicht: (2024)
von: Xiao, Zhenyuan, et al.
Veröffentlicht: (2024)
Mamba-based Decoder-Only Approach with Bidirectional Speech Modeling for Speech Recognition
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2024)
MASV: Speaker Verification with Global and Local Context Mamba
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Are Mamba-based Audio Foundation Models the Best Fit for Non-Verbal Emotion Recognition?
von: Akhtar, Mohd Mujtaba, et al.
Veröffentlicht: (2025)
von: Akhtar, Mohd Mujtaba, et al.
Veröffentlicht: (2025)
SPGM: Prioritizing Local Features for enhanced speech separation performance
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
DualSep: A Light-weight dual-encoder convolutional recurrent network for real-time in-car speech separation
von: Wang, Ziqian, et al.
Veröffentlicht: (2024)
von: Wang, Ziqian, et al.
Veröffentlicht: (2024)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
KS-Net: Multi-band joint speech restoration and enhancement network for 2024 ICASSP SSI Challenge
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution
von: Lee, Yongjoon, et al.
Veröffentlicht: (2024)
von: Lee, Yongjoon, et al.
Veröffentlicht: (2024)
MambaRate: Speech Quality Assessment Across Different Sampling Rates
von: Kakoulidis, Panos, et al.
Veröffentlicht: (2025)
von: Kakoulidis, Panos, et al.
Veröffentlicht: (2025)
Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement
von: Chao, Rong, et al.
Veröffentlicht: (2025)
von: Chao, Rong, et al.
Veröffentlicht: (2025)
TF-Mamba: A Time-Frequency Network for Sound Source Localization
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
SAM: A Mamba-2 State-Space Audio-Language Model
von: Lee, Taehan, et al.
Veröffentlicht: (2025)
von: Lee, Taehan, et al.
Veröffentlicht: (2025)
Multichannel blind speech source separation with a disjoint constraint source model
von: Wang, Jianyu, et al.
Veröffentlicht: (2024)
von: Wang, Jianyu, et al.
Veröffentlicht: (2024)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
MambaFoley: Foley Sound Generation using Selective State-Space Models
von: Colombo, Marco Furio, et al.
Veröffentlicht: (2024)
von: Colombo, Marco Furio, et al.
Veröffentlicht: (2024)
Speech-Mamba: Long-Context Speech Recognition with Selective State Spaces Models
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
A Two-Stage Band-Split Mamba-2 Network For Music Separation
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
Exploring synthetic data for cross-speaker style transfer in style representation based TTS
von: Ueda, Lucas H., et al.
Veröffentlicht: (2024)
von: Ueda, Lucas H., et al.
Veröffentlicht: (2024)
TTS-CtrlNet: Time varying emotion aligned text-to-speech generation with ControlNet
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
Non-Intrusive Binaural Speech Intelligibility Prediction Using Mamba for Hearing-Impaired Listeners
von: Yamamoto, Katsuhiko, et al.
Veröffentlicht: (2025)
von: Yamamoto, Katsuhiko, et al.
Veröffentlicht: (2025)
Determined blind source separation via modeling adjacent frequency band correlations in speech signals
von: Wang, Jianyu, et al.
Veröffentlicht: (2025)
von: Wang, Jianyu, et al.
Veröffentlicht: (2025)
Effects of auditory distance cues and reverberation on spatial perception and listening strategies
von: Missoni, Fulvio, et al.
Veröffentlicht: (2025)
von: Missoni, Fulvio, et al.
Veröffentlicht: (2025)
Filling MIDI Velocity using U-Net Image Colorizer
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Omni-directional attention mechanism based on Mamba for speech separation
von: Xue, Ke, et al.
Veröffentlicht: (2026) -
Developing vocal system impaired patient-aimed voice quality assessment approach using ASR representation-included multiple features
von: Dang, Shaoxiang, et al.
Veröffentlicht: (2024) -
SepMamba: State-space models for speaker separation using Mamba
von: Avenstrup, Thor Højhus, et al.
Veröffentlicht: (2024) -
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
von: Wang, Junyu, et al.
Veröffentlicht: (2024) -
Mamba-based Segmentation Model for Speaker Diarization
von: Plaquet, Alexis, et al.
Veröffentlicht: (2024)