A Two-Stage Band-Split Mamba-2 Network For Music Separation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Jinglin, Fang, Yuan, Wang, Jiajie, Zhang, Xueliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vector Quantized Diffusion Model Based Speech Bandwidth Extension
von: Fang, Yuan, et al.
Veröffentlicht: (2024)
von: Fang, Yuan, et al.
Veröffentlicht: (2024)
Attention-Based Beamformer For Multi-Channel Speech Enhancement
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
Multi-Stage Music Source Restoration with BandSplit-RoFormer Separation and HiFi++ GAN
von: Morocutti, Tobias, et al.
Veröffentlicht: (2026)
von: Morocutti, Tobias, et al.
Veröffentlicht: (2026)
Moises-Light: Resource-efficient Band-split U-Net For Music Source Separation
von: Yun-Ning, et al.
Veröffentlicht: (2025)
von: Yun-Ning, et al.
Veröffentlicht: (2025)
MusicMamba: A Dual-Feature Modeling Approach for Generating Chinese Traditional Music with Modal Precision
von: Chen, Jiatao, et al.
Veröffentlicht: (2024)
von: Chen, Jiatao, et al.
Veröffentlicht: (2024)
Improving Music Source Separation with Diffusion and Consistency Refinement
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
3S-TSE: Efficient Three-Stage Target Speaker Extraction for Real-Time and Low-Resource Applications
von: He, Shulin, et al.
Veröffentlicht: (2023)
von: He, Shulin, et al.
Veröffentlicht: (2023)
ACMID: Automatic Curation of Musical Instrument Dataset for 7-Stem Music Source Separation
von: Yu, Ji, et al.
Veröffentlicht: (2025)
von: Yu, Ji, et al.
Veröffentlicht: (2025)
CrossNet: Leveraging Global, Cross-Band, Narrow-Band, and Positional Encoding for Single- and Multi-Channel Speaker Separation
von: Kalkhorani, Vahid Ahmadi, et al.
Veröffentlicht: (2024)
von: Kalkhorani, Vahid Ahmadi, et al.
Veröffentlicht: (2024)
The Whole Is Greater than the Sum of Its Parts: Improving Music Source Separation by Bridging Network
von: Sawata, Ryosuke, et al.
Veröffentlicht: (2023)
von: Sawata, Ryosuke, et al.
Veröffentlicht: (2023)
Learning Separated Representations for Instrument-based Music Similarity
von: Hashizume, Yuka, et al.
Veröffentlicht: (2025)
von: Hashizume, Yuka, et al.
Veröffentlicht: (2025)
BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech Enhancement
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
Exploiting Music Source Separation for Automatic Lyrics Transcription with Whisper
von: Syed, Jaza, et al.
Veröffentlicht: (2025)
von: Syed, Jaza, et al.
Veröffentlicht: (2025)
Multi-Channel Acoustic Echo Cancellation Based on Direction-of-Arrival Estimation
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
Two-stage Audio-Visual Target Speaker Extraction System for Real-Time Processing On Edge Device
von: Li, Zixuan, et al.
Veröffentlicht: (2025)
von: Li, Zixuan, et al.
Veröffentlicht: (2025)
SICRN: Advancing Speech Enhancement through State Space Model and Inplace Convolution Techniques
von: Zhao, Changjiang, et al.
Veröffentlicht: (2024)
von: Zhao, Changjiang, et al.
Veröffentlicht: (2024)
Robust Target Speaker Direction of Arrival Estimation
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
Room Impulse Response as a Prompt for Acoustic Echo Cancellation
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
Simultaneous Music Separation and Generation Using Multi-Track Latent Diffusion Models
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
Musical Source Separation of Brazilian Percussion
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
A Two-Stage Framework in Cross-Spectrum Domain for Real-Time Speech Enhancement
von: Zhang, Yuewei, et al.
Veröffentlicht: (2024)
von: Zhang, Yuewei, et al.
Veröffentlicht: (2024)
Embedding-Based Intrusive Evaluation Metrics for Musical Source Separation Using MERT Representations
von: Bereuter, Paul A., et al.
Veröffentlicht: (2026)
von: Bereuter, Paul A., et al.
Veröffentlicht: (2026)
Music Similarity Representation Learning Focusing on Individual Instruments with Source Separation and Human Preference
von: Imamura, Takehiro, et al.
Veröffentlicht: (2025)
von: Imamura, Takehiro, et al.
Veröffentlicht: (2025)
High Fidelity Text-Guided Music Editing via Single-Stage Flow Matching
von: Lan, Gael Le, et al.
Veröffentlicht: (2024)
von: Lan, Gael Le, et al.
Veröffentlicht: (2024)
Music2Fail: Transfer Music to Failed Recorder Style
von: Leong, Chon In, et al.
Veröffentlicht: (2024)
von: Leong, Chon In, et al.
Veröffentlicht: (2024)
Dual-path Mamba: Short and Long-term Bidirectional Selective Structured State Space Models for Speech Separation
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
Context-Aware Two-Step Training Scheme for Domain Invariant Speech Separation
von: Wang, Wupeng, et al.
Veröffentlicht: (2025)
von: Wang, Wupeng, et al.
Veröffentlicht: (2025)
TF-Mamba: A Time-Frequency Network for Sound Source Localization
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Spatial-Temporal Activity-Informed Diarization and Separation
von: Hsu, Yicheng, et al.
Veröffentlicht: (2024)
von: Hsu, Yicheng, et al.
Veröffentlicht: (2024)
Distortion Recovery: A Two-Stage Method for Guitar Effect Removal
von: Lee, Ying-Shuo, et al.
Veröffentlicht: (2024)
von: Lee, Ying-Shuo, et al.
Veröffentlicht: (2024)
MusicEval: A Generative Music Dataset with Expert Ratings for Automatic Text-to-Music Evaluation
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
Dense-TSNet: Dense Connected Two-Stage Structure for Ultra-Lightweight Speech Enhancement
von: Lin, Zizhen, et al.
Veröffentlicht: (2024)
von: Lin, Zizhen, et al.
Veröffentlicht: (2024)
Subband Splitting: Simple, Efficient and Effective Technique for Solving Block Permutation Problem in Determined Blind Source Separation
von: Matsumoto, Kazuki, et al.
Veröffentlicht: (2024)
von: Matsumoto, Kazuki, et al.
Veröffentlicht: (2024)
Phase Repair for Time-Domain Convolutional Neural Networks in Music Super-Resolution
von: Zhang, Yenan, et al.
Veröffentlicht: (2023)
von: Zhang, Yenan, et al.
Veröffentlicht: (2023)
Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
Mamba-Diffusion Model with Learnable Wavelet for Controllable Symbolic Music Generation
von: Zhang, Jincheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jincheng, et al.
Veröffentlicht: (2025)
MAJL: A Model-Agnostic Joint Learning Framework for Music Source Separation and Pitch Estimation
von: Wei, Haojie, et al.
Veröffentlicht: (2025)
von: Wei, Haojie, et al.
Veröffentlicht: (2025)
Music Source Separation Based on a Lightweight Deep Learning Framework (DTTNET: DUAL-PATH TFC-TDF UNET)
von: Chen, Junyu, et al.
Veröffentlicht: (2023)
von: Chen, Junyu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Vector Quantized Diffusion Model Based Speech Bandwidth Extension
von: Fang, Yuan, et al.
Veröffentlicht: (2024) -
Attention-Based Beamformer For Multi-Channel Speech Enhancement
von: Bai, Jinglin, et al.
Veröffentlicht: (2024) -
Multi-Stage Music Source Restoration with BandSplit-RoFormer Separation and HiFi++ GAN
von: Morocutti, Tobias, et al.
Veröffentlicht: (2026) -
Moises-Light: Resource-efficient Band-split U-Net For Music Source Separation
von: Yun-Ning, et al.
Veröffentlicht: (2025) -
MusicMamba: A Dual-Feature Modeling Approach for Generating Chinese Traditional Music with Modal Precision
von: Chen, Jiatao, et al.
Veröffentlicht: (2024)