Sub-band and Full-band Interactive U-Net with DPRNN for Demixing Cross-talk Stereo Music
Fuente:
arXiv
Salvato in:
| Autori principali: | Yin, Han, Wang, Mou, Bai, Jisheng, Shi, Dongyuan, Gan, Woon-Seng, Chen, Jianfeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AudioLog: LLMs-Powered Long Audio Logging with Hybrid Token-Semantic Contrastive Learning
di: Bai, Jisheng, et al.
Pubblicazione: (2023)
di: Bai, Jisheng, et al.
Pubblicazione: (2023)
A Stabilized Hybrid Active Noise Control Algorithm of GFANC and FxNLMS with Online Clustering
di: Luo, Zhengding, et al.
Pubblicazione: (2026)
di: Luo, Zhengding, et al.
Pubblicazione: (2026)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Music Demixing Track
di: Fabbro, Giorgio, et al.
Pubblicazione: (2023)
di: Fabbro, Giorgio, et al.
Pubblicazione: (2023)
Sound event localization and classification using WASN in Outdoor Environment
di: Zhang, Dongzhe, et al.
Pubblicazione: (2024)
di: Zhang, Dongzhe, et al.
Pubblicazione: (2024)
Description on IEEE ICME 2024 Grand Challenge: Semi-supervised Acoustic Scene Classification under Domain Shift
di: Bai, Jisheng, et al.
Pubblicazione: (2024)
di: Bai, Jisheng, et al.
Pubblicazione: (2024)
A Real-Time Platform for Portable and Scalable Active Noise Mitigation for Construction Machinery
di: Gan, Woon-Seng, et al.
Pubblicazione: (2024)
di: Gan, Woon-Seng, et al.
Pubblicazione: (2024)
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction
di: Yeo, Jin Jie Sean, et al.
Pubblicazione: (2024)
di: Yeo, Jin Jie Sean, et al.
Pubblicazione: (2024)
Improving Speech Enhancement by Cross- and Sub-band Processing with State Space Model
di: Li, Jizhen, et al.
Pubblicazione: (2025)
di: Li, Jizhen, et al.
Pubblicazione: (2025)
AudioSetCaps: An Enriched Audio-Caption Dataset using Automated Generation Pipeline with Large Audio and Language Models
di: Bai, Jisheng, et al.
Pubblicazione: (2024)
di: Bai, Jisheng, et al.
Pubblicazione: (2024)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Cinematic Demixing Track
di: Uhlich, Stefan, et al.
Pubblicazione: (2023)
di: Uhlich, Stefan, et al.
Pubblicazione: (2023)
Directional Selective Fixed-Filter Active Noise Control Based on a Convolutional Neural Network in Reverberant Environments
di: Wang, Boxiang, et al.
Pubblicazione: (2026)
di: Wang, Boxiang, et al.
Pubblicazione: (2026)
A Survey of Integrating Wireless Technology into Active Noise Control
di: Shen, Xiaoyi, et al.
Pubblicazione: (2024)
di: Shen, Xiaoyi, et al.
Pubblicazione: (2024)
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
di: Zhao, Shengkui, et al.
Pubblicazione: (2022)
di: Zhao, Shengkui, et al.
Pubblicazione: (2022)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
di: Yeow, Jun-Wei, et al.
Pubblicazione: (2025)
di: Yeow, Jun-Wei, et al.
Pubblicazione: (2025)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
di: Xiao, Yang, et al.
Pubblicazione: (2024)
di: Xiao, Yang, et al.
Pubblicazione: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
di: Yin, Han, et al.
Pubblicazione: (2024)
di: Yin, Han, et al.
Pubblicazione: (2024)
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
di: Ooi, Kenneth, et al.
Pubblicazione: (2023)
di: Ooi, Kenneth, et al.
Pubblicazione: (2023)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
di: Yan, Haoyin, et al.
Pubblicazione: (2024)
di: Yan, Haoyin, et al.
Pubblicazione: (2024)
Extracting Urban Sound Information for Residential Areas in Smart Cities Using an End-to-End IoT System
di: Tan, Ee-Leng, et al.
Pubblicazione: (2024)
di: Tan, Ee-Leng, et al.
Pubblicazione: (2024)
Towards Ultra-Low-Power Neuromorphic Speech Enhancement with Spiking-FullSubNet
di: Hao, Xiang, et al.
Pubblicazione: (2024)
di: Hao, Xiang, et al.
Pubblicazione: (2024)
KS-Net: Multi-band joint speech restoration and enhancement network for 2024 ICASSP SSI Challenge
di: Yu, Guochen, et al.
Pubblicazione: (2024)
di: Yu, Guochen, et al.
Pubblicazione: (2024)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
di: Yin, Han, et al.
Pubblicazione: (2024)
di: Yin, Han, et al.
Pubblicazione: (2024)
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
di: Ooi, Kenneth, et al.
Pubblicazione: (2022)
di: Ooi, Kenneth, et al.
Pubblicazione: (2022)
Moises-Light: Resource-efficient Band-split U-Net For Music Source Separation
di: Yun-Ning, et al.
Pubblicazione: (2025)
di: Yun-Ning, et al.
Pubblicazione: (2025)
A Survey on Cross-Modal Interaction Between Music and Multimodal Data
di: Li, Sifei, et al.
Pubblicazione: (2025)
di: Li, Sifei, et al.
Pubblicazione: (2025)
Multi-band Frequency Reconstruction for Neural Psychoacoustic Coding
di: Ng, Dianwen, et al.
Pubblicazione: (2025)
di: Ng, Dianwen, et al.
Pubblicazione: (2025)
Joint Feature and Output Distillation for Low-complexity Acoustic Scene Classification
di: Li, Haowen, et al.
Pubblicazione: (2025)
di: Li, Haowen, et al.
Pubblicazione: (2025)
FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement
di: Hao, Xiang, et al.
Pubblicazione: (2020)
di: Hao, Xiang, et al.
Pubblicazione: (2020)
Squeeze-and-Excite ResNet-Conformers for Sound Event Localization, Detection, and Distance Estimation for DCASE 2024 Challenge
di: Yeow, Jun Wei, et al.
Pubblicazione: (2024)
di: Yeow, Jun Wei, et al.
Pubblicazione: (2024)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
di: Gao, Wenmiao, et al.
Pubblicazione: (2025)
di: Gao, Wenmiao, et al.
Pubblicazione: (2025)
SoundCTM: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation
di: Saito, Koichi, et al.
Pubblicazione: (2024)
di: Saito, Koichi, et al.
Pubblicazione: (2024)
Do neonates hear what we measure? Assessing neonatal ward soundscapes at the neonates ears
di: Lam, Bhan, et al.
Pubblicazione: (2025)
di: Lam, Bhan, et al.
Pubblicazione: (2025)
Determined blind source separation via modeling adjacent frequency band correlations in speech signals
di: Wang, Jianyu, et al.
Pubblicazione: (2025)
di: Wang, Jianyu, et al.
Pubblicazione: (2025)
Editing Music with Melody and Text: Using ControlNet for Diffusion Transformer
di: Hou, Siyuan, et al.
Pubblicazione: (2024)
di: Hou, Siyuan, et al.
Pubblicazione: (2024)
MusicHiFi: Fast High-Fidelity Stereo Vocoding
di: Zhu, Ge, et al.
Pubblicazione: (2024)
di: Zhu, Ge, et al.
Pubblicazione: (2024)
What do neural networks listen to? Exploring the crucial bands in Speech Enhancement using Sinc-convolution
di: Ho, Kuan-Hsun, et al.
Pubblicazione: (2024)
di: Ho, Kuan-Hsun, et al.
Pubblicazione: (2024)
ArtifactNet: Detecting AI-Generated Music via Forensic Residual Physics
di: Oh, Heewon
Pubblicazione: (2026)
di: Oh, Heewon
Pubblicazione: (2026)
Audio-Language Models for Audio-Centric Tasks: A Systematic Survey
di: Su, Yi, et al.
Pubblicazione: (2025)
di: Su, Yi, et al.
Pubblicazione: (2025)
Mixed-gradients Distributed Filtered Reference Least Mean Square Algorithm -- A Robust Distributed Multichannel Active Noise Control Algorithm
di: Ji, Junwei, et al.
Pubblicazione: (2025)
di: Ji, Junwei, et al.
Pubblicazione: (2025)
Detecting gamma-band responses to the speech envelope for the ICASSP 2024 Auditory EEG Decoding Signal Processing Grand Challenge
di: Thornton, Mike, et al.
Pubblicazione: (2024)
di: Thornton, Mike, et al.
Pubblicazione: (2024)
Documenti analoghi
-
AudioLog: LLMs-Powered Long Audio Logging with Hybrid Token-Semantic Contrastive Learning
di: Bai, Jisheng, et al.
Pubblicazione: (2023) -
A Stabilized Hybrid Active Noise Control Algorithm of GFANC and FxNLMS with Online Clustering
di: Luo, Zhengding, et al.
Pubblicazione: (2026) -
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Music Demixing Track
di: Fabbro, Giorgio, et al.
Pubblicazione: (2023) -
Sound event localization and classification using WASN in Outdoor Environment
di: Zhang, Dongzhe, et al.
Pubblicazione: (2024) -
Description on IEEE ICME 2024 Grand Challenge: Semi-supervised Acoustic Scene Classification under Domain Shift
di: Bai, Jisheng, et al.
Pubblicazione: (2024)