FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Shengkui, Ma, Bin, Watcharasupat, Karn N., Gan, Woon-Seng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Monaural Speech Enhancement with Complex Convolutional Block Attention Module and Joint Time Frequency Losses
von: Zhao, Shengkui, et al.
Veröffentlicht: (2021)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2021)
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
MossFormer2: Combining Transformer and RNN-Free Recurrent Network for Enhanced Time-Domain Monaural Speech Separation
von: Zhao, Shengkui, et al.
Veröffentlicht: (2023)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2023)
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
Quantifying Spatial Audio Quality Impairment
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
Conditional Latent Diffusion-Based Speech Enhancement Via Dual Context Learning
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
Complex-Cycle-Consistent Diffusion Model for Monaural Speech Enhancement
von: Li, Yi, et al.
Veröffentlicht: (2024)
von: Li, Yi, et al.
Veröffentlicht: (2024)
LORT: Locally Refined Convolution and Taylor Transformer for Monaural Speech Enhancement
von: Wang, Junyu, et al.
Veröffentlicht: (2025)
von: Wang, Junyu, et al.
Veröffentlicht: (2025)
Automating Urban Soundscape Enhancements with AI: In-situ Assessment of Quality and Restorativeness in Traffic-Exposed Residential Areas
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
An Empirical Study on the Impact of Positional Encoding in Transformer-based Monaural Speech Enhancement
von: Zhang, Qiquan, et al.
Veröffentlicht: (2024)
von: Zhang, Qiquan, et al.
Veröffentlicht: (2024)
Separate This, and All of these Things Around It: Music Source Separation via Hyperellipsoidal Queries
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
Spectral Masking with Explicit Time-Context Windowing for Neural Network-Based Monaural Speech Enhancement
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2024)
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2024)
ZipEnhancer: Dual-Path Down-Up Sampling-based Zipformer for Monaural Speech Enhancement
von: Wang, Haoxu, et al.
Veröffentlicht: (2025)
von: Wang, Haoxu, et al.
Veröffentlicht: (2025)
BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech Enhancement
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
von: Wang, Junyu, et al.
Veröffentlicht: (2024)
von: Wang, Junyu, et al.
Veröffentlicht: (2024)
HiFi-SR: A Unified Generative Transformer-Convolutional Adversarial Network for High-Fidelity Speech Super-Resolution
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement
von: Abdulatif, Sherif, et al.
Veröffentlicht: (2022)
von: Abdulatif, Sherif, et al.
Veröffentlicht: (2022)
Towards Decoupling Frontend Enhancement and Backend Recognition in Monaural Robust ASR
von: Yang, Yufeng, et al.
Veröffentlicht: (2024)
von: Yang, Yufeng, et al.
Veröffentlicht: (2024)
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
A Stabilized Hybrid Active Noise Control Algorithm of GFANC and FxNLMS with Online Clustering
von: Luo, Zhengding, et al.
Veröffentlicht: (2026)
von: Luo, Zhengding, et al.
Veröffentlicht: (2026)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
Advancing Electrolaryngeal Speech Enhancement Through Speech-Text Representation Learning
von: Ma, Ding, et al.
Veröffentlicht: (2026)
von: Ma, Ding, et al.
Veröffentlicht: (2026)
Remastering Divide and Remaster: A Cinematic Audio Source Separation Dataset with Multilingual Support
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
Facing the Music: Tackling Singing Voice Separation in Cinematic Audio Source Separation
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
Towards Audio Codec-based Speech Separation
von: Yip, Jia Qi, et al.
Veröffentlicht: (2024)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2024)
Learning Time-Graph Frequency Representation for Monaural Speech Enhancement
von: Wang, Tingting, et al.
Veröffentlicht: (2025)
von: Wang, Tingting, et al.
Veröffentlicht: (2025)
Monaural speech enhancement on drone via Adapter based transfer learning
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
Spiking Structured State Space Model for Monaural Speech Enhancement
von: Du, Yu, et al.
Veröffentlicht: (2023)
von: Du, Yu, et al.
Veröffentlicht: (2023)
Dynamic Frequency-Adaptive Knowledge Distillation for Speech Enhancement
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
Sub-band and Full-band Interactive U-Net with DPRNN for Demixing Cross-talk Stereo Music
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
Online Audio-Visual Autoregressive Speaker Extraction
von: Pan, Zexu, et al.
Veröffentlicht: (2025)
von: Pan, Zexu, et al.
Veröffentlicht: (2025)
Uncertainty Estimation in the Real World: A Study on Music Emotion Recognition
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
Emotional Dimension Control in Language Model-Based Text-to-Speech: Spanning a Broad Spectrum of Human Emotions
von: Zhou, Kun, et al.
Veröffentlicht: (2024)
von: Zhou, Kun, et al.
Veröffentlicht: (2024)
SPGM: Prioritizing Local Features for enhanced speech separation performance
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
Extracting Urban Sound Information for Residential Areas in Smart Cities Using an End-to-End IoT System
von: Tan, Ee-Leng, et al.
Veröffentlicht: (2024)
von: Tan, Ee-Leng, et al.
Veröffentlicht: (2024)
E2E-AEC: Implementing an end-to-end neural network learning approach for acoustic echo cancellation
von: Jiang, Yiheng, et al.
Veröffentlicht: (2026)
von: Jiang, Yiheng, et al.
Veröffentlicht: (2026)
A Real-Time Platform for Portable and Scalable Active Noise Mitigation for Construction Machinery
von: Gan, Woon-Seng, et al.
Veröffentlicht: (2024)
von: Gan, Woon-Seng, et al.
Veröffentlicht: (2024)
Rethinking Speech Representation Aggregation in Speech Enhancement: A Phonetic Mutual Information Perspective
von: Han, Seungu, et al.
Veröffentlicht: (2026)
von: Han, Seungu, et al.
Veröffentlicht: (2026)
Leveraging Local and Global Knowledge Integration with Time-Frequency Calibrated Distillation for Speech Enhancement
von: Cheng, Jiaming, et al.
Veröffentlicht: (2025)
von: Cheng, Jiaming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Monaural Speech Enhancement with Complex Convolutional Block Attention Module and Joint Time Frequency Losses
von: Zhao, Shengkui, et al.
Veröffentlicht: (2021) -
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023) -
MossFormer2: Combining Transformer and RNN-Free Recurrent Network for Enhanced Time-Domain Monaural Speech Separation
von: Zhao, Shengkui, et al.
Veröffentlicht: (2023) -
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022) -
Quantifying Spatial Audio Quality Impairment
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)