Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Dongheon, Pandey, Ashutosh, Parekh, Sanjeel, Wong, Daniel, Donley, Jacob, Xu, Buye, Azcarreta, Juan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ArrayDPS-Refine: Generative Refinement of Discriminative Multi-Channel Speech Enhancement
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026)
Unified Diffusion Refinement for Multi-Channel Speech Enhancement and Separation
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026)
All Neural Low-latency Directional Speech Extraction
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
On the Importance of Neural Wiener Filter for Resource Efficient Multichannel Speech Enhancement
von: Hsieh, Tsun-An, et al.
Veröffentlicht: (2024)
von: Hsieh, Tsun-An, et al.
Veröffentlicht: (2024)
Decoupled Spatial and Temporal Processing for Resource Efficient Multichannel Speech Enhancement
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
Efficient Audiovisual Speech Processing via MUTUD: Multimodal Training and Unimodal Deployment
von: Hong, Joanna, et al.
Veröffentlicht: (2025)
von: Hong, Joanna, et al.
Veröffentlicht: (2025)
A Novel Deep Learning Framework for Efficient Multichannel Acoustic Feedback Control
von: Wu, Yuan-Kuei, et al.
Veröffentlicht: (2025)
von: Wu, Yuan-Kuei, et al.
Veröffentlicht: (2025)
Improving Resource-Efficient Speech Enhancement via Neural Differentiable DSP Vocoder Refinement
von: Guimarães, Heitor R., et al.
Veröffentlicht: (2025)
von: Guimarães, Heitor R., et al.
Veröffentlicht: (2025)
Controlling the Parameterized Multi-channel Wiener Filter using a tiny neural network
von: Grinstein, Eric, et al.
Veröffentlicht: (2025)
von: Grinstein, Eric, et al.
Veröffentlicht: (2025)
Spatially constrained vs. unconstrained filtering in neural spatiospectral filters for multichannel speech enhancement
von: Briegleb, Annika, et al.
Veröffentlicht: (2024)
von: Briegleb, Annika, et al.
Veröffentlicht: (2024)
Sound Event Detection with Boundary-Aware Optimization and Inference
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
von: Schmid, Florian, et al.
Veröffentlicht: (2026)
Conditional Flow Matching for Visually-Guided Acoustic Highlighting
von: Malard, Hugo, et al.
Veröffentlicht: (2026)
von: Malard, Hugo, et al.
Veröffentlicht: (2026)
Dynamic Gated Recurrent Neural Network for Compute-efficient Speech Enhancement
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
DeFTAN-II: Efficient Multichannel Speech Enhancement with Subgroup Processing
von: Lee, Dongheon, et al.
Veröffentlicht: (2023)
von: Lee, Dongheon, et al.
Veröffentlicht: (2023)
FoVNet: Configurable Field-of-View Speech Enhancement with Low Computation and Distortion for Smart Glasses
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
Speech enhancement deep-learning architecture for efficient edge processing
von: Pal, Monisankha, et al.
Veröffentlicht: (2024)
von: Pal, Monisankha, et al.
Veröffentlicht: (2024)
Real-time multichannel deep speech enhancement in hearing aids: Comparing monaural and binaural processing in complex acoustic scenarios
von: Westhausen, Nils L., et al.
Veröffentlicht: (2024)
von: Westhausen, Nils L., et al.
Veröffentlicht: (2024)
Self-Guided Target Sound Extraction and Classification Through Universal Sound Separation Model and Multiple Clues
von: Kwon, Younghoo, et al.
Veröffentlicht: (2025)
von: Kwon, Younghoo, et al.
Veröffentlicht: (2025)
Modulating State Space Model with SlowFast Framework for Compute-Efficient Ultra Low-Latency Speech Enhancement
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
von: Cheng, Longbiao, et al.
Veröffentlicht: (2024)
DeepASA: An Object-Oriented Multi-Purpose Network for Auditory Scene Analysis
von: Lee, Dongheon, et al.
Veröffentlicht: (2025)
von: Lee, Dongheon, et al.
Veröffentlicht: (2025)
I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
Design and Analysis of Binaural Signal Matching with Arbitrary Microphone Arrays and Listener Head Rotations
von: Madmoni, Lior, et al.
Veröffentlicht: (2024)
von: Madmoni, Lior, et al.
Veröffentlicht: (2024)
Ambisonics Encoding For Arbitrary Microphone Arrays Incorporating Residual Channels For Binaural Reproduction
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2024)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2024)
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
Predicting speech intelligibility in older adults for speech enhancement using the Gammachirp Envelope Similarity Index, GESI
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2025)
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2025)
SpatialCodec: Neural Spatial Speech Coding
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2023)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2023)
Performance and Robustness of Signal-Dependent vs. Signal-Independent Binaural Signal Matching with Wearable Microphone Arrays
von: Berger, Ami, et al.
Veröffentlicht: (2024)
von: Berger, Ami, et al.
Veröffentlicht: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
Spatial-Filter-Bank-Based Neural Method for Multichannel Speech Enhancement
von: Zheng, Tianqin, et al.
Veröffentlicht: (2025)
von: Zheng, Tianqin, et al.
Veröffentlicht: (2025)
An Analysis of Joint Nonlinear Spatial Filtering for Spatial Aliasing Reduction
von: Mannanova, Alina, et al.
Veröffentlicht: (2025)
von: Mannanova, Alina, et al.
Veröffentlicht: (2025)
Spatially-Augmented Sequence-to-Sequence Neural Diarization for Meetings
von: Li, Li, et al.
Veröffentlicht: (2025)
von: Li, Li, et al.
Veröffentlicht: (2025)
TokenSE: a Mamba-based discrete token speech enhancement framework for cochlear implants
von: Chiang, Hsin-Tien, et al.
Veröffentlicht: (2026)
von: Chiang, Hsin-Tien, et al.
Veröffentlicht: (2026)
Universal Spatial Audio Transcoder
von: Sagasti, Amaia, et al.
Veröffentlicht: (2024)
von: Sagasti, Amaia, et al.
Veröffentlicht: (2024)
Audio-Visual Speech Enhancement for Spatial Audio - Spatial-VisualVoice and the MAVE Database
von: Yaffe, Danielle, et al.
Veröffentlicht: (2025)
von: Yaffe, Danielle, et al.
Veröffentlicht: (2025)
Unsupervised speech enhancement with spectral kurtosis and double deep priors
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
SPGM: Prioritizing Local Features for enhanced speech separation performance
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
Non-invasive electromyographic speech neuroprosthesis: a geometric perspective
von: Gowda, Harshavardhana T., et al.
Veröffentlicht: (2025)
von: Gowda, Harshavardhana T., et al.
Veröffentlicht: (2025)
The CHiME-7 UDASE task: Unsupervised domain adaptation for conversational speech enhancement
von: Leglaive, Simon, et al.
Veröffentlicht: (2023)
von: Leglaive, Simon, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
ArrayDPS-Refine: Generative Refinement of Discriminative Multi-Channel Speech Enhancement
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026) -
Unified Diffusion Refinement for Multi-Channel Speech Enhancement and Separation
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2026) -
All Neural Low-latency Directional Speech Extraction
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024) -
On the Importance of Neural Wiener Filter for Resource Efficient Multichannel Speech Enhancement
von: Hsieh, Tsun-An, et al.
Veröffentlicht: (2024) -
Decoupled Spatial and Temporal Processing for Resource Efficient Multichannel Speech Enhancement
von: Pandey, Ashutosh, et al.
Veröffentlicht: (2024)