Deep low-latency joint speech transmission and enhancement over a gaussian channel
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bokaei, Mohammad, Jensen, Jesper, Doclo, Simon, Østergaard, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Steered Response Power Method for Sound Source Localization With Generic Acoustic Models
von: Müller, Kaspar, et al.
Veröffentlicht: (2025)
von: Müller, Kaspar, et al.
Veröffentlicht: (2025)
Head Orientation Estimation with Distributed Microphones Using Speech Radiation Patterns
von: Müller, Kaspar, et al.
Veröffentlicht: (2023)
von: Müller, Kaspar, et al.
Veröffentlicht: (2023)
Assisted RTF-Vector-Based Binaural Direction of Arrival Estimation Exploiting a Calibrated External Microphone Array
von: Fejgin, Daniel, et al.
Veröffentlicht: (2022)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2022)
DNN-Based Online Source Counting Based on Spatial Generalized Magnitude Squared Coherence
von: Gode, Henri, et al.
Veröffentlicht: (2026)
von: Gode, Henri, et al.
Veröffentlicht: (2026)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
Joint Minimum Processing Beamforming and Near-end Listening Enhancement
von: Fuglsig, Andreas J., et al.
Veröffentlicht: (2023)
von: Fuglsig, Andreas J., et al.
Veröffentlicht: (2023)
Subjective quality evaluation of personalized own voice reconstruction systems
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2025)
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2025)
MambAttention: Mamba with Multi-Head Attention for Generalizable Single-Channel Speech Enhancement
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enhancement
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
Single-channel speech enhancement by using psychoacoustical model inspired fusion framework
von: Samui, Suman
Veröffentlicht: (2022)
von: Samui, Suman
Veröffentlicht: (2022)
Modeling of Speech-dependent Own Voice Transfer Characteristics for Hearables with In-ear Microphones
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
Speech-dependent Modeling of Own Voice Transfer Characteristics for In-ear Microphones in Hearables
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
Exploiting an External Microphone for Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2025)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2025)
KS-Net: Multi-band joint speech restoration and enhancement network for 2024 ICASSP SSI Challenge
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
Modeling strategies for speech enhancement in the latent space of a neural audio codec
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
Diffusion-Based Speech Enhancement in Matched and Mismatched Conditions Using a Heun-Based Sampler
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
A two-step approach for speech enhancement in low-SNR scenarios using cyclostationary beamforming and DNNs
von: Bologni, Giovanni, et al.
Veröffentlicht: (2026)
von: Bologni, Giovanni, et al.
Veröffentlicht: (2026)
BRUDEX Database: Binaural Room Impulse Responses with Uniformly Distributed External Microphones
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
Single-channel speech enhancement using learnable loss mixup
von: Chang, Oscar, et al.
Veröffentlicht: (2023)
von: Chang, Oscar, et al.
Veröffentlicht: (2023)
Investigating the Design Space of Diffusion Models for Speech Enhancement
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
von: Serre, Thomas, et al.
Veröffentlicht: (2024)
The CHiME-7 UDASE task: Unsupervised domain adaptation for conversational speech enhancement
von: Leglaive, Simon, et al.
Veröffentlicht: (2023)
von: Leglaive, Simon, et al.
Veröffentlicht: (2023)
Fast-Converging Distributed Signal Estimation in Topology-Unconstrained Wireless Acoustic Sensor Networks
von: Didier, Paul, et al.
Veröffentlicht: (2025)
von: Didier, Paul, et al.
Veröffentlicht: (2025)
Unsupervised speech enhancement with spectral kurtosis and double deep priors
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
von: Ohnaka, Hien, et al.
Veröffentlicht: (2024)
SPGM: Prioritizing Local Features for enhanced speech separation performance
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2023)
Spatially Selective Active Noise Control for Open-fitting Hearables with Acausal Optimization
von: Xiao, Tong, et al.
Veröffentlicht: (2025)
von: Xiao, Tong, et al.
Veröffentlicht: (2025)
Array Geometry-Robust Attention-Based Neural Beamformer for Moving Speakers
von: Tammen, Marvin, et al.
Veröffentlicht: (2024)
von: Tammen, Marvin, et al.
Veröffentlicht: (2024)
Reference Microphone Selection for Guided Source Separation based on the Normalized L-p Norm
von: Lohmann, Anselm, et al.
Veröffentlicht: (2025)
von: Lohmann, Anselm, et al.
Veröffentlicht: (2025)
Time-domain sound field estimation using kernel ridge regression
von: Brunnström, Jesper, et al.
Veröffentlicht: (2025)
von: Brunnström, Jesper, et al.
Veröffentlicht: (2025)
Knowledge boosting during low-latency inference
von: Srinivas, Vidya, et al.
Veröffentlicht: (2024)
von: Srinivas, Vidya, et al.
Veröffentlicht: (2024)
Monaural speech enhancement on drone via Adapter based transfer learning
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
GDiffuSE: Diffusion-based speech enhancement with noise model guidance
von: Yanir, Efrayim, et al.
Veröffentlicht: (2025)
von: Yanir, Efrayim, et al.
Veröffentlicht: (2025)
A low latency attention module for streaming self-supervised speech representation learning
von: Ma, Jianbo, et al.
Veröffentlicht: (2023)
von: Ma, Jianbo, et al.
Veröffentlicht: (2023)
Using RLHF to align speech enhancement approaches to mean-opinion quality scores
von: Kumar, Anurag, et al.
Veröffentlicht: (2024)
von: Kumar, Anurag, et al.
Veröffentlicht: (2024)
Sound Zone Control Robust To Sound Speed Change
von: Bhattacharjee, Sankha Subhra, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Sankha Subhra, et al.
Veröffentlicht: (2024)
Online Single-Channel Audio-Based Sound Speed Estimation for Robust Multi-Channel Audio Control
von: Fuglsig, Andreas Jonas, et al.
Veröffentlicht: (2026)
von: Fuglsig, Andreas Jonas, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Steered Response Power Method for Sound Source Localization With Generic Acoustic Models
von: Müller, Kaspar, et al.
Veröffentlicht: (2025) -
Head Orientation Estimation with Distributed Microphones Using Speech Radiation Patterns
von: Müller, Kaspar, et al.
Veröffentlicht: (2023) -
Assisted RTF-Vector-Based Binaural Direction of Arrival Estimation Exploiting a Calibrated External Microphone Array
von: Fejgin, Daniel, et al.
Veröffentlicht: (2022) -
DNN-Based Online Source Counting Based on Spatial Generalized Magnitude Squared Coherence
von: Gode, Henri, et al.
Veröffentlicht: (2026) -
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)