Reverse Attention for Lightweight Speech Enhancement on Edge Devices
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ojha, Shuubham, Gervits, Felix, Espy-Wilson, Carol |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
RealClass: A Framework for Classroom Speech Simulation with Public Datasets and Game Engines
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
Improving Speech Inversion Through Self-Supervised Embeddings and Enhanced Tract Variables
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2023)
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2023)
Articulation-Informed ASR: Integrating Articulatory Features into ASR via Auxiliary Speech Inversion and Cross-Attention Fusion
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
Kid-Whisper: Towards Bridging the Performance Gap in Automatic Speech Recognition for Children VS. Adults
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2023)
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2023)
DroFiT: A Lightweight Band-fused Frequency Attention Toward Real-time UAV Speech Enhancement
von: Lee, Jeongmin, et al.
Veröffentlicht: (2025)
von: Lee, Jeongmin, et al.
Veröffentlicht: (2025)
Study of Lightweight Transformer Architectures for Single-Channel Speech Enhancement
von: Zhao, Haixin, et al.
Veröffentlicht: (2025)
von: Zhao, Haixin, et al.
Veröffentlicht: (2025)
GALD-SE: Guided Anisotropic Lightweight Diffusion for Efficient Speech Enhancement
von: Wang, Chengzhong, et al.
Veröffentlicht: (2024)
von: Wang, Chengzhong, et al.
Veröffentlicht: (2024)
Attention-Based Beamformer For Multi-Channel Speech Enhancement
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
A Lightweight Fourier-based Network for Binaural Speech Enhancement with Spatial Cue Preservation
von: Lu, Xikun, et al.
Veröffentlicht: (2025)
von: Lu, Xikun, et al.
Veröffentlicht: (2025)
Dual-View Predictive Diffusion: Lightweight Speech Enhancement via Spectrogram-Image Synergy
von: Xue, Ke, et al.
Veröffentlicht: (2026)
von: Xue, Ke, et al.
Veröffentlicht: (2026)
Dense-TSNet: Dense Connected Two-Stage Structure for Ultra-Lightweight Speech Enhancement
von: Lin, Zizhen, et al.
Veröffentlicht: (2024)
von: Lin, Zizhen, et al.
Veröffentlicht: (2024)
From Weak Labels to Strong Results: Utilizing 5,000 Hours of Noisy Classroom Transcripts with Minimal Accurate Data
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
Speech Enhancement with Overlapped-Frame Information Fusion and Causal Self-Attention
von: Zhang, Yuewei, et al.
Veröffentlicht: (2025)
von: Zhang, Yuewei, et al.
Veröffentlicht: (2025)
Robust One-step Speech Enhancement via Consistency Distillation
von: Xu, Liang, et al.
Veröffentlicht: (2025)
von: Xu, Liang, et al.
Veröffentlicht: (2025)
LABNet: A Lightweight Attentive Beamforming Network for Ad-hoc Multichannel Microphone Invariant Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2025)
von: Yan, Haoyin, et al.
Veröffentlicht: (2025)
SimClass: A Classroom Speech Dataset Generated via Game Engine Simulation For Automatic Speech Recognition Research
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)
Learnable Pulse Accumulation for On-Device Speech Recognition: How Much Attention Do You Need?
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
CPT-Boosted Wav2vec2.0: Towards Noise Robust Speech Recognition for Classroom Environments
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2024)
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2024)
Speech-Based Estimation of Schizophrenia Severity Using Feature Fusion
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
Plugin Speech Enhancement: A Universal Speech Enhancement Framework Inspired by Dynamic Neural Network
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
Robust Speech Recognition with Schrödinger Bridge-Based Speech Enhancement
von: Nasretdinov, Rauf, et al.
Veröffentlicht: (2025)
von: Nasretdinov, Rauf, et al.
Veröffentlicht: (2025)
Absorbing Discrete Diffusion for Speech Enhancement
von: Gonzalez, Philippe
Veröffentlicht: (2026)
von: Gonzalez, Philippe
Veröffentlicht: (2026)
LMFCA-Net: A Lightweight Model for Multi-Channel Speech Enhancement with Efficient Narrow-Band and Cross-Band Attention
von: Zhang, Yaokai, et al.
Veröffentlicht: (2025)
von: Zhang, Yaokai, et al.
Veröffentlicht: (2025)
Objective and Subjective Evaluation of Diffusion-Based Speech Enhancement for Dysarthric Speech
von: de Groot, Dimme, et al.
Veröffentlicht: (2025)
von: de Groot, Dimme, et al.
Veröffentlicht: (2025)
Advancing Electrolaryngeal Speech Enhancement Through Speech-Text Representation Learning
von: Ma, Ding, et al.
Veröffentlicht: (2026)
von: Ma, Ding, et al.
Veröffentlicht: (2026)
LISTEN: Lightweight Industrial Sound-representable Transformer for Edge Notification
von: Han, Changheon, et al.
Veröffentlicht: (2025)
von: Han, Changheon, et al.
Veröffentlicht: (2025)
Re-Parameterization of Lightweight Transformer for On-Device Speech Emotion Recognition
von: Zhang, Zixing, et al.
Veröffentlicht: (2024)
von: Zhang, Zixing, et al.
Veröffentlicht: (2024)
DISPATCH: Distilling Selective Patches for Speech Enhancement
von: Kim, Dohwan, et al.
Veröffentlicht: (2025)
von: Kim, Dohwan, et al.
Veröffentlicht: (2025)
Variational Autoencoder for Personalized Pathological Speech Enhancement
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
von: Hou, Mingchi, et al.
Veröffentlicht: (2025)
Universal Speech Enhancement with Regression and Generative Mamba
von: Chao, Rong, et al.
Veröffentlicht: (2025)
von: Chao, Rong, et al.
Veröffentlicht: (2025)
ICASSP 2026 URGENT Speech Enhancement Challenge
von: Li, Chenda, et al.
Veröffentlicht: (2026)
von: Li, Chenda, et al.
Veröffentlicht: (2026)
Investigating Training Objectives for Generative Speech Enhancement
von: Richter, Julius, et al.
Veröffentlicht: (2024)
von: Richter, Julius, et al.
Veröffentlicht: (2024)
Geneses: Unified Generative Speech Enhancement and Separation
von: Asai, Kohei, et al.
Veröffentlicht: (2026)
von: Asai, Kohei, et al.
Veröffentlicht: (2026)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
Using Speech Foundational Models in Loss Functions for Hearing Aid Speech Enhancement
von: Sutherland, Robert, et al.
Veröffentlicht: (2024)
von: Sutherland, Robert, et al.
Veröffentlicht: (2024)
Rethinking Speech Representation Aggregation in Speech Enhancement: A Phonetic Mutual Information Perspective
von: Han, Seungu, et al.
Veröffentlicht: (2026)
von: Han, Seungu, et al.
Veröffentlicht: (2026)
Rethinking Processing Distortions: Disentangling the Impact of Speech Enhancement Errors on Speech Recognition Performance
von: Ochiai, Tsubasa, et al.
Veröffentlicht: (2024)
von: Ochiai, Tsubasa, et al.
Veröffentlicht: (2024)
Adaptive Convolution for CNN-based Speech Enhancement Models
von: Wang, Dahan, et al.
Veröffentlicht: (2025)
von: Wang, Dahan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024) -
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026) -
RealClass: A Framework for Classroom Speech Simulation with Public Datasets and Game Engines
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025) -
Improving Speech Inversion Through Self-Supervised Embeddings and Enhanced Tract Variables
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2023) -
Articulation-Informed ASR: Integrating Articulatory Features into ASR via Auxiliary Speech Inversion and Cross-Attention Fusion
von: Attia, Ahmed Adel, et al.
Veröffentlicht: (2025)