Guardado en:
| Autores principales: | Avenstrup, Thor Højhus, Elek, Boldizsár, Mádi, István László, Schin, András Bence, Mørup, Morten, Jensen, Bjørn Sand, Olsen, Kenny Falkær |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2410.20997 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Knowing When to Quit: Probabilistic Early Exits for Speech Separation
por: Olsen, Kenny Falkær, et al.
Publicado: (2025)
por: Olsen, Kenny Falkær, et al.
Publicado: (2025)
Omni-directional attention mechanism based on Mamba for speech separation
por: Xue, Ke, et al.
Publicado: (2026)
por: Xue, Ke, et al.
Publicado: (2026)
U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation
por: Dang, Shaoxiang, et al.
Publicado: (2024)
por: Dang, Shaoxiang, et al.
Publicado: (2024)
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
por: Wang, Junyu, et al.
Publicado: (2024)
por: Wang, Junyu, et al.
Publicado: (2024)
Self-Supervised Learning for Time Series: A Review & Critique of FITS
por: Eefsen, Andreas Løvendahl, et al.
Publicado: (2024)
por: Eefsen, Andreas Løvendahl, et al.
Publicado: (2024)
TC-BiMamba: Trans-Chunk bidirectionally within BiMamba for unified streaming and non-streaming ASR
por: She, Qingshun, et al.
Publicado: (2026)
por: She, Qingshun, et al.
Publicado: (2026)
MambAttention: Mamba with Multi-Head Attention for Generalizable Single-Channel Speech Enhancement
por: Kühne, Nikolai Lund, et al.
Publicado: (2025)
por: Kühne, Nikolai Lund, et al.
Publicado: (2025)
BiCrossMamba-ST: Speech Deepfake Detection with Bidirectional Mamba Spectro-Temporal Cross-Attention
por: Kheir, Yassine El, et al.
Publicado: (2025)
por: Kheir, Yassine El, et al.
Publicado: (2025)
Online speaker diarization of meetings guided by speech separation
por: Gruttadauria, Elio, et al.
Publicado: (2024)
por: Gruttadauria, Elio, et al.
Publicado: (2024)
Exploring the Capability of Mamba in Speech Applications
por: Miyazaki, Koichi, et al.
Publicado: (2024)
por: Miyazaki, Koichi, et al.
Publicado: (2024)
Mamba in Speech: Towards an Alternative to Self-Attention
por: Zhang, Xiangyu, et al.
Publicado: (2024)
por: Zhang, Xiangyu, et al.
Publicado: (2024)
Mamba-based Segmentation Model for Speaker Diarization
por: Plaquet, Alexis, et al.
Publicado: (2024)
por: Plaquet, Alexis, et al.
Publicado: (2024)
Universal Speech Enhancement with Regression and Generative Mamba
por: Chao, Rong, et al.
Publicado: (2025)
por: Chao, Rong, et al.
Publicado: (2025)
Rethinking Mamba in Speech Processing by Self-Supervised Models
por: Zhang, Xiangyu, et al.
Publicado: (2024)
por: Zhang, Xiangyu, et al.
Publicado: (2024)
Hierarchical speaker representation for target speaker extraction
por: He, Shulin, et al.
Publicado: (2022)
por: He, Shulin, et al.
Publicado: (2022)
Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enhancement
por: Kühne, Nikolai Lund, et al.
Publicado: (2025)
por: Kühne, Nikolai Lund, et al.
Publicado: (2025)
An Investigation of Incorporating Mamba for Speech Enhancement
por: Chao, Rong, et al.
Publicado: (2024)
por: Chao, Rong, et al.
Publicado: (2024)
MASV: Speaker Verification with Global and Local Context Mamba
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
Text adaptation for speaker verification with speaker-text factorized embeddings
por: Yang, Yexin, et al.
Publicado: (2025)
por: Yang, Yexin, et al.
Publicado: (2025)
Improving curriculum learning for target speaker extraction with synthetic speakers
por: Liu, Yun, et al.
Publicado: (2024)
por: Liu, Yun, et al.
Publicado: (2024)
Quantifying the effect of speech pathology on automatic and human speaker verification
por: Halpern, Bence Mark, et al.
Publicado: (2024)
por: Halpern, Bence Mark, et al.
Publicado: (2024)
Mamba for Streaming ASR Combined with Unimodal Aggregation
por: Fang, Ying, et al.
Publicado: (2024)
por: Fang, Ying, et al.
Publicado: (2024)
MamTra: A Hybrid Mamba-Transformer Backbone for Speech Synthesis
por: Nguyen, Tan Dat, et al.
Publicado: (2026)
por: Nguyen, Tan Dat, et al.
Publicado: (2026)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
por: Ding, Hanyu, et al.
Publicado: (2025)
por: Ding, Hanyu, et al.
Publicado: (2025)
Target speaker anonymization in multi-speaker recordings
por: Tomashenko, Natalia, et al.
Publicado: (2025)
por: Tomashenko, Natalia, et al.
Publicado: (2025)
MambaRate: Speech Quality Assessment Across Different Sampling Rates
por: Kakoulidis, Panos, et al.
Publicado: (2025)
por: Kakoulidis, Panos, et al.
Publicado: (2025)
Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement
por: Chao, Rong, et al.
Publicado: (2025)
por: Chao, Rong, et al.
Publicado: (2025)
TF-Mamba: A Time-Frequency Network for Sound Source Localization
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
por: Gao, Wenmiao, et al.
Publicado: (2025)
por: Gao, Wenmiao, et al.
Publicado: (2025)
SAM: A Mamba-2 State-Space Audio-Language Model
por: Lee, Taehan, et al.
Publicado: (2025)
por: Lee, Taehan, et al.
Publicado: (2025)
XLSR-MamBo: Scaling the Hybrid Mamba-Attention Backbone for Audio Deepfake Detection
por: Ng, Kwok-Ho, et al.
Publicado: (2026)
por: Ng, Kwok-Ho, et al.
Publicado: (2026)
Triage knowledge distillation for speaker verification
por: Kim, Ju-ho, et al.
Publicado: (2026)
por: Kim, Ju-ho, et al.
Publicado: (2026)
Privacy-oriented manipulation of speaker representations
por: Teixeira, Francisco, et al.
Publicado: (2023)
por: Teixeira, Francisco, et al.
Publicado: (2023)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
por: Lee, Dongheon, et al.
Publicado: (2024)
por: Lee, Dongheon, et al.
Publicado: (2024)
TAME: Temporal Audio-based Mamba for Enhanced Drone Trajectory Estimation and Classification
por: Xiao, Zhenyuan, et al.
Publicado: (2024)
por: Xiao, Zhenyuan, et al.
Publicado: (2024)
MambaFoley: Foley Sound Generation using Selective State-Space Models
por: Colombo, Marco Furio, et al.
Publicado: (2024)
por: Colombo, Marco Furio, et al.
Publicado: (2024)
Mamba-based Decoder-Only Approach with Bidirectional Speech Modeling for Speech Recognition
por: Masuyama, Yoshiki, et al.
Publicado: (2024)
por: Masuyama, Yoshiki, et al.
Publicado: (2024)
Speech-Mamba: Long-Context Speech Recognition with Selective State Spaces Models
por: Gao, Xiaoxue, et al.
Publicado: (2024)
por: Gao, Xiaoxue, et al.
Publicado: (2024)
A Two-Stage Band-Split Mamba-2 Network For Music Separation
por: Bai, Jinglin, et al.
Publicado: (2024)
por: Bai, Jinglin, et al.
Publicado: (2024)
Fake-Mamba: Real-Time Speech Deepfake Detection Using Bidirectional Mamba as Self-Attention's Alternative
por: Xuan, Xi, et al.
Publicado: (2025)
por: Xuan, Xi, et al.
Publicado: (2025)
Ejemplares similares
-
Knowing When to Quit: Probabilistic Early Exits for Speech Separation
por: Olsen, Kenny Falkær, et al.
Publicado: (2025) -
Omni-directional attention mechanism based on Mamba for speech separation
por: Xue, Ke, et al.
Publicado: (2026) -
U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation
por: Dang, Shaoxiang, et al.
Publicado: (2024) -
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
por: Wang, Junyu, et al.
Publicado: (2024) -
Self-Supervised Learning for Time Series: A Review & Critique of FITS
por: Eefsen, Andreas Løvendahl, et al.
Publicado: (2024)