TF-Mamba: A Time-Frequency Network for Sound Source Localization
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Yang, Das, Rohan Kumar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Where's That Voice Coming? Continual Learning for Sound Source Localization
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
XLSR-Mamba: A Dual-Column Bidirectional State Space Model for Spoofing Attack Detection
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
AdaKWS: Towards Robust Keyword Spotting with Test-Time Adaptation
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
RawTFNet: A Lightweight CNN Architecture for Speech Anti-spoofing
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
IPDnet: A Universal Direct-Path IPD Estimation Network for Sound Source Localization
por: Wang, Yabo, et al.
Publicado: (2024)
por: Wang, Yabo, et al.
Publicado: (2024)
Multilingual Source Tracing of Speech Deepfakes: A First Benchmark
por: Xuan, Xi, et al.
Publicado: (2025)
por: Xuan, Xi, et al.
Publicado: (2025)
Steered Response Power for Sound Source Localization: A Tutorial Review
por: Grinstein, Eric, et al.
Publicado: (2024)
por: Grinstein, Eric, et al.
Publicado: (2024)
Fast Algorithm for Moving Sound Source
por: Yang, Dong
Publicado: (2025)
por: Yang, Dong
Publicado: (2025)
Analytic Class Incremental Learning for Sound Source Localization with Privacy Protection
por: Qian, Xinyuan, et al.
Publicado: (2024)
por: Qian, Xinyuan, et al.
Publicado: (2024)
A Steered Response Power Method for Sound Source Localization With Generic Acoustic Models
por: Müller, Kaspar, et al.
Publicado: (2025)
por: Müller, Kaspar, et al.
Publicado: (2025)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
por: Yang, Ningyuan, et al.
Publicado: (2026)
por: Yang, Ningyuan, et al.
Publicado: (2026)
Multi-modal Speech Enhancement with Limited Electromyography Channels
por: Feng, Fuyuan, et al.
Publicado: (2025)
por: Feng, Fuyuan, et al.
Publicado: (2025)
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
por: Yin, Jun, et al.
Publicado: (2025)
por: Yin, Jun, et al.
Publicado: (2025)
SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation
por: Mu, Da, et al.
Publicado: (2024)
por: Mu, Da, et al.
Publicado: (2024)
MASV: Speaker Verification with Global and Local Context Mamba
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
Dual Knowledge Distillation for Efficient Sound Event Detection
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Leveraging Sound Source Trajectories for Universal Sound Separation
por: Wu, Donghang, et al.
Publicado: (2024)
por: Wu, Donghang, et al.
Publicado: (2024)
EnvSDD: Benchmarking Environmental Sound Deepfake Detection
por: Yin, Han, et al.
Publicado: (2025)
por: Yin, Han, et al.
Publicado: (2025)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
por: Gao, Wenmiao, et al.
Publicado: (2025)
por: Gao, Wenmiao, et al.
Publicado: (2025)
TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement
por: Saijo, Kohei, et al.
Publicado: (2024)
por: Saijo, Kohei, et al.
Publicado: (2024)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
por: Lee, Dongheon, et al.
Publicado: (2024)
por: Lee, Dongheon, et al.
Publicado: (2024)
MambaFoley: Foley Sound Generation using Selective State-Space Models
por: Colombo, Marco Furio, et al.
Publicado: (2024)
por: Colombo, Marco Furio, et al.
Publicado: (2024)
Frequency Dynamic Convolutions for Sound Event Detection
por: Nam, Hyeonuk
Publicado: (2025)
por: Nam, Hyeonuk
Publicado: (2025)
A Few-Shot Learning Approach for Sound Source Distance Estimation Using Relation Networks
por: Sobhdel, Amirreza, et al.
Publicado: (2021)
por: Sobhdel, Amirreza, et al.
Publicado: (2021)
Towards Understanding of Frequency Dependence on Sound Event Detection
por: Nam, Hyeonuk, et al.
Publicado: (2025)
por: Nam, Hyeonuk, et al.
Publicado: (2025)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
por: Hu, Jinbo, et al.
Publicado: (2024)
por: Hu, Jinbo, et al.
Publicado: (2024)
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
por: Gao, Wenmiao, et al.
Publicado: (2025)
por: Gao, Wenmiao, et al.
Publicado: (2025)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
por: Lee, Gyeong-Tae
Publicado: (2025)
por: Lee, Gyeong-Tae
Publicado: (2025)
Self Training and Ensembling Frequency Dependent Networks with Coarse Prediction Pooling and Sound Event Bounding Boxes
por: Nam, Hyeonuk, et al.
Publicado: (2024)
por: Nam, Hyeonuk, et al.
Publicado: (2024)
An Efficient GPU-based Implementation for Noise Robust Sound Source Localization
por: Lin, Zirui, et al.
Publicado: (2025)
por: Lin, Zirui, et al.
Publicado: (2025)
Leveraging Local and Global Knowledge Integration with Time-Frequency Calibrated Distillation for Speech Enhancement
por: Cheng, Jiaming, et al.
Publicado: (2025)
por: Cheng, Jiaming, et al.
Publicado: (2025)
Device Feature based on Graph Fourier Transformation with Logarithmic Processing For Detection of Replay Speech Attacks
por: He, Mingrui, et al.
Publicado: (2024)
por: He, Mingrui, et al.
Publicado: (2024)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
por: Chen, Yuanjian, et al.
Publicado: (2025)
por: Chen, Yuanjian, et al.
Publicado: (2025)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
por: Wang, Jiang, et al.
Publicado: (2025)
por: Wang, Jiang, et al.
Publicado: (2025)
Ejemplares similares
-
Where's That Voice Coming? Continual Learning for Sound Source Localization
por: Xiao, Yang, et al.
Publicado: (2024) -
XLSR-Mamba: A Dual-Column Bidirectional State Space Model for Spoofing Attack Detection
por: Xiao, Yang, et al.
Publicado: (2024) -
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
por: Xiao, Yang, et al.
Publicado: (2024) -
WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
por: Xiao, Yang, et al.
Publicado: (2024) -
Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing
por: Xiao, Yang, et al.
Publicado: (2025)