Where's That Voice Coming? Continual Learning for Sound Source Localization
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Yang, Das, Rohan Kumar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TF-Mamba: A Time-Frequency Network for Sound Source Localization
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
XLSR-Mamba: A Dual-Column Bidirectional State Space Model for Spoofing Attack Detection
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
RawTFNet: A Lightweight CNN Architecture for Speech Anti-spoofing
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
AdaKWS: Towards Robust Keyword Spotting with Test-Time Adaptation
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
Analytic Class Incremental Learning for Sound Source Localization with Privacy Protection
por: Qian, Xinyuan, et al.
Publicado: (2024)
por: Qian, Xinyuan, et al.
Publicado: (2024)
Multilingual Source Tracing of Speech Deepfakes: A First Benchmark
por: Xuan, Xi, et al.
Publicado: (2025)
por: Xuan, Xi, et al.
Publicado: (2025)
IPDnet: A Universal Direct-Path IPD Estimation Network for Sound Source Localization
por: Wang, Yabo, et al.
Publicado: (2024)
por: Wang, Yabo, et al.
Publicado: (2024)
Fast Algorithm for Moving Sound Source
por: Yang, Dong
Publicado: (2025)
por: Yang, Dong
Publicado: (2025)
Steered Response Power for Sound Source Localization: A Tutorial Review
por: Grinstein, Eric, et al.
Publicado: (2024)
por: Grinstein, Eric, et al.
Publicado: (2024)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
por: Yang, Ningyuan, et al.
Publicado: (2026)
por: Yang, Ningyuan, et al.
Publicado: (2026)
Multi-modal Speech Enhancement with Limited Electromyography Channels
por: Feng, Fuyuan, et al.
Publicado: (2025)
por: Feng, Fuyuan, et al.
Publicado: (2025)
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
por: Yin, Jun, et al.
Publicado: (2025)
por: Yin, Jun, et al.
Publicado: (2025)
Continual Learning for Acoustic Event Classification
por: Xiao, Yang
Publicado: (2025)
por: Xiao, Yang
Publicado: (2025)
Dual Knowledge Distillation for Efficient Sound Event Detection
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Leveraging Sound Source Trajectories for Universal Sound Separation
por: Wu, Donghang, et al.
Publicado: (2024)
por: Wu, Donghang, et al.
Publicado: (2024)
A Steered Response Power Method for Sound Source Localization With Generic Acoustic Models
por: Müller, Kaspar, et al.
Publicado: (2025)
por: Müller, Kaspar, et al.
Publicado: (2025)
EnvSDD: Benchmarking Environmental Sound Deepfake Detection
por: Yin, Han, et al.
Publicado: (2025)
por: Yin, Han, et al.
Publicado: (2025)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
por: Hu, Jinbo, et al.
Publicado: (2023)
por: Hu, Jinbo, et al.
Publicado: (2023)
AFT: An Exemplar-Free Class Incremental Learning Method for Environmental Sound Classification
por: Chen, Xinyi, et al.
Publicado: (2025)
por: Chen, Xinyi, et al.
Publicado: (2025)
An Efficient GPU-based Implementation for Noise Robust Sound Source Localization
por: Lin, Zirui, et al.
Publicado: (2025)
por: Lin, Zirui, et al.
Publicado: (2025)
AnalyticKWS: Towards Exemplar-Free Analytic Class Incremental Learning for Small-footprint Keyword Spotting
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
Device Feature based on Graph Fourier Transformation with Logarithmic Processing For Detection of Replay Speech Attacks
por: He, Mingrui, et al.
Publicado: (2024)
por: He, Mingrui, et al.
Publicado: (2024)
A Few-Shot Learning Approach for Sound Source Distance Estimation Using Relation Networks
por: Sobhdel, Amirreza, et al.
Publicado: (2021)
por: Sobhdel, Amirreza, et al.
Publicado: (2021)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
por: Chen, Yuanjian, et al.
Publicado: (2025)
por: Chen, Yuanjian, et al.
Publicado: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
por: Yu, Hogeon
Publicado: (2025)
por: Yu, Hogeon
Publicado: (2025)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
por: Wang, Jiang, et al.
Publicado: (2025)
por: Wang, Jiang, et al.
Publicado: (2025)
SingMOS: An extensive Open-Source Singing Voice Dataset for MOS Prediction
por: Tang, Yuxun, et al.
Publicado: (2024)
por: Tang, Yuxun, et al.
Publicado: (2024)
Physics-Informed Transfer Learning for Data-Driven Sound Source Reconstruction in Near-Field Acoustic Holography
por: Luan, Xinmeng, et al.
Publicado: (2025)
por: Luan, Xinmeng, et al.
Publicado: (2025)
Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
por: Bai, Ye, et al.
Publicado: (2024)
por: Bai, Ye, et al.
Publicado: (2024)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
por: Zhang, Xueping, et al.
Publicado: (2025)
por: Zhang, Xueping, et al.
Publicado: (2025)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
por: Kim, Taewoo, et al.
Publicado: (2024)
por: Kim, Taewoo, et al.
Publicado: (2024)
Zero- and Few-shot Sound Event Localization and Detection
por: Shimada, Kazuki, et al.
Publicado: (2023)
por: Shimada, Kazuki, et al.
Publicado: (2023)
Audio Simulation for Sound Source Localization in Virtual Evironment
por: Di Yuan, Yi, et al.
Publicado: (2024)
por: Di Yuan, Yi, et al.
Publicado: (2024)
Attacking Voice Anonymization Systems with Augmented Feature and Speaker Identity Difference
por: Zhang, Yanzhe, et al.
Publicado: (2024)
por: Zhang, Yanzhe, et al.
Publicado: (2024)
Ejemplares similares
-
TF-Mamba: A Time-Frequency Network for Sound Source Localization
por: Xiao, Yang, et al.
Publicado: (2024) -
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
por: Xiao, Yang, et al.
Publicado: (2024) -
Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing
por: Xiao, Yang, et al.
Publicado: (2025) -
WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
por: Xiao, Yang, et al.
Publicado: (2024) -
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
por: Yin, Han, et al.
Publicado: (2024)