Analytic Class Incremental Learning for Sound Source Localization with Privacy Protection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qian, Xinyuan, Yue, Xianghu, Wang, Jiadong, Zhuang, Huiping, Li, Haizhou |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AnalyticKWS: Towards Exemplar-Free Analytic Class Incremental Learning for Small-footprint Keyword Spotting
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Audio-Visual Target Speaker Extraction with Reverse Selective Auditory Attention
von: Tao, Ruijie, et al.
Veröffentlicht: (2024)
von: Tao, Ruijie, et al.
Veröffentlicht: (2024)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
AFT: An Exemplar-Free Class Incremental Learning Method for Environmental Sound Classification
von: Chen, Xinyi, et al.
Veröffentlicht: (2025)
von: Chen, Xinyi, et al.
Veröffentlicht: (2025)
UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook
von: Jiang, Yidi, et al.
Veröffentlicht: (2025)
von: Jiang, Yidi, et al.
Veröffentlicht: (2025)
AV-SSAN: Audio-Visual Selective DoA Estimation through Explicit Multi-Band Semantic-Spatial Alignment
von: Chen, Yu, et al.
Veröffentlicht: (2025)
von: Chen, Yu, et al.
Veröffentlicht: (2025)
Where's That Voice Coming? Continual Learning for Sound Source Localization
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Class-Incremental Learning for Multi-Label Audio Classification
von: Mulimani, Manjunath, et al.
Veröffentlicht: (2024)
von: Mulimani, Manjunath, et al.
Veröffentlicht: (2024)
Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
IPDnet: A Universal Direct-Path IPD Estimation Network for Sound Source Localization
von: Wang, Yabo, et al.
Veröffentlicht: (2024)
von: Wang, Yabo, et al.
Veröffentlicht: (2024)
AudioCIL: A Python Toolbox for Audio Class-Incremental Learning with Multiple Scenes
von: Xu, Qisheng, et al.
Veröffentlicht: (2024)
von: Xu, Qisheng, et al.
Veröffentlicht: (2024)
CoAVT: A Cognition-Inspired Unified Audio-Visual-Text Pre-Training Model for Multimodal Processing
von: Yue, Xianghu, et al.
Veröffentlicht: (2024)
von: Yue, Xianghu, et al.
Veröffentlicht: (2024)
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
TF-Mamba: A Time-Frequency Network for Sound Source Localization
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Steered Response Power for Sound Source Localization: A Tutorial Review
von: Grinstein, Eric, et al.
Veröffentlicht: (2024)
von: Grinstein, Eric, et al.
Veröffentlicht: (2024)
Human-Inspired Computing for Robust and Efficient Audio-Visual Speech Recognition
von: Liu, Qianhui, et al.
Veröffentlicht: (2024)
von: Liu, Qianhui, et al.
Veröffentlicht: (2024)
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
von: Yin, Jun, et al.
Veröffentlicht: (2025)
von: Yin, Jun, et al.
Veröffentlicht: (2025)
Region-Specific Audio Tagging for Spatial Sound
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
Leveraging Sound Source Trajectories for Universal Sound Separation
von: Wu, Donghang, et al.
Veröffentlicht: (2024)
von: Wu, Donghang, et al.
Veröffentlicht: (2024)
A Steered Response Power Method for Sound Source Localization With Generic Acoustic Models
von: Müller, Kaspar, et al.
Veröffentlicht: (2025)
von: Müller, Kaspar, et al.
Veröffentlicht: (2025)
Mamba in Speech: Towards an Alternative to Self-Attention
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
Investigating Effective Speaker Property Privacy Protection in Federated Learning for Speech Emotion Recognition
von: Tan, Chao, et al.
Veröffentlicht: (2024)
von: Tan, Chao, et al.
Veröffentlicht: (2024)
M-Vec: Matryoshka Speaker Embeddings with Flexible Dimensions
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
VoiceBench: Benchmarking LLM-Based Voice Assistants
von: Chen, Yiming, et al.
Veröffentlicht: (2024)
von: Chen, Yiming, et al.
Veröffentlicht: (2024)
Fast Algorithm for Moving Sound Source
von: Yang, Dong
Veröffentlicht: (2025)
von: Yang, Dong
Veröffentlicht: (2025)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
SoundCollage: Automated Discovery of New Classes in Audio Datasets
von: Choi, Ryuhaerang, et al.
Veröffentlicht: (2024)
von: Choi, Ryuhaerang, et al.
Veröffentlicht: (2024)
An Efficient GPU-based Implementation for Noise Robust Sound Source Localization
von: Lin, Zirui, et al.
Veröffentlicht: (2025)
von: Lin, Zirui, et al.
Veröffentlicht: (2025)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
A Few-Shot Learning Approach for Sound Source Distance Estimation Using Relation Networks
von: Sobhdel, Amirreza, et al.
Veröffentlicht: (2021)
von: Sobhdel, Amirreza, et al.
Veröffentlicht: (2021)
Hierarchical Control of Emotion Rendering in Speech Synthesis
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
Hierarchical Emotion Prediction and Control in Text-to-Speech Synthesis
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
Fine-Grained Quantitative Emotion Editing for Speech Generation
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
PhiNet: Speaker Verification with Phonetic Interpretability
von: Ma, Yi, et al.
Veröffentlicht: (2026)
von: Ma, Yi, et al.
Veröffentlicht: (2026)
AudioRAG: A Challenging Benchmark for Audio Reasoning and Information Retrieval
von: Lin, Jingru, et al.
Veröffentlicht: (2026)
von: Lin, Jingru, et al.
Veröffentlicht: (2026)
Multi-Step Prediction and Control of Hierarchical Emotion Distribution in Text-to-Speech Synthesis
von: Inoue, Sho, et al.
Veröffentlicht: (2025)
von: Inoue, Sho, et al.
Veröffentlicht: (2025)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
von: Yu, Hogeon
Veröffentlicht: (2025)
von: Yu, Hogeon
Veröffentlicht: (2025)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation
von: Mu, Da, et al.
Veröffentlicht: (2024)
von: Mu, Da, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AnalyticKWS: Towards Exemplar-Free Analytic Class Incremental Learning for Small-footprint Keyword Spotting
von: Xiao, Yang, et al.
Veröffentlicht: (2025) -
Audio-Visual Target Speaker Extraction with Reverse Selective Auditory Attention
von: Tao, Ruijie, et al.
Veröffentlicht: (2024) -
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
von: Xiao, Yang, et al.
Veröffentlicht: (2024) -
AFT: An Exemplar-Free Class Incremental Learning Method for Environmental Sound Classification
von: Chen, Xinyi, et al.
Veröffentlicht: (2025) -
UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook
von: Jiang, Yidi, et al.
Veröffentlicht: (2025)