Leveraging Sound Source Trajectories for Universal Sound Separation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Donghang, Wu, Xihong, Qu, Tianshu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cross-attention Inspired Selective State Space Models for Target Sound Extraction
von: Wu, Donghang, et al.
Veröffentlicht: (2024)
von: Wu, Donghang, et al.
Veröffentlicht: (2024)
TSE-PI: Target Sound Extraction under Reverberant Environments with Pitch Information
von: Wang, Yiwen, et al.
Veröffentlicht: (2024)
von: Wang, Yiwen, et al.
Veröffentlicht: (2024)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds
von: Hasumi, Takuya, et al.
Veröffentlicht: (2025)
von: Hasumi, Takuya, et al.
Veröffentlicht: (2025)
Fast Algorithm for Moving Sound Source
von: Yang, Dong
Veröffentlicht: (2025)
von: Yang, Dong
Veröffentlicht: (2025)
IPDnet: A Universal Direct-Path IPD Estimation Network for Sound Source Localization
von: Wang, Yabo, et al.
Veröffentlicht: (2024)
von: Wang, Yabo, et al.
Veröffentlicht: (2024)
DiffSound: Differentiable Modal Sound Rendering and Inverse Rendering for Diverse Inference Tasks
von: Jin, Xutong, et al.
Veröffentlicht: (2024)
von: Jin, Xutong, et al.
Veröffentlicht: (2024)
FlowSep: Language-Queried Sound Separation with Rectified Flow Matching
von: Yuan, Yi, et al.
Veröffentlicht: (2024)
von: Yuan, Yi, et al.
Veröffentlicht: (2024)
Leveraging Audio-Only Data for Text-Queried Target Sound Extraction
von: Saijo, Kohei, et al.
Veröffentlicht: (2024)
von: Saijo, Kohei, et al.
Veröffentlicht: (2024)
Where's That Voice Coming? Continual Learning for Sound Source Localization
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Sound Source Separation Using Latent Variational Block-Wise Disentanglement
von: Helwani, Karim, et al.
Veröffentlicht: (2024)
von: Helwani, Karim, et al.
Veröffentlicht: (2024)
Universal Sound Separation with Self-Supervised Audio Masked Autoencoder
von: Zhao, Junqi, et al.
Veröffentlicht: (2024)
von: Zhao, Junqi, et al.
Veröffentlicht: (2024)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
Sound Zone Control Robust To Sound Speed Change
von: Bhattacharjee, Sankha Subhra, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Sankha Subhra, et al.
Veröffentlicht: (2024)
Analytic Class Incremental Learning for Sound Source Localization with Privacy Protection
von: Qian, Xinyuan, et al.
Veröffentlicht: (2024)
von: Qian, Xinyuan, et al.
Veröffentlicht: (2024)
TF-Mamba: A Time-Frequency Network for Sound Source Localization
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Steered Response Power for Sound Source Localization: A Tutorial Review
von: Grinstein, Eric, et al.
Veröffentlicht: (2024)
von: Grinstein, Eric, et al.
Veröffentlicht: (2024)
DENSE: Dynamic Embedding Causal Target Speech Extraction
von: Wang, Yiwen, et al.
Veröffentlicht: (2024)
von: Wang, Yiwen, et al.
Veröffentlicht: (2024)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
Evaluating Sound Similarity Metrics for Differentiable, Iterative Sound-Matching
von: Salimi, Amir, et al.
Veröffentlicht: (2025)
von: Salimi, Amir, et al.
Veröffentlicht: (2025)
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
von: Yin, Jun, et al.
Veröffentlicht: (2025)
von: Yin, Jun, et al.
Veröffentlicht: (2025)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
Sound Field Translation and Mixed Source Model for Virtual Applications with Perceptual Validation
von: Birnie, Lachlan, et al.
Veröffentlicht: (2020)
von: Birnie, Lachlan, et al.
Veröffentlicht: (2020)
Codec-SUPERB: An In-Depth Analysis of Sound Codec Models
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
A Steered Response Power Method for Sound Source Localization With Generic Acoustic Models
von: Müller, Kaspar, et al.
Veröffentlicht: (2025)
von: Müller, Kaspar, et al.
Veröffentlicht: (2025)
A Few-Shot Learning Approach for Sound Source Distance Estimation Using Relation Networks
von: Sobhdel, Amirreza, et al.
Veröffentlicht: (2021)
von: Sobhdel, Amirreza, et al.
Veröffentlicht: (2021)
Diffuse Sound Field Synthesis
von: Zotter, Franz, et al.
Veröffentlicht: (2024)
von: Zotter, Franz, et al.
Veröffentlicht: (2024)
Sound Event Bounding Boxes
von: Ebbers, Janek, et al.
Veröffentlicht: (2024)
von: Ebbers, Janek, et al.
Veröffentlicht: (2024)
Fractional Fourier Sound Synthesis
von: Gutiérrez, Esteban, et al.
Veröffentlicht: (2025)
von: Gutiérrez, Esteban, et al.
Veröffentlicht: (2025)
SoundBeam meets M2D: Target Sound Extraction with Audio Foundation Model
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
Physics-Informed Transfer Learning for Data-Driven Sound Source Reconstruction in Near-Field Acoustic Holography
von: Luan, Xinmeng, et al.
Veröffentlicht: (2025)
von: Luan, Xinmeng, et al.
Veröffentlicht: (2025)
SoundLoCD: An Efficient Conditional Discrete Contrastive Latent Diffusion Model for Text-to-Sound Generation
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
A Detailed Audio-Text Data Simulation Pipeline using Single-Event Sounds
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
DiveSound: LLM-Assisted Automatic Taxonomy Construction for Diverse Audio Generation
von: Li, Baihan, et al.
Veröffentlicht: (2024)
von: Li, Baihan, et al.
Veröffentlicht: (2024)
Sound Field Synthesis with Acoustic Waves
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Cross-attention Inspired Selective State Space Models for Target Sound Extraction
von: Wu, Donghang, et al.
Veröffentlicht: (2024) -
TSE-PI: Target Sound Extraction under Reverberant Environments with Pitch Information
von: Wang, Yiwen, et al.
Veröffentlicht: (2024) -
Exploring Text-Queried Sound Event Detection with Audio Source Separation
von: Yin, Han, et al.
Veröffentlicht: (2024) -
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
von: Yin, Han, et al.
Veröffentlicht: (2024) -
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)