Salvato in:
| Autori principali: | Liang, Chengsi, Sun, Yao, Thomas, Christo Kurisummoottil, Mohjazi, Lina, Saad, Walid |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2407.12203 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Description and Discussion on DCASE 2026 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2026)
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2026)
Sound Separation and Classification with Object and Semantic Guidance
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
BUT Systems for Environmental Sound Deepfake Detection in the ESDD 2026 Challenge
di: Peng, Junyi, et al.
Pubblicazione: (2025)
di: Peng, Junyi, et al.
Pubblicazione: (2025)
On-device Internet of Sounds Sonification with Wavetable Synthesis Techniques for Soil Moisture Monitoring in Water Scarcity Contexts
di: Roddy, Stephen
Pubblicazione: (2025)
di: Roddy, Stephen
Pubblicazione: (2025)
Robust Semantic Communications for Speech Transmission
di: Weng, Zhenzi, et al.
Pubblicazione: (2024)
di: Weng, Zhenzi, et al.
Pubblicazione: (2024)
Domain-Agnostic Incremental Learning for Sound Classification. A DCASE 2026 Challenge task
di: Casciotti, Riccardo, et al.
Pubblicazione: (2026)
di: Casciotti, Riccardo, et al.
Pubblicazione: (2026)
Semantic Communication with Hopfield Memories
di: Nasreddine, Karim, et al.
Pubblicazione: (2025)
di: Nasreddine, Karim, et al.
Pubblicazione: (2025)
Importance-Weighted Domain Adaptation for Sound Source Tracking
di: Zhong, Bingxiang, et al.
Pubblicazione: (2025)
di: Zhong, Bingxiang, et al.
Pubblicazione: (2025)
ACES: Evaluating Automated Audio Captioning Models on the Semantics of Sounds
di: Wijngaard, Gijs, et al.
Pubblicazione: (2024)
di: Wijngaard, Gijs, et al.
Pubblicazione: (2024)
Large Model Empowered Streaming Speech Semantic Communications
di: Weng, Zhenzi, et al.
Pubblicazione: (2025)
di: Weng, Zhenzi, et al.
Pubblicazione: (2025)
Baseline Systems and Evaluation Metrics for Spatial Semantic Segmentation of Sound Scenes
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2025)
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2025)
Class-Aware Permutation-Invariant Signal-to-Distortion Ratio for Semantic Segmentation of Sound Scene with Same-Class Sources
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2026)
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2026)
Squeeze-and-Excite ResNet-Conformers for Sound Event Localization, Detection, and Distance Estimation for DCASE 2024 Challenge
di: Yeow, Jun Wei, et al.
Pubblicazione: (2024)
di: Yeow, Jun Wei, et al.
Pubblicazione: (2024)
A Dual-Path Framework with Frequency-and-Time Excited Network for Anomalous Sound Detection
di: Zhang, Yucong, et al.
Pubblicazione: (2024)
di: Zhang, Yucong, et al.
Pubblicazione: (2024)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Cinematic Demixing Track
di: Uhlich, Stefan, et al.
Pubblicazione: (2023)
di: Uhlich, Stefan, et al.
Pubblicazione: (2023)
Technical Report of Nomi Team in the Environmental Sound Deepfake Detection Challenge 2026
di: Mawalim, Candy Olivia, et al.
Pubblicazione: (2025)
di: Mawalim, Candy Olivia, et al.
Pubblicazione: (2025)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Music Demixing Track
di: Fabbro, Giorgio, et al.
Pubblicazione: (2023)
di: Fabbro, Giorgio, et al.
Pubblicazione: (2023)
Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection
di: Liang, Jinhua, et al.
Pubblicazione: (2024)
di: Liang, Jinhua, et al.
Pubblicazione: (2024)
ILD-VIT: A Unified Vision Transformer Architecture for Detection of Interstitial Lung Disease from Respiratory Sounds
di: Hota, Soubhagya Ranjan, et al.
Pubblicazione: (2025)
di: Hota, Soubhagya Ranjan, et al.
Pubblicazione: (2025)
Semantic Communications for Speech Recognition
di: Weng, Zhenzi, et al.
Pubblicazione: (2021)
di: Weng, Zhenzi, et al.
Pubblicazione: (2021)
AudSemThinker: Enhancing Audio-Language Models through Reasoning over Semantics of Sound
di: Wijngaard, Gijs, et al.
Pubblicazione: (2025)
di: Wijngaard, Gijs, et al.
Pubblicazione: (2025)
SoundSculpt: Direction and Semantics Driven Ambisonic Target Sound Extraction
di: Chen, Tuochao, et al.
Pubblicazione: (2025)
di: Chen, Tuochao, et al.
Pubblicazione: (2025)
Adaptive Differential Denoising for Respiratory Sounds Classification
di: Dong, Gaoyang, et al.
Pubblicazione: (2025)
di: Dong, Gaoyang, et al.
Pubblicazione: (2025)
Audio Generation Through Score-Based Generative Modeling: Design Principles and Implementation
di: Zhu, Ge, et al.
Pubblicazione: (2025)
di: Zhu, Ge, et al.
Pubblicazione: (2025)
HSDreport: Heart Sound Diagnosis with Echocardiography Reports
di: Zhao, Zihan, et al.
Pubblicazione: (2024)
di: Zhao, Zihan, et al.
Pubblicazione: (2024)
ParaStyleTTS: Toward Efficient and Robust Paralinguistic Style Control for Expressive Text-to-Speech Generation
di: Lou, Haowei, et al.
Pubblicazione: (2025)
di: Lou, Haowei, et al.
Pubblicazione: (2025)
Self-Guided Target Sound Extraction and Classification Through Universal Sound Separation Model and Multiple Clues
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
di: Kwon, Younghoo, et al.
Pubblicazione: (2025)
Sound Field Estimation: Theories and Applications
di: Ueno, Natsuki, et al.
Pubblicazione: (2025)
di: Ueno, Natsuki, et al.
Pubblicazione: (2025)
The Database and Benchmark for the Source Speaker Tracing Challenge 2024
di: Li, Ze, et al.
Pubblicazione: (2024)
di: Li, Ze, et al.
Pubblicazione: (2024)
Location-Oriented Sound Event Localization and Detection with Spatial Mapping and Regression Localization
di: Zhang, Xueping, et al.
Pubblicazione: (2025)
di: Zhang, Xueping, et al.
Pubblicazione: (2025)
StreamAAD: Decoding Spatial Auditory Attention with a Streaming Architecture
di: Qiu, Zelin, et al.
Pubblicazione: (2024)
di: Qiu, Zelin, et al.
Pubblicazione: (2024)
NPU-NTU System for Voice Privacy 2024 Challenge
di: Yao, Jixun, et al.
Pubblicazione: (2024)
di: Yao, Jixun, et al.
Pubblicazione: (2024)
Sound event detection based on auxiliary decoder and maximum probability aggregation for DCASE Challenge 2024 Task 4
di: Son, Sang Won, et al.
Pubblicazione: (2024)
di: Son, Sang Won, et al.
Pubblicazione: (2024)
Sound Scene Synthesis at the DCASE 2024 Challenge
di: Lagrange, Mathieu, et al.
Pubblicazione: (2025)
di: Lagrange, Mathieu, et al.
Pubblicazione: (2025)
Text-Queried Target Sound Event Localization
di: Zhao, Jinzheng, et al.
Pubblicazione: (2024)
di: Zhao, Jinzheng, et al.
Pubblicazione: (2024)
Causal Spatio-Temporal Sound Field Reconstruction
di: Sundström, David, et al.
Pubblicazione: (2026)
di: Sundström, David, et al.
Pubblicazione: (2026)
A Generalist Audio Foundation Model for Comprehensive Body Sound Auscultation
di: Wang, Pingjie, et al.
Pubblicazione: (2024)
di: Wang, Pingjie, et al.
Pubblicazione: (2024)
Leveraging Sound Source Trajectories for Universal Sound Separation
di: Wu, Donghang, et al.
Pubblicazione: (2024)
di: Wu, Donghang, et al.
Pubblicazione: (2024)
Sound Zone Control Robust To Sound Speed Change
di: Bhattacharjee, Sankha Subhra, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Sankha Subhra, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Description and Discussion on DCASE 2026 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2026) -
Sound Separation and Classification with Object and Semantic Guidance
di: Kwon, Younghoo, et al.
Pubblicazione: (2025) -
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025) -
BUT Systems for Environmental Sound Deepfake Detection in the ESDD 2026 Challenge
di: Peng, Junyi, et al.
Pubblicazione: (2025) -
On-device Internet of Sounds Sonification with Wavetable Synthesis Techniques for Soil Moisture Monitoring in Water Scarcity Contexts
di: Roddy, Stephen
Pubblicazione: (2025)