Description and analysis of novelties introduced in DCASE Task 4 2022 on the baseline system
Fuente:
arXiv
Salvato in:
| Autori principali: | Ronchini, Francesca, Cornell, Samuele, Serizel, Romain, Turpault, Nicolas, Fonseca, Eduardo, Ellis, Daniel P. W. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The impact of non-target events in synthetic soundscapes for sound event detection
di: Ronchini, Francesca, et al.
Pubblicazione: (2021)
di: Ronchini, Francesca, et al.
Pubblicazione: (2021)
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
di: Ronchini, Francesca, et al.
Pubblicazione: (2023)
di: Ronchini, Francesca, et al.
Pubblicazione: (2023)
A benchmark of state-of-the-art sound event detection systems evaluated on synthetic soundscapes
di: Ronchini, Francesca, et al.
Pubblicazione: (2022)
di: Ronchini, Francesca, et al.
Pubblicazione: (2022)
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
A decade of DCASE: Achievements, practices, evaluations and future challenges
di: Mesaros, Annamaria, et al.
Pubblicazione: (2024)
di: Mesaros, Annamaria, et al.
Pubblicazione: (2024)
Diffused Responsibility: Analyzing the Energy Consumption of Generative Text-to-Audio Diffusion Models
di: Passoni, Riccardo, et al.
Pubblicazione: (2025)
di: Passoni, Riccardo, et al.
Pubblicazione: (2025)
Domain-Invariant Representation Learning of Bird Sounds
di: Moummad, Ilyass, et al.
Pubblicazione: (2024)
di: Moummad, Ilyass, et al.
Pubblicazione: (2024)
Angular Distance Distribution Loss for Audio Classification
di: Almudévar, Antonio, et al.
Pubblicazione: (2024)
di: Almudévar, Antonio, et al.
Pubblicazione: (2024)
Cross-Talk Speech Reduction, by Separation, for Separation
di: Wang, Zhong-Qiu, et al.
Pubblicazione: (2026)
di: Wang, Zhong-Qiu, et al.
Pubblicazione: (2026)
Self-Supervised Learning for Few-Shot Bird Sound Classification
di: Moummad, Ilyass, et al.
Pubblicazione: (2023)
di: Moummad, Ilyass, et al.
Pubblicazione: (2023)
Regularized Contrastive Pre-training for Few-shot Bioacoustic Sound Detection
di: Moummad, Ilyass, et al.
Pubblicazione: (2023)
di: Moummad, Ilyass, et al.
Pubblicazione: (2023)
A Phoneme-Scale Assessment of Multichannel Speech Enhancement Algorithms
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2024)
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2024)
Evaluating Multichannel Speech Enhancement Algorithms at the Phoneme Scale Across Genders
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2025)
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2025)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2026)
di: Nishida, Tomoya, et al.
Pubblicazione: (2026)
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2025)
di: Nishida, Tomoya, et al.
Pubblicazione: (2025)
Energy Consumption Trends in Sound Event Detection Systems
di: Douwes, Constance, et al.
Pubblicazione: (2024)
di: Douwes, Constance, et al.
Pubblicazione: (2024)
MAPSS: Manifold-based Assessment of Perceptual Source Separation
di: Ivry, Amir, et al.
Pubblicazione: (2025)
di: Ivry, Amir, et al.
Pubblicazione: (2025)
Sound event localization and detection based on crnn using rectangular filters and channel rotation data augmentation
di: Ronchini, Francesca, et al.
Pubblicazione: (2020)
di: Ronchini, Francesca, et al.
Pubblicazione: (2020)
Sound event detection based on auxiliary decoder and maximum probability aggregation for DCASE Challenge 2024 Task 4
di: Son, Sang Won, et al.
Pubblicazione: (2024)
di: Son, Sang Won, et al.
Pubblicazione: (2024)
Diffusion-based Generative Modeling with Discriminative Guidance for Streamable Speech Enhancement
di: Li, Chenda, et al.
Pubblicazione: (2024)
di: Li, Chenda, et al.
Pubblicazione: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
di: Xiao, Yang, et al.
Pubblicazione: (2024)
di: Xiao, Yang, et al.
Pubblicazione: (2024)
Tracking of Intermittent and Moving Speakers : Dataset and Metrics
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
AISTAT lab system for DCASE2025 Task6: Language-based audio retrieval
di: Kim, Hyun Jun, et al.
Pubblicazione: (2025)
di: Kim, Hyun Jun, et al.
Pubblicazione: (2025)
Latent Watermarking of Audio Generative Models
di: Roman, Robin San, et al.
Pubblicazione: (2024)
di: Roman, Robin San, et al.
Pubblicazione: (2024)
Description and Discussion on DCASE 2026 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2026)
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2026)
Mixture of Mixups for Multi-label Classification of Rare Anuran Sounds
di: Moummad, Ilyass, et al.
Pubblicazione: (2024)
di: Moummad, Ilyass, et al.
Pubblicazione: (2024)
PAGURI: a user experience study of creative interaction with text-to-music models
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Frequency-Weighted Training Losses for Phoneme-Level DNN-based Speech Enhancement
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2025)
di: Monir, Nasser-Eddine, et al.
Pubblicazione: (2025)
Description and Discussion on DCASE 2024 Challenge Task 2: First-Shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2024)
di: Nishida, Tomoya, et al.
Pubblicazione: (2024)
MambaFoley: Foley Sound Generation using Selective State-Space Models
di: Colombo, Marco Furio, et al.
Pubblicazione: (2024)
di: Colombo, Marco Furio, et al.
Pubblicazione: (2024)
Low-Complexity Acoustic Scene Classification with Device Information in the DCASE 2025 Challenge
di: Schmid, Florian, et al.
Pubblicazione: (2025)
di: Schmid, Florian, et al.
Pubblicazione: (2025)
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
di: Schmid, Florian, et al.
Pubblicazione: (2024)
di: Schmid, Florian, et al.
Pubblicazione: (2024)
BioDCASE 2026 Challenge Baseline for Cross-Domain Mosquito Species Classification
di: Hou, Yuanbo, et al.
Pubblicazione: (2026)
di: Hou, Yuanbo, et al.
Pubblicazione: (2026)
Speaker Embeddings to Improve Tracking of Intermittent and Moving Speakers
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
Handling Domain Shifts for Anomalous Sound Detection: A Review of DCASE-Related Work
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2025)
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2025)
Synthetic training set generation using text-to-audio models for environmental sound classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Generating Data with Text-to-Speech and Large-Language Models for Conversational Speech Recognition
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
Sound Scene Synthesis at the DCASE 2024 Challenge
di: Lagrange, Mathieu, et al.
Pubblicazione: (2025)
di: Lagrange, Mathieu, et al.
Pubblicazione: (2025)
Towards Low-Latency Tracking of Multiple Speakers With Short-Context Speaker Embeddings
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The impact of non-target events in synthetic soundscapes for sound event detection
di: Ronchini, Francesca, et al.
Pubblicazione: (2021) -
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
di: Cornell, Samuele, et al.
Pubblicazione: (2024) -
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
di: Ronchini, Francesca, et al.
Pubblicazione: (2023) -
A benchmark of state-of-the-art sound event detection systems evaluated on synthetic soundscapes
di: Ronchini, Francesca, et al.
Pubblicazione: (2022) -
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)