Sound Scene Synthesis at the DCASE 2024 Challenge
Fuente:
arXiv
Salvato in:
| Autori principali: | Lagrange, Mathieu, Lee, Junwon, Tailleur, Modan, Heller, Laurie M., Choi, Keunwoo, McFee, Brian, Imoto, Keisuke, Okamoto, Yuki |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Challenge on Sound Scene Synthesis: Evaluating Text-to-Audio Generation
di: Lee, Junwon, et al.
Pubblicazione: (2024)
di: Lee, Junwon, et al.
Pubblicazione: (2024)
Correlation of Fréchet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependant
di: Tailleur, Modan, et al.
Pubblicazione: (2024)
di: Tailleur, Modan, et al.
Pubblicazione: (2024)
Detection of Deepfake Environmental Audio
di: Ouajdi, Hafsa, et al.
Pubblicazione: (2024)
di: Ouajdi, Hafsa, et al.
Pubblicazione: (2024)
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
di: Tailleur, Modan, et al.
Pubblicazione: (2025)
di: Tailleur, Modan, et al.
Pubblicazione: (2025)
Towards better visualizations of urban sound environments: insights from interviews
di: Tailleur, Modan, et al.
Pubblicazione: (2024)
di: Tailleur, Modan, et al.
Pubblicazione: (2024)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
di: Imoto, Keisuke
Pubblicazione: (2025)
di: Imoto, Keisuke
Pubblicazione: (2025)
Construction and Analysis of Impression Caption Dataset for Environmental Sounds
di: Okamoto, Yuki, et al.
Pubblicazione: (2024)
di: Okamoto, Yuki, et al.
Pubblicazione: (2024)
Handling Domain Shifts for Anomalous Sound Detection: A Review of DCASE-Related Work
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2025)
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2025)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2026)
di: Nishida, Tomoya, et al.
Pubblicazione: (2026)
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2025)
di: Nishida, Tomoya, et al.
Pubblicazione: (2025)
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
di: Schmid, Florian, et al.
Pubblicazione: (2024)
di: Schmid, Florian, et al.
Pubblicazione: (2024)
Description and Discussion on DCASE 2024 Challenge Task 2: First-Shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
di: Nishida, Tomoya, et al.
Pubblicazione: (2024)
di: Nishida, Tomoya, et al.
Pubblicazione: (2024)
Description and Discussion on DCASE 2025 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
di: Yasuda, Masahiro, et al.
Pubblicazione: (2025)
Sound event detection based on auxiliary decoder and maximum probability aggregation for DCASE Challenge 2024 Task 4
di: Son, Sang Won, et al.
Pubblicazione: (2024)
di: Son, Sang Won, et al.
Pubblicazione: (2024)
Learning to Solve Inverse Problems for Perceptual Sound Matching
di: Han, Han, et al.
Pubblicazione: (2023)
di: Han, Han, et al.
Pubblicazione: (2023)
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
di: Koga, Naoki, et al.
Pubblicazione: (2024)
di: Koga, Naoki, et al.
Pubblicazione: (2024)
Low-Complexity Acoustic Scene Classification with Device Information in the DCASE 2025 Challenge
di: Schmid, Florian, et al.
Pubblicazione: (2025)
di: Schmid, Florian, et al.
Pubblicazione: (2025)
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
How Much Does Machine Identity Matter in Anomalous Sound Detection at Test Time?
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2026)
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2026)
Context-Aware Query Refinement for Target Sound Extraction: Handling Partially Matched Queries
di: Sato, Ryo, et al.
Pubblicazione: (2025)
di: Sato, Ryo, et al.
Pubblicazione: (2025)
Spatial Scaper: A Library to Simulate and Augment Soundscapes for Sound Event Localization and Detection in Realistic Rooms
di: Roman, Iran R., et al.
Pubblicazione: (2024)
di: Roman, Iran R., et al.
Pubblicazione: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
di: Xiao, Yang, et al.
Pubblicazione: (2024)
di: Xiao, Yang, et al.
Pubblicazione: (2024)
AudioBERTScore: Objective Evaluation of Environmental Sound Synthesis Based on Similarity of Audio embedding Sequences
di: Kishi, Minoru, et al.
Pubblicazione: (2025)
di: Kishi, Minoru, et al.
Pubblicazione: (2025)
Trainingless Adaptation of Pretrained Models for Environmental Sound Classification
di: Tonami, Noriyuki, et al.
Pubblicazione: (2024)
di: Tonami, Noriyuki, et al.
Pubblicazione: (2024)
BioDCASE 2026 Challenge Baseline for Cross-Domain Mosquito Species Classification
di: Hou, Yuanbo, et al.
Pubblicazione: (2026)
di: Hou, Yuanbo, et al.
Pubblicazione: (2026)
Hybrid Losses for Hierarchical Embedding Learning
di: Tian, Haokun, et al.
Pubblicazione: (2025)
di: Tian, Haokun, et al.
Pubblicazione: (2025)
Machine listening in a neonatal intensive care unit
di: Tailleur, Modan, et al.
Pubblicazione: (2024)
di: Tailleur, Modan, et al.
Pubblicazione: (2024)
A decade of DCASE: Achievements, practices, evaluations and future challenges
di: Mesaros, Annamaria, et al.
Pubblicazione: (2024)
di: Mesaros, Annamaria, et al.
Pubblicazione: (2024)
Vision Language Models Are Few-Shot Audio Spectrogram Classifiers
di: Dixit, Satvik, et al.
Pubblicazione: (2024)
di: Dixit, Satvik, et al.
Pubblicazione: (2024)
Discrete Tokens Exhibit Interlanguage Speech Intelligibility Benefit: an Analytical Study Towards Accent-robust ASR Only with Native Speech Data
di: Onda, Kentaro, et al.
Pubblicazione: (2025)
di: Onda, Kentaro, et al.
Pubblicazione: (2025)
Prosodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora
di: Onda, Kentaro, et al.
Pubblicazione: (2025)
di: Onda, Kentaro, et al.
Pubblicazione: (2025)
Description and analysis of novelties introduced in DCASE Task 4 2022 on the baseline system
di: Ronchini, Francesca, et al.
Pubblicazione: (2022)
di: Ronchini, Francesca, et al.
Pubblicazione: (2022)
TALKPLAY: Multimodal Music Recommendation with Large Language Models
di: Doh, Seungheon, et al.
Pubblicazione: (2025)
di: Doh, Seungheon, et al.
Pubblicazione: (2025)
KAD: No More FAD! An Effective and Efficient Evaluation Metric for Audio Generation
di: Chung, Yoonjin, et al.
Pubblicazione: (2025)
di: Chung, Yoonjin, et al.
Pubblicazione: (2025)
S-KEY: Self-supervised Learning of Major and Minor Keys from Audio
di: Kong, Yuexuan, et al.
Pubblicazione: (2025)
di: Kong, Yuexuan, et al.
Pubblicazione: (2025)
Neurobench: DCASE 2020 Acoustic Scene Classification benchmark on XyloAudio 2
di: Ke, Weijie, et al.
Pubblicazione: (2024)
di: Ke, Weijie, et al.
Pubblicazione: (2024)
SONAR: Self-Distilled Continual Pre-training for Domain Adaptive Audio Representation
di: Zhang, Yizhou, et al.
Pubblicazione: (2025)
di: Zhang, Yizhou, et al.
Pubblicazione: (2025)
Refining Knowledge Transfer on Audio-Image Temporal Agreement for Audio-Text Cross Retrieval
di: Tsubaki, Shunsuke, et al.
Pubblicazione: (2024)
di: Tsubaki, Shunsuke, et al.
Pubblicazione: (2024)
AISTAT lab system for DCASE2025 Task6: Language-based audio retrieval
di: Kim, Hyun Jun, et al.
Pubblicazione: (2025)
di: Kim, Hyun Jun, et al.
Pubblicazione: (2025)
SoundCompass: Navigating Target Sound Extraction With Effective Directional Clue Integration In Complex Acoustic Scenes
di: Choi, Dayun, et al.
Pubblicazione: (2025)
di: Choi, Dayun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Challenge on Sound Scene Synthesis: Evaluating Text-to-Audio Generation
di: Lee, Junwon, et al.
Pubblicazione: (2024) -
Correlation of Fréchet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependant
di: Tailleur, Modan, et al.
Pubblicazione: (2024) -
Detection of Deepfake Environmental Audio
di: Ouajdi, Hafsa, et al.
Pubblicazione: (2024) -
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
di: Tailleur, Modan, et al.
Pubblicazione: (2025) -
Towards better visualizations of urban sound environments: insights from interviews
di: Tailleur, Modan, et al.
Pubblicazione: (2024)