Heterogeneous sound classification with the Broad Sound Taxonomy and Dataset
Fuente:
arXiv
Salvato in:
| Autori principali: | Anastasopoulou, Panagiota, Torrey, Jessica, Serra, Xavier, Font, Frederic |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Statistics-Driven Differentiable Approach for Sound Texture Synthesis and Analysis
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025)
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025)
Fractional Fourier Sound Synthesis
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025)
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025)
Multi-Speaker Conversational Audio Deepfake: Taxonomy, Dataset and Pilot Study
di: Ahmed, Alabi, et al.
Pubblicazione: (2026)
di: Ahmed, Alabi, et al.
Pubblicazione: (2026)
A sound description: Exploring prompt templates and class descriptions to enhance zero-shot audio classification
di: Olvera, Michel, et al.
Pubblicazione: (2024)
di: Olvera, Michel, et al.
Pubblicazione: (2024)
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
di: Bibbó, Gabriel, et al.
Pubblicazione: (2024)
di: Bibbó, Gabriel, et al.
Pubblicazione: (2024)
CycleGuardian: A Framework for Automatic RespiratorySound classification Based on Improved Deep clustering and Contrastive Learning
di: Chu, Yun, et al.
Pubblicazione: (2025)
di: Chu, Yun, et al.
Pubblicazione: (2025)
BTS: Bridging Text and Sound Modalities for Metadata-Aided Respiratory Sound Classification
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
Detect Any Sound: Open-Vocabulary Sound Event Detection with Multi-Modal Queries
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
Detecting abnormal heart sound using mobile phones and on-device IConNet
di: Vu, Linh, et al.
Pubblicazione: (2024)
di: Vu, Linh, et al.
Pubblicazione: (2024)
Epic-Sounds: A Large-scale Dataset of Actions That Sound
di: Huh, Jaesung, et al.
Pubblicazione: (2023)
di: Huh, Jaesung, et al.
Pubblicazione: (2023)
SoundCompass: Navigating Target Sound Extraction With Effective Directional Clue Integration In Complex Acoustic Scenes
di: Choi, Dayun, et al.
Pubblicazione: (2025)
di: Choi, Dayun, et al.
Pubblicazione: (2025)
Sound Classification of Four Insect Classes
di: Wang, Yinxuan, et al.
Pubblicazione: (2024)
di: Wang, Yinxuan, et al.
Pubblicazione: (2024)
Sound Check: Auditing Audio Datasets
di: Agnew, William, et al.
Pubblicazione: (2024)
di: Agnew, William, et al.
Pubblicazione: (2024)
Joint Learning of Emotions in Music and Generalized Sounds
di: Simonetta, Federico, et al.
Pubblicazione: (2024)
di: Simonetta, Federico, et al.
Pubblicazione: (2024)
A Holistic Evaluation of Piano Sound Quality
di: Zhou, Monan, et al.
Pubblicazione: (2023)
di: Zhou, Monan, et al.
Pubblicazione: (2023)
Auditory Intelligence: Understanding the World Through Sound
di: Nam, Hyeonuk
Pubblicazione: (2025)
di: Nam, Hyeonuk
Pubblicazione: (2025)
Sound Scene Synthesis at the DCASE 2024 Challenge
di: Lagrange, Mathieu, et al.
Pubblicazione: (2025)
di: Lagrange, Mathieu, et al.
Pubblicazione: (2025)
Abstract Sound Fusion with Unconditional Inversion Models
di: Liu, Jing, et al.
Pubblicazione: (2025)
di: Liu, Jing, et al.
Pubblicazione: (2025)
SonicSim: A customizable simulation platform for speech processing in moving sound source scenarios
di: Li, Kai, et al.
Pubblicazione: (2024)
di: Li, Kai, et al.
Pubblicazione: (2024)
Investigation into respiratory sound classification for an imbalanced data set using hybrid LSTM-KAN architectures
di: K. V, Nithinkumar, et al.
Pubblicazione: (2026)
di: K. V, Nithinkumar, et al.
Pubblicazione: (2026)
EvMic: Event-based Non-contact sound recovery from effective spatial-temporal modeling
di: Yin, Hao, et al.
Pubblicazione: (2025)
di: Yin, Hao, et al.
Pubblicazione: (2025)
Heart Sound Segmentation Using Deep Learning Techniques
di: Madine, Manas
Pubblicazione: (2024)
di: Madine, Manas
Pubblicazione: (2024)
ASD-Diffusion: Anomalous Sound Detection with Diffusion Models
di: Zhang, Fengrun, et al.
Pubblicazione: (2024)
di: Zhang, Fengrun, et al.
Pubblicazione: (2024)
EnvSDD: Benchmarking Environmental Sound Deepfake Detection
di: Yin, Han, et al.
Pubblicazione: (2025)
di: Yin, Han, et al.
Pubblicazione: (2025)
Leveraging Language Model Capabilities for Sound Event Detection
di: Wang, Hualei, et al.
Pubblicazione: (2023)
di: Wang, Hualei, et al.
Pubblicazione: (2023)
EZhouNet:A framework based on graph neural network and anchor interval for the respiratory sound event detection
di: Chu, Yun, et al.
Pubblicazione: (2025)
di: Chu, Yun, et al.
Pubblicazione: (2025)
DiffMoog: a Differentiable Modular Synthesizer for Sound Matching
di: Uzrad, Noy, et al.
Pubblicazione: (2024)
di: Uzrad, Noy, et al.
Pubblicazione: (2024)
Deep Generic Representations for Domain-Generalized Anomalous Sound Detection
di: Saengthong, Phurich, et al.
Pubblicazione: (2024)
di: Saengthong, Phurich, et al.
Pubblicazione: (2024)
Universal Sound Separation with Self-Supervised Audio Masked Autoencoder
di: Zhao, Junqi, et al.
Pubblicazione: (2024)
di: Zhao, Junqi, et al.
Pubblicazione: (2024)
Contrastive Learning with Spectrum Information Augmentation in Abnormal Sound Detection
di: Meng, Xinxin, et al.
Pubblicazione: (2025)
di: Meng, Xinxin, et al.
Pubblicazione: (2025)
FlexSED: Towards Open-Vocabulary Sound Event Detection
di: Hai, Jiarui, et al.
Pubblicazione: (2025)
di: Hai, Jiarui, et al.
Pubblicazione: (2025)
IndieFake Dataset: A Benchmark Dataset for Audio Deepfake Detection
di: Kumar, Abhay, et al.
Pubblicazione: (2025)
di: Kumar, Abhay, et al.
Pubblicazione: (2025)
Multichannel-to-Multichannel Target Sound Extraction Using Direction and Timestamp Clues
di: Choi, Dayun, et al.
Pubblicazione: (2024)
di: Choi, Dayun, et al.
Pubblicazione: (2024)
Towards Assessing Data Replication in Music Generation with Music Similarity Metrics on Raw Audio
di: Batlle-Roca, Roser, et al.
Pubblicazione: (2024)
di: Batlle-Roca, Roser, et al.
Pubblicazione: (2024)
PC-MCL: Patient-Consistent Multi-Cycle Learning with multi-label bias correction for respiratory sound classification
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2026)
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2026)
Studying the Effect of Audio Filters in Pre-Trained Models for Environmental Sound Classification
di: Dawn, Aditya, et al.
Pubblicazione: (2024)
di: Dawn, Aditya, et al.
Pubblicazione: (2024)
RepAugment: Input-Agnostic Representation-Level Augmentation for Respiratory Sound Classification
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
di: Kim, June-Woo, et al.
Pubblicazione: (2024)
Heterogeneous bimodal attention fusion for speech emotion recognition
di: Luo, Jiachen, et al.
Pubblicazione: (2025)
di: Luo, Jiachen, et al.
Pubblicazione: (2025)
Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles
di: Toikkanen, Miika, et al.
Pubblicazione: (2025)
di: Toikkanen, Miika, et al.
Pubblicazione: (2025)
Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis
di: Doerfler, Robin, et al.
Pubblicazione: (2026)
di: Doerfler, Robin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Statistics-Driven Differentiable Approach for Sound Texture Synthesis and Analysis
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025) -
Fractional Fourier Sound Synthesis
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025) -
Multi-Speaker Conversational Audio Deepfake: Taxonomy, Dataset and Pilot Study
di: Ahmed, Alabi, et al.
Pubblicazione: (2026) -
A sound description: Exploring prompt templates and class descriptions to enhance zero-shot audio classification
di: Olvera, Michel, et al.
Pubblicazione: (2024) -
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
di: Bibbó, Gabriel, et al.
Pubblicazione: (2024)