Trainingless Adaptation of Pretrained Models for Environmental Sound Classification
Fuente:
arXiv
Guardado en:
| Autores principales: | Tonami, Noriyuki, Kohno, Wataru, Imoto, Keisuke, Yajima, Yoshiyuki, Mishima, Sakiko, Kondo, Reishi, Hino, Tomoyuki |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Event Classification by Physics-informed Inpainting for Distributed Multichannel Acoustic Sensor with Partially Degraded Channels
por: Tonami, Noriyuki, et al.
Publicado: (2026)
por: Tonami, Noriyuki, et al.
Publicado: (2026)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
por: Imoto, Keisuke
Publicado: (2025)
por: Imoto, Keisuke
Publicado: (2025)
Construction and Analysis of Impression Caption Dataset for Environmental Sounds
por: Okamoto, Yuki, et al.
Publicado: (2024)
por: Okamoto, Yuki, et al.
Publicado: (2024)
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
por: Koga, Naoki, et al.
Publicado: (2024)
por: Koga, Naoki, et al.
Publicado: (2024)
How Much Does Machine Identity Matter in Anomalous Sound Detection at Test Time?
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
Context-Aware Query Refinement for Target Sound Extraction: Handling Partially Matched Queries
por: Sato, Ryo, et al.
Publicado: (2025)
por: Sato, Ryo, et al.
Publicado: (2025)
Handling Domain Shifts for Anomalous Sound Detection: A Review of DCASE-Related Work
por: Wilkinghoff, Kevin, et al.
Publicado: (2025)
por: Wilkinghoff, Kevin, et al.
Publicado: (2025)
Correlation of Fréchet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependant
por: Tailleur, Modan, et al.
Publicado: (2024)
por: Tailleur, Modan, et al.
Publicado: (2024)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
por: Nishida, Tomoya, et al.
Publicado: (2026)
por: Nishida, Tomoya, et al.
Publicado: (2026)
Discrete Tokens Exhibit Interlanguage Speech Intelligibility Benefit: an Analytical Study Towards Accent-robust ASR Only with Native Speech Data
por: Onda, Kentaro, et al.
Publicado: (2025)
por: Onda, Kentaro, et al.
Publicado: (2025)
Prosodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora
por: Onda, Kentaro, et al.
Publicado: (2025)
por: Onda, Kentaro, et al.
Publicado: (2025)
AFT: An Exemplar-Free Class Incremental Learning Method for Environmental Sound Classification
por: Chen, Xinyi, et al.
Publicado: (2025)
por: Chen, Xinyi, et al.
Publicado: (2025)
SONAR: Self-Distilled Continual Pre-training for Domain Adaptive Audio Representation
por: Zhang, Yizhou, et al.
Publicado: (2025)
por: Zhang, Yizhou, et al.
Publicado: (2025)
Refining Knowledge Transfer on Audio-Image Temporal Agreement for Audio-Text Cross Retrieval
por: Tsubaki, Shunsuke, et al.
Publicado: (2024)
por: Tsubaki, Shunsuke, et al.
Publicado: (2024)
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
por: Nishida, Tomoya, et al.
Publicado: (2025)
por: Nishida, Tomoya, et al.
Publicado: (2025)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
por: Gao, Wenmiao, et al.
Publicado: (2025)
por: Gao, Wenmiao, et al.
Publicado: (2025)
Sound Scene Synthesis at the DCASE 2024 Challenge
por: Lagrange, Mathieu, et al.
Publicado: (2025)
por: Lagrange, Mathieu, et al.
Publicado: (2025)
Adaptive Differential Denoising for Respiratory Sounds Classification
por: Dong, Gaoyang, et al.
Publicado: (2025)
por: Dong, Gaoyang, et al.
Publicado: (2025)
Studying the Effect of Audio Filters in Pre-Trained Models for Environmental Sound Classification
por: Dawn, Aditya, et al.
Publicado: (2024)
por: Dawn, Aditya, et al.
Publicado: (2024)
Evaluating CNN with Stacked Feature Representations and Audio Spectrogram Transformer Models for Sound Classification
por: Dehaghania, Parinaz Binandeh, et al.
Publicado: (2026)
por: Dehaghania, Parinaz Binandeh, et al.
Publicado: (2026)
Infrastructure-less Localization from Indoor Environmental Sounds Based on Spectral Decomposition and Spatial Likelihood Model
por: Ogiso, Satoki, et al.
Publicado: (2024)
por: Ogiso, Satoki, et al.
Publicado: (2024)
Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification
por: Sims, Ysobel, et al.
Publicado: (2024)
por: Sims, Ysobel, et al.
Publicado: (2024)
Disentangling Dual-Encoder Masked Autoencoder for Respiratory Sound Classification
por: Wei, Peidong, et al.
Publicado: (2025)
por: Wei, Peidong, et al.
Publicado: (2025)
Technical Report of Nomi Team in the Environmental Sound Deepfake Detection Challenge 2026
por: Mawalim, Candy Olivia, et al.
Publicado: (2025)
por: Mawalim, Candy Olivia, et al.
Publicado: (2025)
Enhancing Zero-shot Audio Classification using Sound Attribute Knowledge from Large Language Models
por: Xu, Xuenan, et al.
Publicado: (2024)
por: Xu, Xuenan, et al.
Publicado: (2024)
Domain Adaptation Method and Modality Gap Impact in Audio-Text Models for Prototypical Sound Classification
por: Acevedo, Emiliano, et al.
Publicado: (2025)
por: Acevedo, Emiliano, et al.
Publicado: (2025)
AudioBERTScore: Objective Evaluation of Environmental Sound Synthesis Based on Similarity of Audio embedding Sequences
por: Kishi, Minoru, et al.
Publicado: (2025)
por: Kishi, Minoru, et al.
Publicado: (2025)
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
por: Tailleur, Modan, et al.
Publicado: (2025)
por: Tailleur, Modan, et al.
Publicado: (2025)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
por: Lee, Dongheon, et al.
Publicado: (2024)
por: Lee, Dongheon, et al.
Publicado: (2024)
MUDAS: Mote-scale Unsupervised Domain Adaptation in Multi-label Sound Classification
por: Yun, Jihoon, et al.
Publicado: (2025)
por: Yun, Jihoon, et al.
Publicado: (2025)
Automatic Sound Event Detection and Classification of Great Ape Calls Using Neural Networks
por: Jiang, Zifan, et al.
Publicado: (2023)
por: Jiang, Zifan, et al.
Publicado: (2023)
Sound-Based Spin Estimation in Table Tennis: Dataset and Real-Time Classification Pipeline
por: Gossard, Thomas, et al.
Publicado: (2024)
por: Gossard, Thomas, et al.
Publicado: (2024)
DOA-Aware Audio-Visual Self-Supervised Learning for Sound Event Localization and Detection
por: Fujita, Yoto, et al.
Publicado: (2024)
por: Fujita, Yoto, et al.
Publicado: (2024)
Active Learning of Non-semantic Speech Tasks with Pretrained Models
por: Lee, Harlin, et al.
Publicado: (2022)
por: Lee, Harlin, et al.
Publicado: (2022)
Improving Controllability and Editability for Pretrained Text-to-Music Generation Models
por: Zhang, Yixiao
Publicado: (2024)
por: Zhang, Yixiao
Publicado: (2024)
LatentVoiceGrad: Nonparallel Voice Conversion with Latent Diffusion/Flow-Matching Models
por: Kameoka, Hirokazu, et al.
Publicado: (2025)
por: Kameoka, Hirokazu, et al.
Publicado: (2025)
SoundBeam meets M2D: Target Sound Extraction with Audio Foundation Model
por: Hernandez-Olivan, Carlos, et al.
Publicado: (2024)
por: Hernandez-Olivan, Carlos, et al.
Publicado: (2024)
Description and Discussion on DCASE 2024 Challenge Task 2: First-Shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
por: Nishida, Tomoya, et al.
Publicado: (2024)
por: Nishida, Tomoya, et al.
Publicado: (2024)
Enhancing Speaker-Independent Dysarthric Speech Severity Classification with DSSCNet and Cross-Corpus Adaptation
por: Roy, Arnab Kumar, et al.
Publicado: (2025)
por: Roy, Arnab Kumar, et al.
Publicado: (2025)
SoundLoCD: An Efficient Conditional Discrete Contrastive Latent Diffusion Model for Text-to-Sound Generation
por: Niu, Xinlei, et al.
Publicado: (2024)
por: Niu, Xinlei, et al.
Publicado: (2024)
Ejemplares similares
-
Event Classification by Physics-informed Inpainting for Distributed Multichannel Acoustic Sensor with Partially Degraded Channels
por: Tonami, Noriyuki, et al.
Publicado: (2026) -
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
por: Imoto, Keisuke
Publicado: (2025) -
Construction and Analysis of Impression Caption Dataset for Environmental Sounds
por: Okamoto, Yuki, et al.
Publicado: (2024) -
LEAD Dataset: How Can Labels for Sound Event Detection Vary Depending on Annotators?
por: Koga, Naoki, et al.
Publicado: (2024) -
How Much Does Machine Identity Matter in Anomalous Sound Detection at Test Time?
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)