Retrieval-Augmented Approach for Unsupervised Anomalous Sound Detection and Captioning without Model Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ogura, Ryoya, Nishida, Tomoya, Kawaguchi, Yohei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Timbre Difference Capturing in Anomalous Sound Detection
von: Nishida, Tomoya, et al.
Veröffentlicht: (2024)
von: Nishida, Tomoya, et al.
Veröffentlicht: (2024)
Retaining Mixture Representations for Domain Generalized Anomalous Sound Detection
von: Saengthong, Phurich, et al.
Veröffentlicht: (2025)
von: Saengthong, Phurich, et al.
Veröffentlicht: (2025)
MIMII-Gen: Generative Modeling Approach for Simulated Evaluation of Anomalous Sound Detection System
von: Purohit, Harsh, et al.
Veröffentlicht: (2024)
von: Purohit, Harsh, et al.
Veröffentlicht: (2024)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
von: Nishida, Tomoya, et al.
Veröffentlicht: (2026)
von: Nishida, Tomoya, et al.
Veröffentlicht: (2026)
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
von: Nishida, Tomoya, et al.
Veröffentlicht: (2025)
von: Nishida, Tomoya, et al.
Veröffentlicht: (2025)
MIMII-Agent: Leveraging LLMs with Function Calling for Relative Evaluation of Anomalous Sound Detection
von: Purohit, Harsh, et al.
Veröffentlicht: (2025)
von: Purohit, Harsh, et al.
Veröffentlicht: (2025)
Stream-based Active Learning for Anomalous Sound Detection in Machine Condition Monitoring
von: Ho, Tuan Vu, et al.
Veröffentlicht: (2024)
von: Ho, Tuan Vu, et al.
Veröffentlicht: (2024)
Description and Discussion on DCASE 2024 Challenge Task 2: First-Shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
von: Nishida, Tomoya, et al.
Veröffentlicht: (2024)
von: Nishida, Tomoya, et al.
Veröffentlicht: (2024)
End-to-End Integration of Speech Emotion Recognition with Voice Activity Detection using Self-Supervised Learning Features
von: Yamashita, Natsuo, et al.
Veröffentlicht: (2024)
von: Yamashita, Natsuo, et al.
Veröffentlicht: (2024)
Improvements of Discriminative Feature Space Training for Anomalous Sound Detection in Unlabeled Conditions
von: Fujimura, Takuya, et al.
Veröffentlicht: (2024)
von: Fujimura, Takuya, et al.
Veröffentlicht: (2024)
First-Shot Unsupervised Anomalous Sound Detection With Unknown Anomalies Estimated by Metadata-Assisted Audio Generation
von: Zhang, Hejing, et al.
Veröffentlicht: (2023)
von: Zhang, Hejing, et al.
Veröffentlicht: (2023)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Temporal Pooling Strategies for Training-Free Anomalous Sound Detection with Self-Supervised Audio Embeddings
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
Exploring Self-Supervised Audio Models for Generalized Anomalous Sound Detection
von: Han, Bing, et al.
Veröffentlicht: (2025)
von: Han, Bing, et al.
Veröffentlicht: (2025)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
ACES: Evaluating Automated Audio Captioning Models on the Semantics of Sounds
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2024)
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2024)
AnoPatch: Towards Better Consistency in Machine Anomalous Sound Detection
von: Jiang, Anbai, et al.
Veröffentlicht: (2024)
von: Jiang, Anbai, et al.
Veröffentlicht: (2024)
Disentangling Hierarchical Features for Anomalous Sound Detection Under Domain Shift
von: Guan, Jian, et al.
Veröffentlicht: (2025)
von: Guan, Jian, et al.
Veröffentlicht: (2025)
Construction and Analysis of Impression Caption Dataset for Environmental Sounds
von: Okamoto, Yuki, et al.
Veröffentlicht: (2024)
von: Okamoto, Yuki, et al.
Veröffentlicht: (2024)
How Much Does Machine Identity Matter in Anomalous Sound Detection at Test Time?
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
Handling Domain Shifts for Anomalous Sound Detection: A Review of DCASE-Related Work
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2025)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2025)
SPO-CLAPScore: Enhancing CLAP-based alignment prediction system with Standardize Preference Optimization, for the first XACLE Challenge
von: Takano, Taisei, et al.
Veröffentlicht: (2026)
von: Takano, Taisei, et al.
Veröffentlicht: (2026)
Mind the Gap: Detecting Cluster Exits for Robust Local Density-Based Score Normalization in Anomalous Sound Detection
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
ASD-Diffusion: Anomalous Sound Detection with Diffusion Models
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024)
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
Unsupervised Improved MVDR Beamforming for Sound Enhancement
von: Kealey, Jacob, et al.
Veröffentlicht: (2024)
von: Kealey, Jacob, et al.
Veröffentlicht: (2024)
AdaProj: Adaptively Scaled Angular Margin Subspace Projections for Anomalous Sound Detection with Auxiliary Classification Tasks
von: Wilkinghoff, Kevin
Veröffentlicht: (2024)
von: Wilkinghoff, Kevin
Veröffentlicht: (2024)
Effective Pre-Training of Audio Transformers for Sound Event Detection
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
Improving Anomalous Sound Detection via Low-Rank Adaptation Fine-Tuning of Pre-Trained Audio Models
von: Zheng, Xinhu, et al.
Veröffentlicht: (2024)
von: Zheng, Xinhu, et al.
Veröffentlicht: (2024)
Improving Anomalous Sound Detection through Pseudo-anomalous Set Selection and Pseudo-label Utilization under Unlabeled Conditions
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
Automatic Inspection Based on Switch Sounds of Electric Point Machines
von: Shibata, Ayano, et al.
Veröffentlicht: (2025)
von: Shibata, Ayano, et al.
Veröffentlicht: (2025)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
Improving Audio-Text Retrieval via Hierarchical Cross-Modal Interaction and Auxiliary Captions
von: Xin, Yifei, et al.
Veröffentlicht: (2023)
von: Xin, Yifei, et al.
Veröffentlicht: (2023)
SLAM-AAC: Enhancing Audio Captioning with Paraphrasing Augmentation and CLAP-Refine through LLMs
von: Chen, Wenxi, et al.
Veröffentlicht: (2024)
von: Chen, Wenxi, et al.
Veröffentlicht: (2024)
Large-scale Contrastive Language-Audio Pretraining with Feature Fusion and Keyword-to-Caption Augmentation
von: Wu, Yusong, et al.
Veröffentlicht: (2022)
von: Wu, Yusong, et al.
Veröffentlicht: (2022)
Sound-VECaps: Improving Audio Generation with Visual Enhanced Captions
von: Yuan, Yi, et al.
Veröffentlicht: (2024)
von: Yuan, Yi, et al.
Veröffentlicht: (2024)
Improving Audio Spectrogram Transformers for Sound Event Detection Through Multi-Stage Training
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
von: Schmid, Florian, et al.
Veröffentlicht: (2024)
Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervision, and LLM Mix-up Augmentation
von: Wu, Shih-Lun, et al.
Veröffentlicht: (2023)
von: Wu, Shih-Lun, et al.
Veröffentlicht: (2023)
Activity-Guided Industrial Anomalous Sound Detection against Interferences
von: Lee, Yunjoo, et al.
Veröffentlicht: (2024)
von: Lee, Yunjoo, et al.
Veröffentlicht: (2024)
Deep Generic Representations for Domain-Generalized Anomalous Sound Detection
von: Saengthong, Phurich, et al.
Veröffentlicht: (2024)
von: Saengthong, Phurich, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Timbre Difference Capturing in Anomalous Sound Detection
von: Nishida, Tomoya, et al.
Veröffentlicht: (2024) -
Retaining Mixture Representations for Domain Generalized Anomalous Sound Detection
von: Saengthong, Phurich, et al.
Veröffentlicht: (2025) -
MIMII-Gen: Generative Modeling Approach for Simulated Evaluation of Anomalous Sound Detection System
von: Purohit, Harsh, et al.
Veröffentlicht: (2024) -
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
von: Nishida, Tomoya, et al.
Veröffentlicht: (2026) -
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
von: Nishida, Tomoya, et al.
Veröffentlicht: (2025)