Machine Anomalous Sound Detection Using Spectral-temporal Modulation Representations Derived from Machine-specific Filterbanks
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Kai, Zaman, Khalid, Li, Xingfeng, Akagi, Masato, Unoki, Masashi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Deepfake Audio Detection Using Self-supervised Fusion Representations
di: Zaman, Khalid, et al.
Pubblicazione: (2026)
di: Zaman, Khalid, et al.
Pubblicazione: (2026)
Spectro-Temporal Modulation Representation Framework for Human-Imitated Speech Detection
di: Zaman, Khalid, et al.
Pubblicazione: (2026)
di: Zaman, Khalid, et al.
Pubblicazione: (2026)
Noise-Aware In-Context Learning for Hallucination Mitigation in ALLMs
di: Huang, Qixuan, et al.
Pubblicazione: (2026)
di: Huang, Qixuan, et al.
Pubblicazione: (2026)
Sub-Band Spectral Matching with Localized Score Aggregation for Robust Anomalous Sound Detection
di: Saengthong, Phurich, et al.
Pubblicazione: (2026)
di: Saengthong, Phurich, et al.
Pubblicazione: (2026)
Improving Anomalous Sound Detection with Attribute-aware Representation from Domain-adaptive Pre-training
di: Fang, Xin, et al.
Pubblicazione: (2025)
di: Fang, Xin, et al.
Pubblicazione: (2025)
ASD-Diffusion: Anomalous Sound Detection with Diffusion Models
di: Zhang, Fengrun, et al.
Pubblicazione: (2024)
di: Zhang, Fengrun, et al.
Pubblicazione: (2024)
Deep Generic Representations for Domain-Generalized Anomalous Sound Detection
di: Saengthong, Phurich, et al.
Pubblicazione: (2024)
di: Saengthong, Phurich, et al.
Pubblicazione: (2024)
CoopASD: Cooperative Machine Anomalous Sound Detection with Privacy Concerns
di: Jiang, Anbai, et al.
Pubblicazione: (2024)
di: Jiang, Anbai, et al.
Pubblicazione: (2024)
Multi-scale Scanning Network for Machine Anomalous Sound Detection
di: Zhang, Yucong, et al.
Pubblicazione: (2025)
di: Zhang, Yucong, et al.
Pubblicazione: (2025)
Infant Cry Detection Using Causal Temporal Representation
di: Fu, Minghao, et al.
Pubblicazione: (2025)
di: Fu, Minghao, et al.
Pubblicazione: (2025)
Environmental Sound Deepfake Detection Using Deep-Learning Framework
di: Pham, Lam, et al.
Pubblicazione: (2026)
di: Pham, Lam, et al.
Pubblicazione: (2026)
Fitting Auditory Filterbanks with Multiresolution Neural Networks
di: Lostanlen, Vincent, et al.
Pubblicazione: (2023)
di: Lostanlen, Vincent, et al.
Pubblicazione: (2023)
MIMII-Gen: Generative Modeling Approach for Simulated Evaluation of Anomalous Sound Detection System
di: Purohit, Harsh, et al.
Pubblicazione: (2024)
di: Purohit, Harsh, et al.
Pubblicazione: (2024)
TLDiffGAN: A Latent Diffusion-GAN Framework with Temporal Information Fusion for Anomalous Sound Detection
di: Ma, Chengyuan, et al.
Pubblicazione: (2026)
di: Ma, Chengyuan, et al.
Pubblicazione: (2026)
Towards Open World Sound Event Detection
di: Hai, P. H., et al.
Pubblicazione: (2026)
di: Hai, P. H., et al.
Pubblicazione: (2026)
Improving Anomalous Sound Detection via Low-Rank Adaptation Fine-Tuning of Pre-Trained Audio Models
di: Zheng, Xinhu, et al.
Pubblicazione: (2024)
di: Zheng, Xinhu, et al.
Pubblicazione: (2024)
Human or Machine? A Preliminary Turing Test for Speech-to-Speech Interaction
di: Li, Xiang, et al.
Pubblicazione: (2026)
di: Li, Xiang, et al.
Pubblicazione: (2026)
NSTR: Neural Spectral Transport Representation for Space-Varying Frequency Fields
di: Versace, Plein
Pubblicazione: (2025)
di: Versace, Plein
Pubblicazione: (2025)
MIMII-Agent: Leveraging LLMs with Function Calling for Relative Evaluation of Anomalous Sound Detection
di: Purohit, Harsh, et al.
Pubblicazione: (2025)
di: Purohit, Harsh, et al.
Pubblicazione: (2025)
'Studies for': A Human-AI Co-Creative Sound Artwork Using a Real-time Multi-channel Sound Generation Model
di: Nagashima, Chihiro, et al.
Pubblicazione: (2025)
di: Nagashima, Chihiro, et al.
Pubblicazione: (2025)
SONAR: Spectral-Contrastive Audio Residuals for Generalizable Deepfake Detection
di: HIdekel, Ido Nitzan, et al.
Pubblicazione: (2025)
di: HIdekel, Ido Nitzan, et al.
Pubblicazione: (2025)
Formula-Supervised Sound Event Detection: Pre-Training Without Real Data
di: Shibata, Yuto, et al.
Pubblicazione: (2025)
di: Shibata, Yuto, et al.
Pubblicazione: (2025)
Exploring Machine Learning and Language Models for Multimodal Depression Detection
di: Hong, Javier Si Zhao, et al.
Pubblicazione: (2025)
di: Hong, Javier Si Zhao, et al.
Pubblicazione: (2025)
Zero-Shot to Zero-Lies: Detecting Bengali Deepfake Audio through Transfer Learning
di: Samu, Most. Sharmin Sultana, et al.
Pubblicazione: (2025)
di: Samu, Most. Sharmin Sultana, et al.
Pubblicazione: (2025)
TopSeg: A Multi-Scale Topological Framework for Data-Efficient Heart Sound Segmentation
di: Zhang, Peihong, et al.
Pubblicazione: (2025)
di: Zhang, Peihong, et al.
Pubblicazione: (2025)
A Machine Learning Approach for MIDI to Guitar Tablature Conversion
di: Kaliakatsos-Papakostas, Maximos, et al.
Pubblicazione: (2025)
di: Kaliakatsos-Papakostas, Maximos, et al.
Pubblicazione: (2025)
Melody or Machine: Detecting Synthetic Music with Dual-Stream Contrastive Learning
di: Batra, Arnesh, et al.
Pubblicazione: (2025)
di: Batra, Arnesh, et al.
Pubblicazione: (2025)
Pediatric Asthma Detection with Googles HeAR Model: An AI-Driven Respiratory Sound Classifier
di: Ehtesham, Abul, et al.
Pubblicazione: (2025)
di: Ehtesham, Abul, et al.
Pubblicazione: (2025)
Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations
di: Han, Yichen, et al.
Pubblicazione: (2025)
di: Han, Yichen, et al.
Pubblicazione: (2025)
Detect Any Sound: Open-Vocabulary Sound Event Detection with Multi-Modal Queries
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine
di: Kuznetsova, Anastasia, et al.
Pubblicazione: (2025)
di: Kuznetsova, Anastasia, et al.
Pubblicazione: (2025)
MFF-EINV2: Multi-scale Feature Fusion across Spectral-Spatial-Temporal Domains for Sound Event Localization and Detection
di: Mu, Da, et al.
Pubblicazione: (2024)
di: Mu, Da, et al.
Pubblicazione: (2024)
Teaching LLMs Music Theory with In-Context Learning and Chain-of-Thought Prompting: Pedagogical Strategies for Machines
di: Pond, Liam, et al.
Pubblicazione: (2025)
di: Pond, Liam, et al.
Pubblicazione: (2025)
Semi-Supervised Diseased Detection from Speech Dialogues with Multi-Level Data Modeling
di: Li, Xingyuan, et al.
Pubblicazione: (2026)
di: Li, Xingyuan, et al.
Pubblicazione: (2026)
Linguistic and Audio Embedding-Based Machine Learning for Alzheimer's Dementia and Mild Cognitive Impairment Detection: Insights from the PROCESS Challenge
di: Devahi, Adharsha Sam Edwin Sam, et al.
Pubblicazione: (2025)
di: Devahi, Adharsha Sam Edwin Sam, et al.
Pubblicazione: (2025)
MARS-Sep: Multimodal-Aligned Reinforced Sound Separation
di: Zhang, Zihan, et al.
Pubblicazione: (2025)
di: Zhang, Zihan, et al.
Pubblicazione: (2025)
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
di: Bibbó, Gabriel, et al.
Pubblicazione: (2024)
di: Bibbó, Gabriel, et al.
Pubblicazione: (2024)
Analytic Incremental Learning For Sound Source Localization With Imbalance Rectification
di: Fan, Zexia, et al.
Pubblicazione: (2026)
di: Fan, Zexia, et al.
Pubblicazione: (2026)
AnoPatch: Towards Better Consistency in Machine Anomalous Sound Detection
di: Jiang, Anbai, et al.
Pubblicazione: (2024)
di: Jiang, Anbai, et al.
Pubblicazione: (2024)
Quantum Kernels for Audio Deepfake Detection Using Spectrogram Patch Features
di: Amin, Lisan Al, et al.
Pubblicazione: (2026)
di: Amin, Lisan Al, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Deepfake Audio Detection Using Self-supervised Fusion Representations
di: Zaman, Khalid, et al.
Pubblicazione: (2026) -
Spectro-Temporal Modulation Representation Framework for Human-Imitated Speech Detection
di: Zaman, Khalid, et al.
Pubblicazione: (2026) -
Noise-Aware In-Context Learning for Hallucination Mitigation in ALLMs
di: Huang, Qixuan, et al.
Pubblicazione: (2026) -
Sub-Band Spectral Matching with Localized Score Aggregation for Robust Anomalous Sound Detection
di: Saengthong, Phurich, et al.
Pubblicazione: (2026) -
Improving Anomalous Sound Detection with Attribute-aware Representation from Domain-adaptive Pre-training
di: Fang, Xin, et al.
Pubblicazione: (2025)