Classification of Adventitious Sounds Combining Cochleogram and Vision Transformers
Fuente:
arXiv
Salvato in:
| Autori principali: | Mang, Loredana Daria, Martinez, Francisco David Gonzalez, Munoz, Damian Martinez, Galan, Sebastian Garcia, Cortina, Raquel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An incremental algorithm based on multichannel non-negative matrix partial co-factorization for ambient denoising in auscultation
di: Cruz, Juan De La Torre, et al.
Pubblicazione: (2024)
di: Cruz, Juan De La Torre, et al.
Pubblicazione: (2024)
ILD-VIT: A Unified Vision Transformer Architecture for Detection of Interstitial Lung Disease from Respiratory Sounds
di: Hota, Soubhagya Ranjan, et al.
Pubblicazione: (2025)
di: Hota, Soubhagya Ranjan, et al.
Pubblicazione: (2025)
An ambient denoising method based on multi-channel non-negative matrix factorization for wheezing detection
di: Muñoz-Montoro, Antonio J., et al.
Pubblicazione: (2024)
di: Muñoz-Montoro, Antonio J., et al.
Pubblicazione: (2024)
Unsupervised detection and classification of heartbeats using the dissimilarity matrix in PCG signals
di: Torre-Cruz, J., et al.
Pubblicazione: (2024)
di: Torre-Cruz, J., et al.
Pubblicazione: (2024)
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
Sound field estimation with moving microphones using kernel ridge regression
di: Brunnström, Jesper, et al.
Pubblicazione: (2025)
di: Brunnström, Jesper, et al.
Pubblicazione: (2025)
Robust Fixed-Filter Sound Zone Control with Audio-Based Position Tracking
di: Bhattacharjee, Sankha Subhra, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Sankha Subhra, et al.
Pubblicazione: (2024)
A Zero-Shot Physics-Informed Dictionary Learning Approach for Sound Field Reconstruction
di: Damiano, Stefano, et al.
Pubblicazione: (2024)
di: Damiano, Stefano, et al.
Pubblicazione: (2024)
MASSLOC: A Massive Sound Source Localization System based on Direction-of-Arrival Estimation
di: Fischer, Georg K. J., et al.
Pubblicazione: (2025)
di: Fischer, Georg K. J., et al.
Pubblicazione: (2025)
Exploring Audio-Visual Information Fusion for Sound Event Localization and Detection In Low-Resource Realistic Scenarios
di: Jiang, Ya, et al.
Pubblicazione: (2024)
di: Jiang, Ya, et al.
Pubblicazione: (2024)
SpeechMLC: Speech Multi-label Classification
di: Kim, Miseul, et al.
Pubblicazione: (2025)
di: Kim, Miseul, et al.
Pubblicazione: (2025)
FUN-SSL: Full-band Layer Followed by U-Net with Narrow-band Layers for Multiple Moving Sound Source Localization
di: Choi, Yuseon, et al.
Pubblicazione: (2025)
di: Choi, Yuseon, et al.
Pubblicazione: (2025)
Parameter-Efficient Fine-Tuning of Foundation Models for CLP Speech Classification
di: Bhattacharjee, Susmita, et al.
Pubblicazione: (2025)
di: Bhattacharjee, Susmita, et al.
Pubblicazione: (2025)
Detection of manatee vocalisations using the Audio Spectrogram Transformer
di: Schiappacasse, Stefano, et al.
Pubblicazione: (2024)
di: Schiappacasse, Stefano, et al.
Pubblicazione: (2024)
Optimizing Domain-Adaptive Self-Supervised Learning for Clinical Voice-Based Disease Classification
di: Liu, Weixin, et al.
Pubblicazione: (2026)
di: Liu, Weixin, et al.
Pubblicazione: (2026)
Independent Feature Enhanced Crossmodal Fusion for Match-Mismatch Classification of Speech Stimulus and EEG Response
di: Fan, Shitong, et al.
Pubblicazione: (2024)
di: Fan, Shitong, et al.
Pubblicazione: (2024)
Comparison of Tiny Machine Learning Techniques for Embedded Acoustic Emission Analysis
di: Muthumala, Uditha, et al.
Pubblicazione: (2024)
di: Muthumala, Uditha, et al.
Pubblicazione: (2024)
String Sound Synthesizer on GPU-accelerated Finite Difference Scheme
di: Lee, Jin Woo, et al.
Pubblicazione: (2023)
di: Lee, Jin Woo, et al.
Pubblicazione: (2023)
Reduce Computational Complexity for Continuous Wavelet Transform in Acoustic Recognition Using Hop Size
di: Phan, Dang Thoai
Pubblicazione: (2024)
di: Phan, Dang Thoai
Pubblicazione: (2024)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
di: Berghi, Davide, et al.
Pubblicazione: (2025)
di: Berghi, Davide, et al.
Pubblicazione: (2025)
Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier
di: Dumpis, Martynas, et al.
Pubblicazione: (2026)
di: Dumpis, Martynas, et al.
Pubblicazione: (2026)
Joint Spectrogram Separation and TDOA Estimation using Optimal Transport
di: Fabiani, Linda, et al.
Pubblicazione: (2025)
di: Fabiani, Linda, et al.
Pubblicazione: (2025)
On the Invariance of Cross-Correlation Peak Positions Under Monotonic Signal Transformations, with Application to Fast Time Difference Estimation
di: Ueno, Natsuki, et al.
Pubblicazione: (2025)
di: Ueno, Natsuki, et al.
Pubblicazione: (2025)
Audio Compression using Periodic Gabor with Biorthogonal Exchange: Implementation Using the Zak Transform
di: Alimi, Roger, et al.
Pubblicazione: (2025)
di: Alimi, Roger, et al.
Pubblicazione: (2025)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
Online Similarity-and-Independence-Aware Beamformer for Low-latency Target Sound Extraction
di: Hiroe, Atsuo
Pubblicazione: (2023)
di: Hiroe, Atsuo
Pubblicazione: (2023)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
di: Torabi, Yasaman, et al.
Pubblicazione: (2025)
di: Torabi, Yasaman, et al.
Pubblicazione: (2025)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
di: Yao, Shengshi, et al.
Pubblicazione: (2025)
di: Yao, Shengshi, et al.
Pubblicazione: (2025)
Mind the Prompt: Prompting Strategies in Audio Generations for Improving Sound Classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
di: Ronchini, Francesca, et al.
Pubblicazione: (2025)
Decomposing the Influence of Physical Acoustic Modeling on Neural Personal Sound Zone Rendering: An Ablation Study
di: Jiang, Hao, et al.
Pubblicazione: (2026)
di: Jiang, Hao, et al.
Pubblicazione: (2026)
Automatic Voice Classification Of Autistic Subjects
di: Vacca, Jessica, et al.
Pubblicazione: (2024)
di: Vacca, Jessica, et al.
Pubblicazione: (2024)
Interfacing PDM MEMS microphones with PFM spiking systems: Application for Neuromorphic Auditory Sensors
di: Jimenez-Fernandez, Angel, et al.
Pubblicazione: (2019)
di: Jimenez-Fernandez, Angel, et al.
Pubblicazione: (2019)
Improving snore detection under limited dataset through harmonic/percussive source separation and convolutional neural networks
di: Gonzalez-Martinez, F. D., et al.
Pubblicazione: (2024)
di: Gonzalez-Martinez, F. D., et al.
Pubblicazione: (2024)
Classification of Heart Sounds Using Multi-Branch Deep Convolutional Network and LSTM-CNN
di: Latifi, Seyed Amir, et al.
Pubblicazione: (2024)
di: Latifi, Seyed Amir, et al.
Pubblicazione: (2024)
Room Impulse Response Estimation using Optimal Transport: Simulation-Informed Inference
di: Sundström, David, et al.
Pubblicazione: (2024)
di: Sundström, David, et al.
Pubblicazione: (2024)
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
Analytical model for the relation between signal bandwidth and spatial resolution in Steered-Response Power Phase Transform (SRP-PHAT) maps
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
Ambisonics Encoder for Wearable Array with Improved Binaural Reproduction
di: Gayer, Yhonatan, et al.
Pubblicazione: (2025)
di: Gayer, Yhonatan, et al.
Pubblicazione: (2025)
Comparison of Classification Algorithms for COVID19 Detection using Cough Acoustic Signals
di: Erdoğan, Yunus Emre, et al.
Pubblicazione: (2022)
di: Erdoğan, Yunus Emre, et al.
Pubblicazione: (2022)
Cough-E: A multimodal, privacy-preserving cough detection algorithm for the edge
di: Albini, Stefano, et al.
Pubblicazione: (2024)
di: Albini, Stefano, et al.
Pubblicazione: (2024)
Documenti analoghi
-
An incremental algorithm based on multichannel non-negative matrix partial co-factorization for ambient denoising in auscultation
di: Cruz, Juan De La Torre, et al.
Pubblicazione: (2024) -
ILD-VIT: A Unified Vision Transformer Architecture for Detection of Interstitial Lung Disease from Respiratory Sounds
di: Hota, Soubhagya Ranjan, et al.
Pubblicazione: (2025) -
An ambient denoising method based on multi-channel non-negative matrix factorization for wheezing detection
di: Muñoz-Montoro, Antonio J., et al.
Pubblicazione: (2024) -
Unsupervised detection and classification of heartbeats using the dissimilarity matrix in PCG signals
di: Torre-Cruz, J., et al.
Pubblicazione: (2024) -
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
di: Huang, Wei, et al.
Pubblicazione: (2025)