Prototypical Contrastive Learning For Improved Few-Shot Audio Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sgouropoulos, Christos, Nikou, Christos, Vlachos, Stefanos, Theiou, Vasileios, Foukanelis, Christos, Giannakopoulos, Theodoros |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Contrastive and Transfer Learning for Effective Audio Fingerprinting through a Real-World Evaluation Protocol
von: Nikou, Christos, et al.
Veröffentlicht: (2025)
von: Nikou, Christos, et al.
Veröffentlicht: (2025)
Improving Audio Classification by Transitioning from Zero- to Few-Shot
von: Taylor, James, et al.
Veröffentlicht: (2025)
von: Taylor, James, et al.
Veröffentlicht: (2025)
Learning Music Audio Representations With Limited Data
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2023)
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2023)
InterGridNet: An Electric Network Frequency Approach for Audio Source Location Classification Using Convolutional Neural Networks
von: Korgialas, Christos, et al.
Veröffentlicht: (2025)
von: Korgialas, Christos, et al.
Veröffentlicht: (2025)
APEX: Audio Prototype EXplanations for Classification Tasks
von: Kawa, Piotr, et al.
Veröffentlicht: (2026)
von: Kawa, Piotr, et al.
Veröffentlicht: (2026)
On the Transferability of Large-Scale Self-Supervision to Few-Shot Audio Classification
von: Heggan, Calum, et al.
Veröffentlicht: (2024)
von: Heggan, Calum, et al.
Veröffentlicht: (2024)
Towards Robust Few-shot Class Incremental Learning in Audio Classification using Contrastive Representation
von: Singh, Riyansha, et al.
Veröffentlicht: (2024)
von: Singh, Riyansha, et al.
Veröffentlicht: (2024)
Audio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue Abilities
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
Self-Supervised Learning for Few-Shot Bird Sound Classification
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
Improving Multimodal Learning with Multi-Loss Gradient Modulation
von: Kontras, Konstantinos, et al.
Veröffentlicht: (2024)
von: Kontras, Konstantinos, et al.
Veröffentlicht: (2024)
Leveraging Prediction Entropy for Automatic Prompt Weighting in Zero-Shot Audio-Language Classification
von: Khoury, Karim El, et al.
Veröffentlicht: (2026)
von: Khoury, Karim El, et al.
Veröffentlicht: (2026)
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
AudioMosaic: Contrastive Masked Audio Representation Learning
von: Huang, Hanxun, et al.
Veröffentlicht: (2026)
von: Huang, Hanxun, et al.
Veröffentlicht: (2026)
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
Multi-label Zero-Shot Audio Classification with Temporal Attention
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
Latent Space Disentanglement via Activation Steering for Interpretable Attribute Control in Symbolic Music Generation
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
Beyond the Hook: Predicting Billboard Hot 100 Chart Inclusion with Machine Learning from Streaming, Audio Signals, and Perceptual Features
von: Mountzouris, Christos
Veröffentlicht: (2025)
von: Mountzouris, Christos
Veröffentlicht: (2025)
Contrastive Learning from Synthetic Audio Doppelgängers
von: Cherep, Manuel, et al.
Veröffentlicht: (2024)
von: Cherep, Manuel, et al.
Veröffentlicht: (2024)
Text Classification: Neural Networks VS Machine Learning Models VS Pre-trained Models
von: Petridis, Christos
Veröffentlicht: (2024)
von: Petridis, Christos
Veröffentlicht: (2024)
TSPE: Task-Specific Prompt Ensemble for Improved Zero-Shot Audio Classification
von: Anand, Nishit, et al.
Veröffentlicht: (2024)
von: Anand, Nishit, et al.
Veröffentlicht: (2024)
Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation
von: Zarkadis, Iakovos-Christos, et al.
Veröffentlicht: (2026)
von: Zarkadis, Iakovos-Christos, et al.
Veröffentlicht: (2026)
Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
CVSM: Contrastive Vocal Similarity Modeling
von: Garoufis, Christos, et al.
Veröffentlicht: (2025)
von: Garoufis, Christos, et al.
Veröffentlicht: (2025)
Multimodal Attention Merging for Improved Speech Recognition and Audio Event Classification
von: Sundar, Anirudh S., et al.
Veröffentlicht: (2023)
von: Sundar, Anirudh S., et al.
Veröffentlicht: (2023)
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
Zero Shot Audio to Audio Emotion Transfer With Speaker Disentanglement
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
Adaptive Discovery of Interpretable Audio Attributes with Multimodal LLMs for Low-Resource Classification
von: Yoshimura, Kosuke, et al.
Veröffentlicht: (2026)
von: Yoshimura, Kosuke, et al.
Veröffentlicht: (2026)
PACE: Pretrained Audio Continual Learning
von: Li, Chang, et al.
Veröffentlicht: (2026)
von: Li, Chang, et al.
Veröffentlicht: (2026)
Supervised Contrastive Representation Learning: Landscape Analysis with Unconstrained Features
von: Behnia, Tina, et al.
Veröffentlicht: (2024)
von: Behnia, Tina, et al.
Veröffentlicht: (2024)
REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion
von: Petsangourakis, Giorgos, et al.
Veröffentlicht: (2025)
von: Petsangourakis, Giorgos, et al.
Veröffentlicht: (2025)
JEPA as a Neural Tokenizer: Learning Robust Speech Representations with Density Adaptive Attention
von: Ioannides, Georgios, et al.
Veröffentlicht: (2025)
von: Ioannides, Georgios, et al.
Veröffentlicht: (2025)
Audio Flamingo Sound-CoT Technical Report: Improving Chain-of-Thought Reasoning in Sound Understanding
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
EnvId: A Metric Learning Approach for Forensic Few-Shot Identification of Unseen Environments
von: Moussa, Denise, et al.
Veröffentlicht: (2024)
von: Moussa, Denise, et al.
Veröffentlicht: (2024)
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
How to Label Resynthesized Audio: The Dual Role of Neural Audio Codecs in Audio Deepfake Detection
von: Xiao, Yixuan, et al.
Veröffentlicht: (2026)
von: Xiao, Yixuan, et al.
Veröffentlicht: (2026)
Improving Out-of-Domain Audio Deepfake Detection via Layer Selection and Fusion of SSL-Based Countermeasures
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
ADNAC: Audio Denoiser using Neural Audio Codec
von: Jimon, Daniel, et al.
Veröffentlicht: (2025)
von: Jimon, Daniel, et al.
Veröffentlicht: (2025)
When Denoising Hinders: Revisiting Zero-Shot ASR with SAM-Audio and Whisper
von: Islam, Akif, et al.
Veröffentlicht: (2026)
von: Islam, Akif, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Contrastive and Transfer Learning for Effective Audio Fingerprinting through a Real-World Evaluation Protocol
von: Nikou, Christos, et al.
Veröffentlicht: (2025) -
Improving Audio Classification by Transitioning from Zero- to Few-Shot
von: Taylor, James, et al.
Veröffentlicht: (2025) -
Learning Music Audio Representations With Limited Data
von: Plachouras, Christos, et al.
Veröffentlicht: (2025) -
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2023) -
InterGridNet: An Electric Network Frequency Approach for Audio Source Location Classification Using Convolutional Neural Networks
von: Korgialas, Christos, et al.
Veröffentlicht: (2025)