LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Papaioannou, Charilaos, Benetos, Emmanouil, Potamianos, Alexandros |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Universal Music Representations? Evaluating Foundation Models on World Music Corpora
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2025)
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2025)
CultureMERT: Continual Pre-Training for Cross-Cultural Music Representation Learning
von: Kanatas, Angelos-Nikolaos, et al.
Veröffentlicht: (2025)
von: Kanatas, Angelos-Nikolaos, et al.
Veröffentlicht: (2025)
Learning Music Audio Representations With Limited Data
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
YourMT3+: Multi-instrument Music Transcription with Enhanced Transformer Architectures and Cross-dataset Stem Augmentation
von: Chang, Sungkyun, et al.
Veröffentlicht: (2024)
von: Chang, Sungkyun, et al.
Veröffentlicht: (2024)
RUMAA: Repeat-Aware Unified Music Audio Analysis for Score-Performance Alignment, Transcription, and Mistake Detection
von: Chang, Sungkyun, et al.
Veröffentlicht: (2025)
von: Chang, Sungkyun, et al.
Veröffentlicht: (2025)
Generalized Multi-Source Inference for Text Conditioned Music Diffusion Models
von: Postolache, Emilian, et al.
Veröffentlicht: (2024)
von: Postolache, Emilian, et al.
Veröffentlicht: (2024)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
MuChoMusic: Evaluating Music Understanding in Multimodal Audio-Language Models
von: Weck, Benno, et al.
Veröffentlicht: (2024)
von: Weck, Benno, et al.
Veröffentlicht: (2024)
SCRAPL: Scattering Transform with Random Paths for Machine Learning
von: Mitcheltree, Christopher, et al.
Veröffentlicht: (2026)
von: Mitcheltree, Christopher, et al.
Veröffentlicht: (2026)
Acoustic identification of individual animals with hierarchical contrastive learning
von: Nolasco, Ines, et al.
Veröffentlicht: (2024)
von: Nolasco, Ines, et al.
Veröffentlicht: (2024)
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
von: Singh, Shubhr, et al.
Veröffentlicht: (2025)
von: Singh, Shubhr, et al.
Veröffentlicht: (2025)
Audio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue Abilities
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
CMI-Bench: A Comprehensive Benchmark for Evaluating Music Instruction Following
von: Ma, Yinghao, et al.
Veröffentlicht: (2025)
von: Ma, Yinghao, et al.
Veröffentlicht: (2025)
A Data-Driven Analysis of Robust Automatic Piano Transcription
von: Edwards, Drew, et al.
Veröffentlicht: (2024)
von: Edwards, Drew, et al.
Veröffentlicht: (2024)
Hierarchical Label Propagation: A Model-Size-Dependent Performance Booster for AudioSet Tagging
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
An Experimental Comparison Of Multi-view Self-supervised Methods For Music Tagging
von: Meseguer-Brocal, Gabriel, et al.
Veröffentlicht: (2024)
von: Meseguer-Brocal, Gabriel, et al.
Veröffentlicht: (2024)
On the Transferability of Large-Scale Self-Supervision to Few-Shot Audio Classification
von: Heggan, Calum, et al.
Veröffentlicht: (2024)
von: Heggan, Calum, et al.
Veröffentlicht: (2024)
Semantic-Aware Interpretable Multimodal Music Auto-Tagging
von: Patakis, Andreas, et al.
Veröffentlicht: (2025)
von: Patakis, Andreas, et al.
Veröffentlicht: (2025)
Enhancing Lyrics Transcription on Music Mixtures with Consistency Loss
von: Huang, Jiawen, et al.
Veröffentlicht: (2025)
von: Huang, Jiawen, et al.
Veröffentlicht: (2025)
Multi-Stage Music Source Restoration with BandSplit-RoFormer Separation and HiFi++ GAN
von: Morocutti, Tobias, et al.
Veröffentlicht: (2026)
von: Morocutti, Tobias, et al.
Veröffentlicht: (2026)
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
Multi-label Zero-Shot Audio Classification with Temporal Attention
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
Self-Supervised Learning for Few-Shot Bird Sound Classification
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
The Rarity of Musical Audio Signals Within the Space of Possible Audio Generation
von: Collins, Nick
Veröffentlicht: (2024)
von: Collins, Nick
Veröffentlicht: (2024)
Geo-ATBench: A Benchmark for Geospatial Audio Tagging with Geospatial Semantic Context
von: Hou, Yuanbo, et al.
Veröffentlicht: (2026)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2026)
Zero Shot Audio to Audio Emotion Transfer With Speaker Disentanglement
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
Domain-Invariant Representation Learning of Bird Sounds
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
EditGen: Harnessing Cross-Attention Control for Instruction-Based Auto-Regressive Audio Editing
von: Sioros, Vassilis, et al.
Veröffentlicht: (2025)
von: Sioros, Vassilis, et al.
Veröffentlicht: (2025)
Do Foundational Audio Encoders Understand Music Structure?
von: Toyama, Keisuke, et al.
Veröffentlicht: (2025)
von: Toyama, Keisuke, et al.
Veröffentlicht: (2025)
Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation
von: Garcia, Nelly, et al.
Veröffentlicht: (2026)
von: Garcia, Nelly, et al.
Veröffentlicht: (2026)
Leveraging LLM Embeddings for Cross Dataset Label Alignment and Zero Shot Music Emotion Prediction
von: Liu, Renhang, et al.
Veröffentlicht: (2024)
von: Liu, Renhang, et al.
Veröffentlicht: (2024)
Towards Robust Few-shot Class Incremental Learning in Audio Classification using Contrastive Representation
von: Singh, Riyansha, et al.
Veröffentlicht: (2024)
von: Singh, Riyansha, et al.
Veröffentlicht: (2024)
Audio Processing using Pattern Recognition for Music Genre Classification
von: Chatterjee, Sivangi, et al.
Veröffentlicht: (2024)
von: Chatterjee, Sivangi, et al.
Veröffentlicht: (2024)
Music2Latent: Consistency Autoencoders for Latent Audio Compression
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
EnvId: A Metric Learning Approach for Forensic Few-Shot Identification of Unseen Environments
von: Moussa, Denise, et al.
Veröffentlicht: (2024)
von: Moussa, Denise, et al.
Veröffentlicht: (2024)
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
Music Boomerang: Reusing Diffusion Models for Data Augmentation and Audio Manipulation
von: Fichtinger, Alexander, et al.
Veröffentlicht: (2025)
von: Fichtinger, Alexander, et al.
Veröffentlicht: (2025)
LiLAC: A Lightweight Latent ControlNet for Musical Audio Generation
von: Baker, Tom, et al.
Veröffentlicht: (2025)
von: Baker, Tom, et al.
Veröffentlicht: (2025)
Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model
von: Huang, Jiawen, et al.
Veröffentlicht: (2024)
von: Huang, Jiawen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Universal Music Representations? Evaluating Foundation Models on World Music Corpora
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2025) -
CultureMERT: Continual Pre-Training for Cross-Cultural Music Representation Learning
von: Kanatas, Angelos-Nikolaos, et al.
Veröffentlicht: (2025) -
Learning Music Audio Representations With Limited Data
von: Plachouras, Christos, et al.
Veröffentlicht: (2025) -
YourMT3+: Multi-instrument Music Transcription with Enhanced Transformer Architectures and Cross-dataset Stem Augmentation
von: Chang, Sungkyun, et al.
Veröffentlicht: (2024) -
RUMAA: Repeat-Aware Unified Music Audio Analysis for Score-Performance Alignment, Transcription, and Mistake Detection
von: Chang, Sungkyun, et al.
Veröffentlicht: (2025)