Saved in:
| Main Authors: | Plachouras, Christos, Guinot, Julien, Fazekas, George, Quinton, Elio, Benetos, Emmanouil, Pauwels, Johan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.06224 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Music Audio Representations With Limited Data
by: Plachouras, Christos, et al.
Published: (2025)
by: Plachouras, Christos, et al.
Published: (2025)
MuChoMusic: Evaluating Music Understanding in Multimodal Audio-Language Models
by: Weck, Benno, et al.
Published: (2024)
by: Weck, Benno, et al.
Published: (2024)
Semi-Supervised Contrastive Learning of Musical Representations
by: Guinot, Julien, et al.
Published: (2024)
by: Guinot, Julien, et al.
Published: (2024)
GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models
by: Guinot, Julien, et al.
Published: (2025)
by: Guinot, Julien, et al.
Published: (2025)
Leave-One-EquiVariant: Alleviating invariance-related information loss in contrastive music representations
by: Guinot, Julien, et al.
Published: (2024)
by: Guinot, Julien, et al.
Published: (2024)
SLAP: Siamese Language-Audio Pretraining Without Negative Samples for Music Understanding
by: Guinot, Julien, et al.
Published: (2025)
by: Guinot, Julien, et al.
Published: (2025)
Universal Music Representations? Evaluating Foundation Models on World Music Corpora
by: Papaioannou, Charilaos, et al.
Published: (2025)
by: Papaioannou, Charilaos, et al.
Published: (2025)
RUMAA: Repeat-Aware Unified Music Audio Analysis for Score-Performance Alignment, Transcription, and Mistake Detection
by: Chang, Sungkyun, et al.
Published: (2025)
by: Chang, Sungkyun, et al.
Published: (2025)
CoDiCodec: Unifying Continuous and Discrete Compressed Representations of Audio
by: Pasini, Marco, et al.
Published: (2025)
by: Pasini, Marco, et al.
Published: (2025)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
by: Tuncay, Ludovic, et al.
Published: (2025)
by: Tuncay, Ludovic, et al.
Published: (2025)
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
by: Papaioannou, Charilaos, et al.
Published: (2024)
by: Papaioannou, Charilaos, et al.
Published: (2024)
Towards Effective Negation Modeling in Joint Audio-Text Models for Music
by: Vasilakis, Yannis, et al.
Published: (2026)
by: Vasilakis, Yannis, et al.
Published: (2026)
CMI-Bench: A Comprehensive Benchmark for Evaluating Music Instruction Following
by: Ma, Yinghao, et al.
Published: (2025)
by: Ma, Yinghao, et al.
Published: (2025)
Acoustic identification of individual animals with hierarchical contrastive learning
by: Nolasco, Ines, et al.
Published: (2024)
by: Nolasco, Ines, et al.
Published: (2024)
YourMT3+: Multi-instrument Music Transcription with Enhanced Transformer Architectures and Cross-dataset Stem Augmentation
by: Chang, Sungkyun, et al.
Published: (2024)
by: Chang, Sungkyun, et al.
Published: (2024)
SCRAPL: Scattering Transform with Random Paths for Machine Learning
by: Mitcheltree, Christopher, et al.
Published: (2026)
by: Mitcheltree, Christopher, et al.
Published: (2026)
Evaluation of pretrained language models on music understanding
by: Vasilakis, Yannis, et al.
Published: (2024)
by: Vasilakis, Yannis, et al.
Published: (2024)
Comparison of Autoencoder Encodings for ECG Representation in Downstream Prediction Tasks
by: Harvey, Christopher J., et al.
Published: (2024)
by: Harvey, Christopher J., et al.
Published: (2024)
Generalized Multi-Source Inference for Text Conditioned Music Diffusion Models
by: Postolache, Emilian, et al.
Published: (2024)
by: Postolache, Emilian, et al.
Published: (2024)
A Data-Driven Analysis of Robust Automatic Piano Transcription
by: Edwards, Drew, et al.
Published: (2024)
by: Edwards, Drew, et al.
Published: (2024)
Generalized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on Graphs
by: Yu, Xingtong, et al.
Published: (2023)
by: Yu, Xingtong, et al.
Published: (2023)
Task Priors: Enhancing Model Evaluation by Considering the Entire Space of Downstream Tasks
by: Patel, Niket, et al.
Published: (2025)
by: Patel, Niket, et al.
Published: (2025)
Smoke and Mirrors in Causal Downstream Tasks
by: Cadei, Riccardo, et al.
Published: (2024)
by: Cadei, Riccardo, et al.
Published: (2024)
Dataset Representativeness and Downstream Task Fairness
by: Borza, Victor, et al.
Published: (2024)
by: Borza, Victor, et al.
Published: (2024)
I can listen but cannot read: An evaluation of two-tower multimodal systems for instrument recognition
by: Vasilakis, Yannis, et al.
Published: (2024)
by: Vasilakis, Yannis, et al.
Published: (2024)
Coupling Speech Encoders with Downstream Text Models
by: Chelba, Ciprian, et al.
Published: (2024)
by: Chelba, Ciprian, et al.
Published: (2024)
Towards a Unified Framework for Evaluating Explanations
by: Pinto, Juan D., et al.
Published: (2024)
by: Pinto, Juan D., et al.
Published: (2024)
Panprediction: Optimal Predictions for Any Downstream Task and Loss
by: Balakrishnan, Sivaraman, et al.
Published: (2025)
by: Balakrishnan, Sivaraman, et al.
Published: (2025)
Pretraining Induces a Reusable Spectral Basis for Downstream Task Adaptation
by: Yu, Junjie, et al.
Published: (2026)
by: Yu, Junjie, et al.
Published: (2026)
JAZZVAR: A Dataset of Variations found within Solo Piano Performances of Jazz Standards for Music Overpainting
by: Row, Eleanor, et al.
Published: (2023)
by: Row, Eleanor, et al.
Published: (2023)
Music2Latent: Consistency Autoencoders for Latent Audio Compression
by: Pasini, Marco, et al.
Published: (2024)
by: Pasini, Marco, et al.
Published: (2024)
ECG Latent Feature Extraction with Autoencoders for Downstream Prediction Tasks
by: Harvey, Christopher, et al.
Published: (2025)
by: Harvey, Christopher, et al.
Published: (2025)
Handling Missing Data in Downstream Tasks With Distribution-Preserving Guarantees
by: Bordoloi, Rahul, et al.
Published: (2025)
by: Bordoloi, Rahul, et al.
Published: (2025)
Learning Treatment Representations for Downstream Instrumental Variable Regression
by: Lin, Shiangyi, et al.
Published: (2025)
by: Lin, Shiangyi, et al.
Published: (2025)
Attacking Attention of Foundation Models Disrupts Downstream Tasks
by: Silva, Hondamunige Prasanna, et al.
Published: (2025)
by: Silva, Hondamunige Prasanna, et al.
Published: (2025)
Aligning the Evaluation of Probabilistic Predictions with Downstream Value
by: Shahroudi, Novin, et al.
Published: (2025)
by: Shahroudi, Novin, et al.
Published: (2025)
Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation
by: Garcia, Nelly, et al.
Published: (2026)
by: Garcia, Nelly, et al.
Published: (2026)
Task-tailored Pre-processing: Fair Downstream Supervised Learning
by: Sohn, Jinwon, et al.
Published: (2026)
by: Sohn, Jinwon, et al.
Published: (2026)
Music2Latent2: Audio Compression with Summary Embeddings and Autoregressive Decoding
by: Pasini, Marco, et al.
Published: (2025)
by: Pasini, Marco, et al.
Published: (2025)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
by: Yu, Chin-Yun, et al.
Published: (2024)
by: Yu, Chin-Yun, et al.
Published: (2024)
Similar Items
-
Learning Music Audio Representations With Limited Data
by: Plachouras, Christos, et al.
Published: (2025) -
MuChoMusic: Evaluating Music Understanding in Multimodal Audio-Language Models
by: Weck, Benno, et al.
Published: (2024) -
Semi-Supervised Contrastive Learning of Musical Representations
by: Guinot, Julien, et al.
Published: (2024) -
GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models
by: Guinot, Julien, et al.
Published: (2025) -
Leave-One-EquiVariant: Alleviating invariance-related information loss in contrastive music representations
by: Guinot, Julien, et al.
Published: (2024)