Stem-JEPA: A Joint-Embedding Predictive Architecture for Musical Stem Compatibility Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Riou, Alain, Lattner, Stefan, Hadjeres, Gaëtan, Anslow, Michael, Peeters, Geoffroy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-shot Musical Stem Retrieval with Joint-Embedding Predictive Architectures
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
Investigating Design Choices in Joint-Embedding Predictive Architectures for General Audio Representation Learning
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
PESTO: Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2023)
von: Riou, Alain, et al.
Veröffentlicht: (2023)
PESTO: Real-Time Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2025)
von: Riou, Alain, et al.
Veröffentlicht: (2025)
Translation-Equivariant Self-Supervised Learning for Pitch Estimation with Optimal Transport
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
Music2Latent: Consistency Autoencoders for Latent Audio Compression
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
Music2Latent2: Audio Compression with Summary Embeddings and Autoregressive Decoding
von: Pasini, Marco, et al.
Veröffentlicht: (2025)
von: Pasini, Marco, et al.
Veröffentlicht: (2025)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
Automatic Music Sample Identification with Multi-Track Contrastive Learning
von: Riou, Alain, et al.
Veröffentlicht: (2025)
von: Riou, Alain, et al.
Veröffentlicht: (2025)
The Inverse Drum Machine: Source Separation Through Joint Transcription and Analysis-by-Synthesis
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
YourMT3+: Multi-instrument Music Transcription with Enhanced Transformer Architectures and Cross-dataset Stem Augmentation
von: Chang, Sungkyun, et al.
Veröffentlicht: (2024)
von: Chang, Sungkyun, et al.
Veröffentlicht: (2024)
Subtractive Training for Music Stem Insertion using Latent Diffusion Models
von: Villa-Renteria, Ivan, et al.
Veröffentlicht: (2024)
von: Villa-Renteria, Ivan, et al.
Veröffentlicht: (2024)
Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
von: Torres, Bernardo, et al.
Veröffentlicht: (2023)
von: Torres, Bernardo, et al.
Veröffentlicht: (2023)
Singer Identity Representation Learning using Self-Supervised Techniques
von: Torres, Bernardo, et al.
Veröffentlicht: (2024)
von: Torres, Bernardo, et al.
Veröffentlicht: (2024)
Bass Accompaniment Generation via Latent Diffusion
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
An Ensemble Approach to Music Source Separation: A Comparative Analysis of Conventional and Hierarchical Stem Separation
von: Vardhan, Saarth, et al.
Veröffentlicht: (2024)
von: Vardhan, Saarth, et al.
Veröffentlicht: (2024)
Hybrid Losses for Hierarchical Embedding Learning
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
Improving Musical Accompaniment Co-creation via Diffusion Transformers
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
SLAP: Siamese Language-Audio Pretraining Without Negative Samples for Music Understanding
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
StemGen: A music generation model that listens
von: Parker, Julian D., et al.
Veröffentlicht: (2023)
von: Parker, Julian D., et al.
Veröffentlicht: (2023)
Estimating Musical Surprisal in Audio
von: Bjare, Mathias Rose, et al.
Veröffentlicht: (2025)
von: Bjare, Mathias Rose, et al.
Veröffentlicht: (2025)
Episodic fine-tuning prototypical networks for optimization-based few-shot learning: Application to audio classification
von: Zhuang, Xuanyu, et al.
Veröffentlicht: (2024)
von: Zhuang, Xuanyu, et al.
Veröffentlicht: (2024)
CoDiCodec: Unifying Continuous and Discrete Compressed Representations of Audio
von: Pasini, Marco, et al.
Veröffentlicht: (2025)
von: Pasini, Marco, et al.
Veröffentlicht: (2025)
Estimating Musical Surprisal from Audio in Autoregressive Diffusion Model Noise Spaces
von: Bjare, Mathias Rose, et al.
Veröffentlicht: (2025)
von: Bjare, Mathias Rose, et al.
Veröffentlicht: (2025)
Blind estimation of audio effects using an auto-encoder approach and differentiable digital signal processing
von: Peladeau, Côme, et al.
Veröffentlicht: (2023)
von: Peladeau, Côme, et al.
Veröffentlicht: (2023)
A-JEPA: Joint-Embedding Predictive Architecture Can Listen
von: Fei, Zhengcong, et al.
Veröffentlicht: (2023)
von: Fei, Zhengcong, et al.
Veröffentlicht: (2023)
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
ACMID: Automatic Curation of Musical Instrument Dataset for 7-Stem Music Source Separation
von: Yu, Ji, et al.
Veröffentlicht: (2025)
von: Yu, Ji, et al.
Veröffentlicht: (2025)
Continuous Autoregressive Models with Noise Augmentation Avoid Error Accumulation
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
MGE-LDM: Joint Latent Diffusion for Simultaneous Music Generation and Source Extraction
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
Soft Clustering Anchors for Self-Supervised Speech Representation Learning in Joint Embedding Prediction Architectures
von: Ioannides, Georgios, et al.
Veröffentlicht: (2026)
von: Ioannides, Georgios, et al.
Veröffentlicht: (2026)
Tune It Up: Music Genre Transfer and Prediction
von: Samet, Fidan, et al.
Veröffentlicht: (2025)
von: Samet, Fidan, et al.
Veröffentlicht: (2025)
Music Emotion Prediction Using Recurrent Neural Networks
von: Chang, Xinyu, et al.
Veröffentlicht: (2024)
von: Chang, Xinyu, et al.
Veröffentlicht: (2024)
A Contrastive Self-Supervised Learning scheme for beat tracking amenable to few-shot learning
von: Gagnere, Antonin, et al.
Veröffentlicht: (2024)
von: Gagnere, Antonin, et al.
Veröffentlicht: (2024)
Controlling Contrastive Self-Supervised Learning with Knowledge-Driven Multiple Hypothesis: Application to Beat Tracking
von: Gagnere, Antonin, et al.
Veröffentlicht: (2025)
von: Gagnere, Antonin, et al.
Veröffentlicht: (2025)
Assessing the Alignment of Audio Representations with Timbre Similarity Ratings
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
Leveraging LLM Embeddings for Cross Dataset Label Alignment and Zero Shot Music Emotion Prediction
von: Liu, Renhang, et al.
Veröffentlicht: (2024)
von: Liu, Renhang, et al.
Veröffentlicht: (2024)
MusicGen-Stem: Multi-stem music generation and edition through autoregressive modeling
von: Rouard, Simon, et al.
Veröffentlicht: (2025)
von: Rouard, Simon, et al.
Veröffentlicht: (2025)
Unsupervised Blind Joint Dereverberation and Room Acoustics Estimation with Diffusion Models
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
MusicRL: Aligning Music Generation to Human Preferences
von: Cideron, Geoffrey, et al.
Veröffentlicht: (2024)
von: Cideron, Geoffrey, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Zero-shot Musical Stem Retrieval with Joint-Embedding Predictive Architectures
von: Riou, Alain, et al.
Veröffentlicht: (2024) -
Investigating Design Choices in Joint-Embedding Predictive Architectures for General Audio Representation Learning
von: Riou, Alain, et al.
Veröffentlicht: (2024) -
PESTO: Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2023) -
PESTO: Real-Time Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2025) -
Translation-Equivariant Self-Supervised Learning for Pitch Estimation with Optimal Transport
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)