Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wilkins, Julia, Ding, Sivan, Fuentes, Magdalena, Bello, Juan Pablo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Supervised Multi-View Learning for Disentangled Music Audio Representations
von: Wilkins, Julia, et al.
Veröffentlicht: (2024)
von: Wilkins, Julia, et al.
Veröffentlicht: (2024)
Do Music Source Separation Models Preserve Spatial Information in Binaural Audio?
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
Learning Multidimensional Disentangled Representations of Instrumental Sounds for Musical Similarity Assessment
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024)
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
von: Shi, Runwu, et al.
Veröffentlicht: (2024)
von: Shi, Runwu, et al.
Veröffentlicht: (2024)
SONIQUE: Video Background Music Generation Using Unpaired Audio-Visual Data
von: Zhang, Liqian, et al.
Veröffentlicht: (2024)
von: Zhang, Liqian, et al.
Veröffentlicht: (2024)
Musical Source Separation of Brazilian Percussion
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
Quantifying Dimensional Independence in Speech: An Information-Theoretic Framework for Disentangled Representation Learning
von: Kashyap, Bipasha, et al.
Veröffentlicht: (2026)
von: Kashyap, Bipasha, et al.
Veröffentlicht: (2026)
Emotion-Aware Speech Self-Supervised Representation Learning with Intensity Knowledge
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Evaluating Disentangled Representations for Controllable Music Generation
von: Ibáñez-Martínez, Laura, et al.
Veröffentlicht: (2026)
von: Ibáñez-Martínez, Laura, et al.
Veröffentlicht: (2026)
Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision
von: Chen, Yafeng, et al.
Veröffentlicht: (2023)
von: Chen, Yafeng, et al.
Veröffentlicht: (2023)
Disentangled Representation Learning for Environment-agnostic Speaker Recognition
von: Nam, KiHyun, et al.
Veröffentlicht: (2024)
von: Nam, KiHyun, et al.
Veröffentlicht: (2024)
Learning Disentangled Speech Representations with Contrastive Learning and Time-Invariant Retrieval
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
MEDIC: Zero-shot Music Editing with Disentangled Inversion Control
von: Liu, Huadai, et al.
Veröffentlicht: (2024)
von: Liu, Huadai, et al.
Veröffentlicht: (2024)
Learning Separated Representations for Instrument-based Music Similarity
von: Hashizume, Yuka, et al.
Veröffentlicht: (2025)
von: Hashizume, Yuka, et al.
Veröffentlicht: (2025)
VISinger2+: End-to-End Singing Voice Synthesis Augmented by Self-Supervised Learning Representation
von: Yu, Yifeng, et al.
Veröffentlicht: (2024)
von: Yu, Yifeng, et al.
Veröffentlicht: (2024)
Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models
von: Zhou, Yizhi, et al.
Veröffentlicht: (2025)
von: Zhou, Yizhi, et al.
Veröffentlicht: (2025)
Post-Training Quantization for Audio Diffusion Transformers
von: Khandelwal, Tanmay, et al.
Veröffentlicht: (2025)
von: Khandelwal, Tanmay, et al.
Veröffentlicht: (2025)
Self-Supervised Learning of Spatial Acoustic Representation with Cross-Channel Signal Reconstruction and Multi-Channel Conformer
von: Yang, Bing, et al.
Veröffentlicht: (2023)
von: Yang, Bing, et al.
Veröffentlicht: (2023)
Refining Self-Supervised Learnt Speech Representation using Brain Activations
von: Li, Hengyu, et al.
Veröffentlicht: (2024)
von: Li, Hengyu, et al.
Veröffentlicht: (2024)
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
von: Hussein, Amir, et al.
Veröffentlicht: (2025)
von: Hussein, Amir, et al.
Veröffentlicht: (2025)
Semi-Supervised Self-Learning Enhanced Music Emotion Recognition
von: Sun, Yifu, et al.
Veröffentlicht: (2024)
von: Sun, Yifu, et al.
Veröffentlicht: (2024)
Distillation and Pruning for Scalable Self-Supervised Representation-Based Speech Quality Assessment
von: Stahl, Benjamin, et al.
Veröffentlicht: (2025)
von: Stahl, Benjamin, et al.
Veröffentlicht: (2025)
Adapting General Disentanglement-Based Speaker Anonymization for Enhanced Emotion Preservation
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
Learning Expressive Disentangled Speech Representations with Soft Speech Units and Adversarial Style Augmentation
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
Leveraging Self-Supervised Learning for Speaker Diarization
von: Han, Jiangyu, et al.
Veröffentlicht: (2024)
von: Han, Jiangyu, et al.
Veröffentlicht: (2024)
Self-Supervised Disentangled Representation Learning for Robust Target Speech Extraction
von: Mu, Zhaoxi, et al.
Veröffentlicht: (2023)
von: Mu, Zhaoxi, et al.
Veröffentlicht: (2023)
OMAR-RQ: Open Music Audio Representation Model Trained with Multi-Feature Masked Token Prediction
von: Alonso-Jiménez, Pablo, et al.
Veröffentlicht: (2025)
von: Alonso-Jiménez, Pablo, et al.
Veröffentlicht: (2025)
Music Similarity Representation Learning Focusing on Individual Instruments with Source Separation and Human Preference
von: Imamura, Takehiro, et al.
Veröffentlicht: (2025)
von: Imamura, Takehiro, et al.
Veröffentlicht: (2025)
Latent Acoustic Mapping for Direction of Arrival Estimation: A Self-Supervised Approach
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
A Large-Scale Probing Analysis of Speaker-Specific Attributes in Self-Supervised Speech Representations
von: Chiu, Aemon Yat Fei, et al.
Veröffentlicht: (2025)
von: Chiu, Aemon Yat Fei, et al.
Veröffentlicht: (2025)
Speaker Recognition Using Isomorphic Graph Attention Network Based Pooling on Self-Supervised Representation
von: Ge, Zirui, et al.
Veröffentlicht: (2023)
von: Ge, Zirui, et al.
Veröffentlicht: (2023)
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval
von: Stewart, Shanti, et al.
Veröffentlicht: (2024)
von: Stewart, Shanti, et al.
Veröffentlicht: (2024)
Learning Disentangled Speech Representations
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
Low-Resource Self-Supervised Learning with SSL-Enhanced TTS
von: Hsu, Po-chun, et al.
Veröffentlicht: (2023)
von: Hsu, Po-chun, et al.
Veröffentlicht: (2023)
Geometric Analysis of Speech Representation Spaces: Topological Disentanglement and Confound Detection
von: Kashyap, Bipasha, et al.
Veröffentlicht: (2026)
von: Kashyap, Bipasha, et al.
Veröffentlicht: (2026)
Emotion-driven Piano Music Generation via Two-stage Disentanglement and Functional Representation
von: Huang, Jingyue, et al.
Veröffentlicht: (2024)
von: Huang, Jingyue, et al.
Veröffentlicht: (2024)
Pianoroll-Event: A Novel Score Representation for Symbolic Music
von: Qian, Lekai, et al.
Veröffentlicht: (2026)
von: Qian, Lekai, et al.
Veröffentlicht: (2026)
SLAP: Learning Speaker and Health-Related Representations from Natural Language Supervision
von: Ando, Angelika, et al.
Veröffentlicht: (2025)
von: Ando, Angelika, et al.
Veröffentlicht: (2025)
MusicAOG: an Energy-Based Model for Learning and Sampling a Hierarchical Representation of Symbolic Music
von: Qian, Yikai, et al.
Veröffentlicht: (2024)
von: Qian, Yikai, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Self-Supervised Multi-View Learning for Disentangled Music Audio Representations
von: Wilkins, Julia, et al.
Veröffentlicht: (2024) -
Do Music Source Separation Models Preserve Spatial Information in Binaural Audio?
von: Namballa, Richa, et al.
Veröffentlicht: (2025) -
Learning Multidimensional Disentangled Representations of Instrumental Sounds for Musical Similarity Assessment
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024) -
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
von: Shi, Runwu, et al.
Veröffentlicht: (2024) -
SONIQUE: Video Background Music Generation Using Unpaired Audio-Visual Data
von: Zhang, Liqian, et al.
Veröffentlicht: (2024)