Gespeichert in:
| Hauptverfasser: | Li, Jiajun, Xu, Tianze, Chen, Xuesong, Yao, Xinrui, Liu, Shuchang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2405.02801 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
M$^{2}$UGen: Multi-modal Music Understanding and Generation with the Power of Large Language Models
von: Liu, Shansong, et al.
Veröffentlicht: (2023)
von: Liu, Shansong, et al.
Veröffentlicht: (2023)
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
von: Liu, Shansong, et al.
Veröffentlicht: (2024)
von: Liu, Shansong, et al.
Veröffentlicht: (2024)
Large-Scale Training Data Attribution for Music Generative Models via Unlearning
von: Choi, Woosung, et al.
Veröffentlicht: (2025)
von: Choi, Woosung, et al.
Veröffentlicht: (2025)
Music Source Separation Based on a Lightweight Deep Learning Framework (DTTNET: DUAL-PATH TFC-TDF UNET)
von: Chen, Junyu, et al.
Veröffentlicht: (2023)
von: Chen, Junyu, et al.
Veröffentlicht: (2023)
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
Temporal Adaptation of Pre-trained Foundation Models for Music Structure Analysis
von: Zhang, Yixiao, et al.
Veröffentlicht: (2025)
von: Zhang, Yixiao, et al.
Veröffentlicht: (2025)
PianoBART: Symbolic Piano Music Generation and Understanding with Large-Scale Pre-Training
von: Liang, Xiao, et al.
Veröffentlicht: (2024)
von: Liang, Xiao, et al.
Veröffentlicht: (2024)
Multi-Track MusicLDM: Towards Versatile Music Generation with Latent Diffusion Model
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
NotaGen: Advancing Musicality in Symbolic Music Generation with Large Language Model Training Paradigms
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
From Generality to Mastery: Composer-Style Symbolic Music Generation via Large-Scale Pre-training
von: Yao, Mingyang, et al.
Veröffentlicht: (2025)
von: Yao, Mingyang, et al.
Veröffentlicht: (2025)
Melodia: Training-Free Music Editing Guided by Attention Probing in Diffusion Models
von: Yang, Yi, et al.
Veröffentlicht: (2025)
von: Yang, Yi, et al.
Veröffentlicht: (2025)
OMAR-RQ: Open Music Audio Representation Model Trained with Multi-Feature Masked Token Prediction
von: Alonso-Jiménez, Pablo, et al.
Veröffentlicht: (2025)
von: Alonso-Jiménez, Pablo, et al.
Veröffentlicht: (2025)
Simultaneous Music Separation and Generation Using Multi-Track Latent Diffusion Models
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
MusicEval: A Generative Music Dataset with Expert Ratings for Automatic Text-to-Music Evaluation
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
A Lightweight Slot-Attention Framework for Multi-Instrument Multi-Pitch Estimation
von: Taenzer, Michael
Veröffentlicht: (2026)
von: Taenzer, Michael
Veröffentlicht: (2026)
MuseControlLite: Multifunctional Music Generation with Lightweight Conditioners
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2025)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2025)
MusicMamba: A Dual-Feature Modeling Approach for Generating Chinese Traditional Music with Modal Precision
von: Chen, Jiatao, et al.
Veröffentlicht: (2024)
von: Chen, Jiatao, et al.
Veröffentlicht: (2024)
Language Models for Music Medicine Generation
von: Nikolakakis, Emmanouil, et al.
Veröffentlicht: (2024)
von: Nikolakakis, Emmanouil, et al.
Veröffentlicht: (2024)
Speaker Disentanglement of Speech Pre-trained Model Based on Interpretability
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
ACE-Step: A Step Towards Music Generation Foundation Model
von: Gong, Junmin, et al.
Veröffentlicht: (2025)
von: Gong, Junmin, et al.
Veröffentlicht: (2025)
FakeMusicCaps: a Dataset for Detection and Attribution of Synthetic Music Generated via Text-to-Music Models
von: Comanducci, Luca, et al.
Veröffentlicht: (2024)
von: Comanducci, Luca, et al.
Veröffentlicht: (2024)
Watermarking Training Data of Music Generation Models
von: Epple, Pascal, et al.
Veröffentlicht: (2024)
von: Epple, Pascal, et al.
Veröffentlicht: (2024)
MINT: Boosting Audio-Language Model via Multi-Target Pre-Training and Instruction Tuning
von: Zhao, Hang, et al.
Veröffentlicht: (2024)
von: Zhao, Hang, et al.
Veröffentlicht: (2024)
A Diffusion-Based Generative Equalizer for Music Restoration
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
Large Language Models: From Notes to Musical Form
von: Atassi, Lilac
Veröffentlicht: (2024)
von: Atassi, Lilac
Veröffentlicht: (2024)
Multi-modal Speech Enhancement with Limited Electromyography Channels
von: Feng, Fuyuan, et al.
Veröffentlicht: (2025)
von: Feng, Fuyuan, et al.
Veröffentlicht: (2025)
MuseBarControl: Enhancing Fine-Grained Control in Symbolic Music Generation through Pre-Training and Counterfactual Loss
von: Shu, Yangyang, et al.
Veröffentlicht: (2024)
von: Shu, Yangyang, et al.
Veröffentlicht: (2024)
Exploring Prediction Targets in Masked Pre-Training for Speech Foundation Models
von: Chen, Li-Wei, et al.
Veröffentlicht: (2024)
von: Chen, Li-Wei, et al.
Veröffentlicht: (2024)
Training a Perceptual Model for Evaluating Auditory Similarity in Music Adversarial Attack
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
Audio Prompt Adapter: Unleashing Music Editing Abilities for Text-to-Music with Lightweight Finetuning
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2024)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2024)
Adapter-Based Multi-Agent AVSR Extension for Pre-Trained ASR Models
von: Simic, Christopher, et al.
Veröffentlicht: (2025)
von: Simic, Christopher, et al.
Veröffentlicht: (2025)
Improving Controllability and Editability for Pretrained Text-to-Music Generation Models
von: Zhang, Yixiao
Veröffentlicht: (2024)
von: Zhang, Yixiao
Veröffentlicht: (2024)
The Arrow of Time in Music -- Revisiting the Temporal Structure of Music with Distinguishability and Unique Orientability as the Anchor Point
von: Xu, Qi
Veröffentlicht: (2023)
von: Xu, Qi
Veröffentlicht: (2023)
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
MMGER: Multi-modal and Multi-granularity Generative Error Correction with LLM for Joint Accent and Speech Recognition
von: Mu, Bingshen, et al.
Veröffentlicht: (2024)
von: Mu, Bingshen, et al.
Veröffentlicht: (2024)
Let Network Decide What to Learn: Symbolic Music Understanding Model Based on Large-scale Adversarial Pre-training
von: Zhao, Zijian
Veröffentlicht: (2024)
von: Zhao, Zijian
Veröffentlicht: (2024)
MusicAOG: an Energy-Based Model for Learning and Sampling a Hierarchical Representation of Symbolic Music
von: Qian, Yikai, et al.
Veröffentlicht: (2024)
von: Qian, Yikai, et al.
Veröffentlicht: (2024)
Multi-Source Music Generation with Latent Diffusion
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
M$^{2}$UGen: Multi-modal Music Understanding and Generation with the Power of Large Language Models
von: Liu, Shansong, et al.
Veröffentlicht: (2023) -
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
von: Liu, Shansong, et al.
Veröffentlicht: (2024) -
Large-Scale Training Data Attribution for Music Generative Models via Unlearning
von: Choi, Woosung, et al.
Veröffentlicht: (2025) -
Music Source Separation Based on a Lightweight Deep Learning Framework (DTTNET: DUAL-PATH TFC-TDF UNET)
von: Chen, Junyu, et al.
Veröffentlicht: (2023) -
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
von: Bai, Ye, et al.
Veröffentlicht: (2024)