Practical and Reproducible Symbolic Music Generation by Large Language Models with Structural Embeddings
Fuente:
arXiv
Guardado en:
| Autores principales: | Rhyu, Seungyeon, Yang, Kichang, Cho, Sungjun, Kim, Jaehyeon, Lee, Kyogu, Lee, Moontae |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Music De-limiter Networks via Sample-wise Gain Inversion
por: Jeon, Chang-Bin, et al.
Publicado: (2023)
por: Jeon, Chang-Bin, et al.
Publicado: (2023)
CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech
por: Kim, Jaehyeon, et al.
Publicado: (2024)
por: Kim, Jaehyeon, et al.
Publicado: (2024)
Do Captioning Metrics Reflect Music Semantic Alignment?
por: Lee, Jinwoo, et al.
Publicado: (2024)
por: Lee, Jinwoo, et al.
Publicado: (2024)
Wavespace: A Highly Explorable Wavetable Generator
por: Lee, Hazounne, et al.
Publicado: (2024)
por: Lee, Hazounne, et al.
Publicado: (2024)
MGE-LDM: Joint Latent Diffusion for Simultaneous Music Generation and Source Extraction
por: Chae, Yunkee, et al.
Publicado: (2025)
por: Chae, Yunkee, et al.
Publicado: (2025)
DOSE : Drum One-Shot Extraction from Music Mixture
por: Hwang, Suntae, et al.
Publicado: (2025)
por: Hwang, Suntae, et al.
Publicado: (2025)
Vo-Ve: An Explainable Voice-Vector for Speaker Identity Evaluation
por: Lee, Jaejun, et al.
Publicado: (2025)
por: Lee, Jaejun, et al.
Publicado: (2025)
Few-step Adversarial Schrödinger Bridge for Generative Speech Enhancement
por: Han, Seungu, et al.
Publicado: (2025)
por: Han, Seungu, et al.
Publicado: (2025)
Towards Bitrate-Efficient and Noise-Robust Speech Coding with Variable Bitrate RVQ
por: Chae, Yunkee, et al.
Publicado: (2025)
por: Chae, Yunkee, et al.
Publicado: (2025)
Rethinking Speech Representation Aggregation in Speech Enhancement: A Phonetic Mutual Information Perspective
por: Han, Seungu, et al.
Publicado: (2026)
por: Han, Seungu, et al.
Publicado: (2026)
Music Auto-Tagging with Robust Music Representation Learned via Domain Adversarial Training
por: Joung, Haesun, et al.
Publicado: (2024)
por: Joung, Haesun, et al.
Publicado: (2024)
Flexible Control in Symbolic Music Generation via Musical Metadata
por: Han, Sangjun, et al.
Publicado: (2024)
por: Han, Sangjun, et al.
Publicado: (2024)
Musical Word Embedding for Music Tagging and Retrieval
por: Doh, SeungHeon, et al.
Publicado: (2024)
por: Doh, SeungHeon, et al.
Publicado: (2024)
Inverse Nonlinearity Compensation of Hyperelastic Deformation in Dielectric Elastomer for Acoustic Actuation
por: Lee, Jin Woo, et al.
Publicado: (2024)
por: Lee, Jin Woo, et al.
Publicado: (2024)
Reverse Engineering of Music Mixing Graphs with Differentiable Processors and Iterative Pruning
por: Lee, Sungho, et al.
Publicado: (2025)
por: Lee, Sungho, et al.
Publicado: (2025)
Learning Semantic Information from Raw Audio Signal Using Both Contextual and Phonetic Representations
por: Kim, Jaeyeon, et al.
Publicado: (2024)
por: Kim, Jaeyeon, et al.
Publicado: (2024)
Removing Speaker Information from Speech Representation using Variable-Length Soft Pooling
por: Hwang, Injune, et al.
Publicado: (2024)
por: Hwang, Injune, et al.
Publicado: (2024)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
por: Yi, Jayeon, et al.
Publicado: (2024)
por: Yi, Jayeon, et al.
Publicado: (2024)
String Sound Synthesizer on GPU-accelerated Finite Difference Scheme
por: Lee, Jin Woo, et al.
Publicado: (2023)
por: Lee, Jin Woo, et al.
Publicado: (2023)
NotaGen: Advancing Musicality in Symbolic Music Generation with Large Language Model Training Paradigms
por: Wang, Yashan, et al.
Publicado: (2025)
por: Wang, Yashan, et al.
Publicado: (2025)
Generating Symbolic Music from Natural Language Prompts using an LLM-Enhanced Dataset
por: Xu, Weihan, et al.
Publicado: (2024)
por: Xu, Weihan, et al.
Publicado: (2024)
SymPAC: Scalable Symbolic Music Generation With Prompts And Constraints
por: Chen, Haonan, et al.
Publicado: (2024)
por: Chen, Haonan, et al.
Publicado: (2024)
Language Models for Music Medicine Generation
por: Nikolakakis, Emmanouil, et al.
Publicado: (2024)
por: Nikolakakis, Emmanouil, et al.
Publicado: (2024)
Automatic Music Mixing using a Generative Model of Effect Embeddings
por: Moliner, Eloi, et al.
Publicado: (2025)
por: Moliner, Eloi, et al.
Publicado: (2025)
METEOR: Melody-aware Texture-controllable Symbolic Orchestral Music Generation via Transformer VAE
por: Le, Dinh-Viet-Toan, et al.
Publicado: (2024)
por: Le, Dinh-Viet-Toan, et al.
Publicado: (2024)
Joint Audio and Symbolic Conditioning for Temporally Controlled Text-to-Music Generation
por: Tal, Or, et al.
Publicado: (2024)
por: Tal, Or, et al.
Publicado: (2024)
Large Language Models: From Notes to Musical Form
por: Atassi, Lilac
Publicado: (2024)
por: Atassi, Lilac
Publicado: (2024)
Steer-by-prior Editing of Symbolic Music Loops
por: Jonason, Nicolas, et al.
Publicado: (2024)
por: Jonason, Nicolas, et al.
Publicado: (2024)
GRAFX: An Open-Source Library for Audio Processing Graphs in PyTorch
por: Lee, Sungho, et al.
Publicado: (2024)
por: Lee, Sungho, et al.
Publicado: (2024)
Guiding Frame-Level CTC Alignments Using Self-knowledge Distillation
por: Kim, Eungbeom, et al.
Publicado: (2024)
por: Kim, Eungbeom, et al.
Publicado: (2024)
SALM: Spatial Audio Language Model with Structured Embeddings for Understanding and Editing
por: Hu, Jinbo, et al.
Publicado: (2025)
por: Hu, Jinbo, et al.
Publicado: (2025)
MeloTrans: A Text to Symbolic Music Generation Model Following Human Composition Habit
por: Wang, Yutian, et al.
Publicado: (2024)
por: Wang, Yutian, et al.
Publicado: (2024)
Exploring State-Space-Model based Language Model in Music Generation
por: Lee, Wei-Jaw, et al.
Publicado: (2025)
por: Lee, Wei-Jaw, et al.
Publicado: (2025)
SMUG-Explain: A Framework for Symbolic Music Graph Explanations
por: Karystinaios, Emmanouil, et al.
Publicado: (2024)
por: Karystinaios, Emmanouil, et al.
Publicado: (2024)
Pianoroll-Event: A Novel Score Representation for Symbolic Music
por: Qian, Lekai, et al.
Publicado: (2026)
por: Qian, Lekai, et al.
Publicado: (2026)
Large-Scale Training Data Attribution for Music Generative Models via Unlearning
por: Choi, Woosung, et al.
Publicado: (2025)
por: Choi, Woosung, et al.
Publicado: (2025)
Differentiable Acoustic Radiance Transfer
por: Lee, Sungho, et al.
Publicado: (2025)
por: Lee, Sungho, et al.
Publicado: (2025)
Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion Simulation
por: Lee, Jin Woo, et al.
Publicado: (2024)
por: Lee, Jin Woo, et al.
Publicado: (2024)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
por: Kim, Taewoo, et al.
Publicado: (2024)
por: Kim, Taewoo, et al.
Publicado: (2024)
Optimizing Feature Extraction for Symbolic Music
por: Simonetta, Federico, et al.
Publicado: (2023)
por: Simonetta, Federico, et al.
Publicado: (2023)
Ejemplares similares
-
Music De-limiter Networks via Sample-wise Gain Inversion
por: Jeon, Chang-Bin, et al.
Publicado: (2023) -
CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech
por: Kim, Jaehyeon, et al.
Publicado: (2024) -
Do Captioning Metrics Reflect Music Semantic Alignment?
por: Lee, Jinwoo, et al.
Publicado: (2024) -
Wavespace: A Highly Explorable Wavetable Generator
por: Lee, Hazounne, et al.
Publicado: (2024) -
MGE-LDM: Joint Latent Diffusion for Simultaneous Music Generation and Source Extraction
por: Chae, Yunkee, et al.
Publicado: (2025)