Training-Efficient Text-to-Music Generation with State-Space Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Wei-Jaw, Hsieh, Fang-Chih, Chen, Xuanjun, Tsai, Fang-Duo, Yang, Yi-Hsuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring State-Space-Model based Language Model in Music Generation
von: Lee, Wei-Jaw, et al.
Veröffentlicht: (2025)
von: Lee, Wei-Jaw, et al.
Veröffentlicht: (2025)
Audio Prompt Adapter: Unleashing Music Editing Abilities for Text-to-Music with Lightweight Finetuning
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2024)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2024)
MuseControlLite: Multifunctional Music Generation with Lightweight Conditioners
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2025)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2025)
Academic Text-to-Music Grand Challenge: Datasets, Baselines, and Evaluation Methods
von: Hsieh, Fang-Chih, et al.
Veröffentlicht: (2026)
von: Hsieh, Fang-Chih, et al.
Veröffentlicht: (2026)
AnchorSteer: Self-Discovered Concept Injection for Structure-Preserving Music Editing
von: Chang, Chih-Heng, et al.
Veröffentlicht: (2026)
von: Chang, Chih-Heng, et al.
Veröffentlicht: (2026)
Segment-Factorized Full-Song Generation on Symbolic Piano Music
von: Chen, Ping-Yi, et al.
Veröffentlicht: (2025)
von: Chen, Ping-Yi, et al.
Veröffentlicht: (2025)
MusiConGen: Rhythm and Chord Control for Transformer-Based Text-to-Music Generation
von: Lan, Yun-Han, et al.
Veröffentlicht: (2024)
von: Lan, Yun-Han, et al.
Veröffentlicht: (2024)
Emotion-driven Piano Music Generation via Two-stage Disentanglement and Functional Representation
von: Huang, Jingyue, et al.
Veröffentlicht: (2024)
von: Huang, Jingyue, et al.
Veröffentlicht: (2024)
HNote: Extending YNote with Hexadecimal Encoding for Fine-Tuning LLMs in Music Modeling
von: Chu, Hung-Ying, et al.
Veröffentlicht: (2025)
von: Chu, Hung-Ying, et al.
Veröffentlicht: (2025)
Efficient Long-Sequence Diffusion Modeling for Symbolic Music Generation
von: Xu, Jinhan, et al.
Veröffentlicht: (2026)
von: Xu, Jinhan, et al.
Veröffentlicht: (2026)
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
Efficient Training for Cross-lingual Speech Language Models
von: Zhou, Yan, et al.
Veröffentlicht: (2026)
von: Zhou, Yan, et al.
Veröffentlicht: (2026)
Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators
von: Novack, Zachary, et al.
Veröffentlicht: (2026)
von: Novack, Zachary, et al.
Veröffentlicht: (2026)
Music Arena: Live Evaluation for Text-to-Music
von: Kim, Yonghyun, et al.
Veröffentlicht: (2025)
von: Kim, Yonghyun, et al.
Veröffentlicht: (2025)
LoopGen: Training-Free Loopable Music Generation
von: Marincione, Davide, et al.
Veröffentlicht: (2025)
von: Marincione, Davide, et al.
Veröffentlicht: (2025)
MusicFlow: Cascaded Flow Matching for Text Guided Music Generation
von: Prajwal, K R, et al.
Veröffentlicht: (2024)
von: Prajwal, K R, et al.
Veröffentlicht: (2024)
MOSA: Music Motion with Semantic Annotation Dataset for Cross-Modal Music Processing
von: Huang, Yu-Fen, et al.
Veröffentlicht: (2024)
von: Huang, Yu-Fen, et al.
Veröffentlicht: (2024)
The Interpretation Gap in Text-to-Music Generation Models
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
Composer Vector: Style-steering Symbolic Music Generation in a Latent Space
von: Jiang, Xunyi, et al.
Veröffentlicht: (2026)
von: Jiang, Xunyi, et al.
Veröffentlicht: (2026)
TimberAgent: Gram-Guided Retrieval for Executable Music Effect Control
von: He, Shihao, et al.
Veröffentlicht: (2026)
von: He, Shihao, et al.
Veröffentlicht: (2026)
Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding
von: Wang, Tsai-Ning, et al.
Veröffentlicht: (2025)
von: Wang, Tsai-Ning, et al.
Veröffentlicht: (2025)
SyncTrack: Rhythmic Stability and Synchronization in Multi-Track Music Generation
von: Wang, Hongrui, et al.
Veröffentlicht: (2026)
von: Wang, Hongrui, et al.
Veröffentlicht: (2026)
CompLex: Music Theory Lexicon Constructed by Autonomous Agents for Automatic Music Generation
von: Hu, Zhejing, et al.
Veröffentlicht: (2025)
von: Hu, Zhejing, et al.
Veröffentlicht: (2025)
Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation
von: Cheng, Yuqing, et al.
Veröffentlicht: (2026)
von: Cheng, Yuqing, et al.
Veröffentlicht: (2026)
Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores
von: Dai, Congren, et al.
Veröffentlicht: (2025)
von: Dai, Congren, et al.
Veröffentlicht: (2025)
Amadeus: Autoregressive Model with Bidirectional Attribute Modelling for Symbolic Music
von: Su, Hongju, et al.
Veröffentlicht: (2025)
von: Su, Hongju, et al.
Veröffentlicht: (2025)
NotaGen: Advancing Musicality in Symbolic Music Generation with Large Language Model Training Paradigms
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
Learning Frame-Wise Emotion Intensity for Audio-Driven Talking-Head Generation
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation
von: Lu, Shao-Chien, et al.
Veröffentlicht: (2025)
von: Lu, Shao-Chien, et al.
Veröffentlicht: (2025)
MusicSynth: An Automated Pipeline for Generating Violin Fingerboard Animations from Sheet Music Using Optical Music Recognition
von: Kaushik, Abhimanyu
Veröffentlicht: (2026)
von: Kaushik, Abhimanyu
Veröffentlicht: (2026)
AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models
von: Kang, Mintong, et al.
Veröffentlicht: (2026)
von: Kang, Mintong, et al.
Veröffentlicht: (2026)
Practical and Reproducible Symbolic Music Generation by Large Language Models with Structural Embeddings
von: Rhyu, Seungyeon, et al.
Veröffentlicht: (2024)
von: Rhyu, Seungyeon, et al.
Veröffentlicht: (2024)
TMD-Bench: A Multi-Level Evaluation Paradigm for Music-Dance Co-Generation
von: Yang, Xiaoda, et al.
Veröffentlicht: (2026)
von: Yang, Xiaoda, et al.
Veröffentlicht: (2026)
Story2MIDI: Emotionally Aligned Music Generation from Text
von: Shokri, Mohammad, et al.
Veröffentlicht: (2025)
von: Shokri, Mohammad, et al.
Veröffentlicht: (2025)
MusicMagus: Zero-Shot Text-to-Music Editing via Diffusion Models
von: Zhang, Yixiao, et al.
Veröffentlicht: (2024)
von: Zhang, Yixiao, et al.
Veröffentlicht: (2024)
Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation
von: Liu, Jiafeng, et al.
Veröffentlicht: (2026)
von: Liu, Jiafeng, et al.
Veröffentlicht: (2026)
Generating Separated Singing Vocals Using a Diffusion Model Conditioned on Music Mixtures
von: Plaja-Roglans, Genís, et al.
Veröffentlicht: (2025)
von: Plaja-Roglans, Genís, et al.
Veröffentlicht: (2025)
Survey on the Evaluation of Generative Models in Music
von: Lerch, Alexander, et al.
Veröffentlicht: (2025)
von: Lerch, Alexander, et al.
Veröffentlicht: (2025)
Music Style Transfer With Diffusion Model
von: Huang, Hong, et al.
Veröffentlicht: (2024)
von: Huang, Hong, et al.
Veröffentlicht: (2024)
ABC-Eval: Benchmarking Large Language Models on Symbolic Music Understanding and Instruction Following
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Exploring State-Space-Model based Language Model in Music Generation
von: Lee, Wei-Jaw, et al.
Veröffentlicht: (2025) -
Audio Prompt Adapter: Unleashing Music Editing Abilities for Text-to-Music with Lightweight Finetuning
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2024) -
MuseControlLite: Multifunctional Music Generation with Lightweight Conditioners
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2025) -
Academic Text-to-Music Grand Challenge: Datasets, Baselines, and Evaluation Methods
von: Hsieh, Fang-Chih, et al.
Veröffentlicht: (2026) -
AnchorSteer: Self-Discovered Concept Injection for Structure-Preserving Music Editing
von: Chang, Chih-Heng, et al.
Veröffentlicht: (2026)