ByteComposer: a Human-like Melody Composition Method based on Language Model Agent
Fuente:
arXiv
Salvato in:
| Autori principali: | Liang, Xia, Du, Xingjian, Lin, Jiaju, Zou, Pei, Wan, Yuan, Zhu, Bilei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SongComposer: A Large Language Model for Lyric and Melody Generation in Song Composition
di: Ding, Shuangrui, et al.
Pubblicazione: (2024)
di: Ding, Shuangrui, et al.
Pubblicazione: (2024)
MINT: Boosting Audio-Language Model via Multi-Target Pre-Training and Instruction Tuning
di: Zhao, Hang, et al.
Pubblicazione: (2024)
di: Zhao, Hang, et al.
Pubblicazione: (2024)
Automatic Melody Reduction via Shortest Path Finding
di: Wang, Ziyu, et al.
Pubblicazione: (2025)
di: Wang, Ziyu, et al.
Pubblicazione: (2025)
Editing Music with Melody and Text: Using ControlNet for Diffusion Transformer
di: Hou, Siyuan, et al.
Pubblicazione: (2024)
di: Hou, Siyuan, et al.
Pubblicazione: (2024)
Mel-RoFormer for Vocal Separation and Vocal Melody Transcription
di: Wang, Ju-Chiang, et al.
Pubblicazione: (2024)
di: Wang, Ju-Chiang, et al.
Pubblicazione: (2024)
RobustSVC: HuBERT-based Melody Extractor and Adversarial Learning for Robust Singing Voice Conversion
di: Chen, Wei, et al.
Pubblicazione: (2024)
di: Chen, Wei, et al.
Pubblicazione: (2024)
MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection
di: Lu, Tongyu, et al.
Pubblicazione: (2025)
di: Lu, Tongyu, et al.
Pubblicazione: (2025)
MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing
di: Wu, Shangda, et al.
Pubblicazione: (2024)
di: Wu, Shangda, et al.
Pubblicazione: (2024)
SymPAC: Scalable Symbolic Music Generation With Prompts And Constraints
di: Chen, Haonan, et al.
Pubblicazione: (2024)
di: Chen, Haonan, et al.
Pubblicazione: (2024)
Note-Level Singing Melody Transcription for Time-Aligned Musical Score Generation
di: Kim, Leekyung, et al.
Pubblicazione: (2025)
di: Kim, Leekyung, et al.
Pubblicazione: (2025)
Melody-Guided Music Generation
di: Wei, Shaopeng, et al.
Pubblicazione: (2024)
di: Wei, Shaopeng, et al.
Pubblicazione: (2024)
Audio Mamba: Pretrained Audio State Space Model For Audio Tagging
di: Lin, Jiaju, et al.
Pubblicazione: (2024)
di: Lin, Jiaju, et al.
Pubblicazione: (2024)
METEOR: Melody-aware Texture-controllable Symbolic Orchestral Music Generation via Transformer VAE
di: Le, Dinh-Viet-Toan, et al.
Pubblicazione: (2024)
di: Le, Dinh-Viet-Toan, et al.
Pubblicazione: (2024)
Small Tunes Transformer: Exploring Macro & Micro-Level Hierarchies for Skeleton-Conditioned Melody Generation
di: Lv, Yishan, et al.
Pubblicazione: (2024)
di: Lv, Yishan, et al.
Pubblicazione: (2024)
Exploring Tokenization Methods for Multitrack Sheet Music Generation
di: Wang, Yashan, et al.
Pubblicazione: (2024)
di: Wang, Yashan, et al.
Pubblicazione: (2024)
AudioComposer: Towards Fine-grained Audio Generation with Natural Language Descriptions
di: Wang, Yuanyuan, et al.
Pubblicazione: (2024)
di: Wang, Yuanyuan, et al.
Pubblicazione: (2024)
Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints
di: Meng, Hao, et al.
Pubblicazione: (2026)
di: Meng, Hao, et al.
Pubblicazione: (2026)
YingMusic-Singer-Plus: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance
di: Hao, Chunbo, et al.
Pubblicazione: (2026)
di: Hao, Chunbo, et al.
Pubblicazione: (2026)
REFFLY: Melody-Constrained Lyrics Editing Model
di: Zhao, Songyan, et al.
Pubblicazione: (2024)
di: Zhao, Songyan, et al.
Pubblicazione: (2024)
Accompanied Singing Voice Synthesis with Fully Text-controlled Melody
di: Li, Ruiqi, et al.
Pubblicazione: (2024)
di: Li, Ruiqi, et al.
Pubblicazione: (2024)
Joint Learning of Wording and Formatting for Singable Melody-to-Lyric Generation
di: Ou, Longshen, et al.
Pubblicazione: (2023)
di: Ou, Longshen, et al.
Pubblicazione: (2023)
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
di: Zhao, PengYuan, et al.
Pubblicazione: (2024)
di: Zhao, PengYuan, et al.
Pubblicazione: (2024)
Emotion-Driven Melody Harmonization via Melodic Variation and Functional Representation
di: Huang, Jingyue, et al.
Pubblicazione: (2024)
di: Huang, Jingyue, et al.
Pubblicazione: (2024)
MPO: Multidimensional Preference Optimization for Language Model-based Text-to-Speech
di: Xia, Kangxiang, et al.
Pubblicazione: (2025)
di: Xia, Kangxiang, et al.
Pubblicazione: (2025)
Melody predominates over harmony in the evolution of musical scales across 96 countries
di: McBride, John M, et al.
Pubblicazione: (2024)
di: McBride, John M, et al.
Pubblicazione: (2024)
Vocal Melody Construction for Persian Lyrics Using LSTM Recurrent Neural Networks
di: Jafari, Farshad, et al.
Pubblicazione: (2024)
di: Jafari, Farshad, et al.
Pubblicazione: (2024)
"It is okay to be uncommon": Quantizing Sound Event Detection Networks on Hardware Accelerators with Uncommon Sub-Byte Support
di: Wu, Yushu, et al.
Pubblicazione: (2024)
di: Wu, Yushu, et al.
Pubblicazione: (2024)
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
di: Chen, Wei, et al.
Pubblicazione: (2025)
di: Chen, Wei, et al.
Pubblicazione: (2025)
ComposerX: Multi-Agent Symbolic Music Composition with LLMs
di: Deng, Qixin, et al.
Pubblicazione: (2024)
di: Deng, Qixin, et al.
Pubblicazione: (2024)
Speech Enhancement with Overlapped-Frame Information Fusion and Causal Self-Attention
di: Zhang, Yuewei, et al.
Pubblicazione: (2025)
di: Zhang, Yuewei, et al.
Pubblicazione: (2025)
A Two-Stage Framework in Cross-Spectrum Domain for Real-Time Speech Enhancement
di: Zhang, Yuewei, et al.
Pubblicazione: (2024)
di: Zhang, Yuewei, et al.
Pubblicazione: (2024)
SongGLM: Lyric-to-Melody Generation with 2D Alignment Encoding and Multi-Task Pre-Training
di: Yu, Jiaxing, et al.
Pubblicazione: (2024)
di: Yu, Jiaxing, et al.
Pubblicazione: (2024)
A Mamba-based Network for Semi-supervised Singing Melody Extraction Using Confidence Binary Regularization
di: He, Xiaoliang, et al.
Pubblicazione: (2025)
di: He, Xiaoliang, et al.
Pubblicazione: (2025)
CoComposer: LLM Multi-agent Collaborative Music Composition
di: Xing, Peiwen, et al.
Pubblicazione: (2025)
di: Xing, Peiwen, et al.
Pubblicazione: (2025)
On the use of Performer and Agent Attention for Spoken Language Identification
di: dhiman, Jitendra Kumar, et al.
Pubblicazione: (2025)
di: dhiman, Jitendra Kumar, et al.
Pubblicazione: (2025)
PhoenixCodec: Taming Neural Speech Coding for Extreme Low-Resource Scenarios
di: Wan, Zixiang, et al.
Pubblicazione: (2025)
di: Wan, Zixiang, et al.
Pubblicazione: (2025)
Unsupervised Composable Representations for Audio
di: Bindi, Giovanni, et al.
Pubblicazione: (2024)
di: Bindi, Giovanni, et al.
Pubblicazione: (2024)
SqueezeComposer: Temporal Speed-up is A Simple Trick for Long-form Music Composing
di: Chen, Jianyi, et al.
Pubblicazione: (2026)
di: Chen, Jianyi, et al.
Pubblicazione: (2026)
CLaMP 3: Universal Music Information Retrieval Across Unaligned Modalities and Unseen Languages
di: Wu, Shangda, et al.
Pubblicazione: (2025)
di: Wu, Shangda, et al.
Pubblicazione: (2025)
Integrating Text-to-Music Models with Language Models: Composing Long Structured Music Pieces
di: Atassi, Lilac
Pubblicazione: (2024)
di: Atassi, Lilac
Pubblicazione: (2024)
Documenti analoghi
-
SongComposer: A Large Language Model for Lyric and Melody Generation in Song Composition
di: Ding, Shuangrui, et al.
Pubblicazione: (2024) -
MINT: Boosting Audio-Language Model via Multi-Target Pre-Training and Instruction Tuning
di: Zhao, Hang, et al.
Pubblicazione: (2024) -
Automatic Melody Reduction via Shortest Path Finding
di: Wang, Ziyu, et al.
Pubblicazione: (2025) -
Editing Music with Melody and Text: Using ControlNet for Diffusion Transformer
di: Hou, Siyuan, et al.
Pubblicazione: (2024) -
Mel-RoFormer for Vocal Separation and Vocal Melody Transcription
di: Wang, Ju-Chiang, et al.
Pubblicazione: (2024)