YuE: Scaling Open Foundation Models for Long-Form Music Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Ruibin, Lin, Hanfeng, Guo, Shuyue, Zhang, Ge, Pan, Jiahao, Zang, Yongyi, Liu, Haohe, Liang, Yiming, Ma, Wenye, Du, Xingjian, Du, Xinrun, Ye, Zhen, Zheng, Tianyu, Jiang, Zhengxuan, Ma, Yinghao, Liu, Minghao, Tian, Zeyue, Zhou, Ziya, Xue, Liumeng, Qu, Xingwei, Li, Yizhi, Wu, Shangda, Shen, Tianhao, Ma, Ziyang, Zhan, Jun, Wang, Chunhui, Wang, Yatian, Chi, Xiaowei, Zhang, Xinyue, Yang, Zhenzhu, Wang, Xiangzhou, Liu, Shansong, Mei, Lingrui, Li, Peng, Wang, Junjie, Yu, Jianwei, Pang, Guojian, Li, Xu, Wang, Zihao, Zhou, Xiaohuan, Yu, Lijun, Benetos, Emmanouil, Chen, Yong, Lin, Chenghua, Chen, Xie, Xia, Gus, Zhang, Zhaoxiang, Zhang, Chao, Chen, Wenhu, Zhou, Xinyu, Qiu, Xipeng, Dannenberg, Roger, Liu, Jiaheng, Yang, Jian, Huang, Wenhao, Xue, Wei, Tan, Xu, Guo, Yike |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Audio-FLAN: A Preliminary Release
by: Xue, Liumeng, et al.
Published: (2025)
by: Xue, Liumeng, et al.
Published: (2025)
ChatMusician: Understanding and Generating Music Intrinsically with LLM
by: Yuan, Ruibin, et al.
Published: (2024)
by: Yuan, Ruibin, et al.
Published: (2024)
AudioX: A Unified Framework for Anything-to-Audio Generation
by: Tian, Zeyue, et al.
Published: (2025)
by: Tian, Zeyue, et al.
Published: (2025)
Sweat and Deformation‐Resistance Graphite/PVDF/PANI‐Based Temperature Sensor for Real‐Time Body Temperature Monitoring
by: Chen Zhang, et al.
Published: (2024)
by: Chen Zhang, et al.
Published: (2024)
Inference-time Scaling for Diffusion-based Audio Super-resolution
by: Jin, Yizhu, et al.
Published: (2025)
by: Jin, Yizhu, et al.
Published: (2025)
InfoLaw: Information Scaling Laws for Large Language Models with Quality-Weighted Mixture Data and Repetition
by: Liu, Fengze, et al.
Published: (2026)
by: Liu, Fengze, et al.
Published: (2026)
Scaling Law for Quantization-Aware Training
by: Chen, Mengzhao, et al.
Published: (2025)
by: Chen, Mengzhao, et al.
Published: (2025)
S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
by: Ma, Ruotian, et al.
Published: (2025)
by: Ma, Ruotian, et al.
Published: (2025)
AirHunt: Bridging VLM Semantics and Continuous Planning for Efficient Aerial Object Navigation
by: Chen, Xuecheng, et al.
Published: (2026)
by: Chen, Xuecheng, et al.
Published: (2026)
An Initial Investigation of Neural Replay Simulator for Over-the-Air Adversarial Perturbations to Automatic Speaker Verification
by: Li, Jiaqi, et al.
Published: (2023)
by: Li, Jiaqi, et al.
Published: (2023)
S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models
by: Jiang, Feng, et al.
Published: (2025)
by: Jiang, Feng, et al.
Published: (2025)
MusiXQA: Advancing Visual Music Understanding in Multimodal Large Language Models
by: Chen, Jian, et al.
Published: (2025)
by: Chen, Jian, et al.
Published: (2025)
Transfer the linguistic representations from TTS to accent conversion with non-parallel data
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Synthesis of Homochiral N‐Heterocyclic Carbene‐Based Nanosheets for Enhanced Asymmetric Catalysis
by: Xinchao Wang, et al.
Published: (2024)
by: Xinchao Wang, et al.
Published: (2024)
Can LLMs "Reason" in Music? An Evaluation of LLMs' Capability of Music Understanding and Generation
by: Zhou, Ziya, et al.
Published: (2024)
by: Zhou, Ziya, et al.
Published: (2024)
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild
by: Qi, Xingqun, et al.
Published: (2024)
by: Qi, Xingqun, et al.
Published: (2024)
Pericardial effusion: Unusual immunohistochemical expression
by: Ying Liu, et al.
Published: (2024)
by: Ying Liu, et al.
Published: (2024)
Multimodal Fish Feeding Intensity Assessment in Aquaculture
by: Cui, Meng, et al.
Published: (2023)
by: Cui, Meng, et al.
Published: (2023)
Collaborative Performance Prediction for Large Language Models
by: Zhang, Qiyuan, et al.
Published: (2024)
by: Zhang, Qiyuan, et al.
Published: (2024)
Zero-Shot Audio Captioning Using Soft and Hard Prompts
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
MuPT: A Generative Symbolic Music Pretrained Transformer
by: Qu, Xingwei, et al.
Published: (2024)
by: Qu, Xingwei, et al.
Published: (2024)
MathMixup: Boosting LLM Mathematical Reasoning with Difficulty-Controllable Data Synthesis and Curriculum Learning
by: Li, Xuchen, et al.
Published: (2026)
by: Li, Xuchen, et al.
Published: (2026)
Automatic Melody Reduction via Shortest Path Finding
by: Wang, Ziyu, et al.
Published: (2025)
by: Wang, Ziyu, et al.
Published: (2025)
Preserving Full Degradation Details for Blind Image Super-Resolution
by: Liu, Hongda, et al.
Published: (2024)
by: Liu, Hongda, et al.
Published: (2024)
SIRT1 as a potential therapeutic target in pelvic organ prolapse due to protective effects against oxidative stress and cellular senescence in human uterosacral ligament fibroblasts
by: Xinyi Wang, et al.
Published: (2024)
by: Xinyi Wang, et al.
Published: (2024)
White/Sun/Red‐Light Mediated Oxidative Cyclization from Formyl Acetates: Application to Ester Substituted N ‐Doped PAHs
by: Haochen Liu, et al.
Published: (2026)
by: Haochen Liu, et al.
Published: (2026)
Joint Geometric and Trajectory Consistency Learning for One-Step Real-World Super-Resolution
by: Deng, Chengyan, et al.
Published: (2026)
by: Deng, Chengyan, et al.
Published: (2026)
Co$^{3}$Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion
by: Qi, Xingqun, et al.
Published: (2025)
by: Qi, Xingqun, et al.
Published: (2025)
TiKMiX: Take Data Influence into Dynamic Mixture for Language Model Pre-training
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Dietary rutin improves the antidiarrheal capacity of weaned piglets by improving intestinal barrier function, antioxidant capacity and cecal microbiota composition
by: Longfei Ma, et al.
Published: (2024)
by: Longfei Ma, et al.
Published: (2024)
Deep Hashing with Semantic Hash Centers for Image Retrieval
by: Chen, Li, et al.
Published: (2025)
by: Chen, Li, et al.
Published: (2025)
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
by: Zhuo, Le, et al.
Published: (2023)
by: Zhuo, Le, et al.
Published: (2023)
Dissolution of Primary Carbides and Formation and Healing of Kirkendall Voids in Bearing Steel under Pulsed Electric Current
by: Zhongxue Wang, et al.
Published: (2024)
by: Zhongxue Wang, et al.
Published: (2024)
Residual Diffusion Bridge Model for Image Restoration
by: Wang, Hebaixu, et al.
Published: (2025)
by: Wang, Hebaixu, et al.
Published: (2025)
Inside Back Cover: New Transparent Rare‐Earth‐Based Hybrid Glasses: Synthesis, Luminescence, and X‐Ray Imaging Application
by: Jun Wei, et al.
Published: (2025)
by: Jun Wei, et al.
Published: (2025)
New Transparent Rare‐Earth‐Based Hybrid Glasses: Synthesis, Luminescence, and X‐Ray Imaging Application
by: Jun Wei, et al.
Published: (2025)
by: Jun Wei, et al.
Published: (2025)
MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
by: Li, Yizhi, et al.
Published: (2023)
by: Li, Yizhi, et al.
Published: (2023)
SingVisio: Visual Analytics of Diffusion Model for Singing Voice Conversion
by: Xue, Liumeng, et al.
Published: (2024)
by: Xue, Liumeng, et al.
Published: (2024)
Photothermal Conversion of the Oleophilic PVDF/Ti 3 C 2 T x Porous Foam Enables Non‐Aqueous Liquid System Applicable Actuator
by: Ruoqi Chen, et al.
Published: (2024)
by: Ruoqi Chen, et al.
Published: (2024)
ZK-Value: A Practical Zero-Knowledge System for Verifiable Data Valuation
by: Wang, Zhaoyu, et al.
Published: (2026)
by: Wang, Zhaoyu, et al.
Published: (2026)
Similar Items
-
Audio-FLAN: A Preliminary Release
by: Xue, Liumeng, et al.
Published: (2025) -
ChatMusician: Understanding and Generating Music Intrinsically with LLM
by: Yuan, Ruibin, et al.
Published: (2024) -
AudioX: A Unified Framework for Anything-to-Audio Generation
by: Tian, Zeyue, et al.
Published: (2025) -
Sweat and Deformation‐Resistance Graphite/PVDF/PANI‐Based Temperature Sensor for Real‐Time Body Temperature Monitoring
by: Chen Zhang, et al.
Published: (2024) -
Inference-time Scaling for Diffusion-based Audio Super-resolution
by: Jin, Yizhu, et al.
Published: (2025)