ACE-Step 1.5: Pushing the Boundaries of Open-Source Music Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gong, Junmin, Song, Yulin, Zhao, Wenxiao, Wang, Sen, Xu, Shengyuan, Guo, Jing, Yang, Xuerui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ACE-Step: A Step Towards Music Generation Foundation Model
von: Gong, Junmin, et al.
Veröffentlicht: (2025)
von: Gong, Junmin, et al.
Veröffentlicht: (2025)
ConSinger: Efficient High-Fidelity Singing Voice Generation with Minimal Steps
von: Song, Yulin, et al.
Veröffentlicht: (2024)
von: Song, Yulin, et al.
Veröffentlicht: (2024)
HeartMuLa: A Family of Open Sourced Music Foundation Models
von: Yang, Dongchao, et al.
Veröffentlicht: (2026)
von: Yang, Dongchao, et al.
Veröffentlicht: (2026)
Amphion: An Open-Source Audio, Music and Speech Generation Toolkit
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023)
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023)
U3-xi: Pushing the Boundaries of Speaker Recognition by Incorporating Uncertainty
von: Li, Junjie, et al.
Veröffentlicht: (2026)
von: Li, Junjie, et al.
Veröffentlicht: (2026)
OpenACE: An Open Benchmark for Evaluating Audio Coding Performance
von: Coldenhoff, Jozef, et al.
Veröffentlicht: (2024)
von: Coldenhoff, Jozef, et al.
Veröffentlicht: (2024)
Multi-Source Music Generation with Latent Diffusion
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
CtrSVDD: A Benchmark Dataset and Baseline Analysis for Controlled Singing Voice Deepfake Detection
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
ACMID: Automatic Curation of Musical Instrument Dataset for 7-Stem Music Source Separation
von: Yu, Ji, et al.
Veröffentlicht: (2025)
von: Yu, Ji, et al.
Veröffentlicht: (2025)
DSFlow: Dual Supervision and Step-Aware Architecture for One-Step Flow Matching Speech Synthesis
von: Lin, Bin, et al.
Veröffentlicht: (2026)
von: Lin, Bin, et al.
Veröffentlicht: (2026)
MSRBench: A Benchmarking Dataset for Music Source Restoration
von: Zang, Yongyi, et al.
Veröffentlicht: (2025)
von: Zang, Yongyi, et al.
Veröffentlicht: (2025)
Video Echoed in Music: Semantic, Temporal, and Rhythmic Alignment for Video-to-Music Generation
von: Tong, Xinyi, et al.
Veröffentlicht: (2025)
von: Tong, Xinyi, et al.
Veröffentlicht: (2025)
CompLex: Music Theory Lexicon Constructed by Autonomous Agents for Automatic Music Generation
von: Hu, Zhejing, et al.
Veröffentlicht: (2025)
von: Hu, Zhejing, et al.
Veröffentlicht: (2025)
MusicEval: A Generative Music Dataset with Expert Ratings for Automatic Text-to-Music Evaluation
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps
von: Qian, Lekai, et al.
Veröffentlicht: (2026)
von: Qian, Lekai, et al.
Veröffentlicht: (2026)
Source Separation for A Cappella Music
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
SLAM-LLM: A Modular, Open-Source Multimodal Large Language Model Framework and Best Practice for Speech, Language, Audio and Music Processing
von: Ma, Ziyang, et al.
Veröffentlicht: (2026)
von: Ma, Ziyang, et al.
Veröffentlicht: (2026)
Efficient Long-Sequence Diffusion Modeling for Symbolic Music Generation
von: Xu, Jinhan, et al.
Veröffentlicht: (2026)
von: Xu, Jinhan, et al.
Veröffentlicht: (2026)
Diffusion-based Symbolic Music Generation with Structured State Space Models
von: Yuan, Shenghua, et al.
Veröffentlicht: (2025)
von: Yuan, Shenghua, et al.
Veröffentlicht: (2025)
MAGE: Modality-Agnostic Music Generation and Editing
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
MusicDET: Zero-Shot AI-Generated Music Detection
von: Han, Chaolei, et al.
Veröffentlicht: (2026)
von: Han, Chaolei, et al.
Veröffentlicht: (2026)
Streaming Generation for Music Accompaniment
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
Score-informed Music Source Separation: Improving Synthetic-to-real Generalization in Classical Music
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
Presto! Distilling Steps and Layers for Accelerating Music Generation
von: Novack, Zachary, et al.
Veröffentlicht: (2024)
von: Novack, Zachary, et al.
Veröffentlicht: (2024)
WeaveMuse: An Open Agentic System for Multimodal Music Understanding and Generation
von: Karystinaios, Emmanouil
Veröffentlicht: (2025)
von: Karystinaios, Emmanouil
Veröffentlicht: (2025)
Pushing the Limits of End-to-End Diarization
von: Broughton, Samuel J., et al.
Veröffentlicht: (2025)
von: Broughton, Samuel J., et al.
Veröffentlicht: (2025)
Procedural Music Generation Systems in Games
von: Luo, Shangxuan, et al.
Veröffentlicht: (2025)
von: Luo, Shangxuan, et al.
Veröffentlicht: (2025)
CCMusic: An Open and Diverse Database for Chinese Music Information Retrieval Research
von: Zhou, Monan, et al.
Veröffentlicht: (2025)
von: Zhou, Monan, et al.
Veröffentlicht: (2025)
Towards Practical Real-Time Low-Latency Music Source Separation
von: Wu, Junyu, et al.
Veröffentlicht: (2025)
von: Wu, Junyu, et al.
Veröffentlicht: (2025)
Multi-Source Diffusion Models for Simultaneous Music Generation and Separation
von: Mariani, Giorgio, et al.
Veröffentlicht: (2023)
von: Mariani, Giorgio, et al.
Veröffentlicht: (2023)
OpenBEATs: A Fully Open-Source General-Purpose Audio Encoder
von: Bharadwaj, Shikhar, et al.
Veröffentlicht: (2025)
von: Bharadwaj, Shikhar, et al.
Veröffentlicht: (2025)
MeanFlow-Accelerated Multimodal Video-to-Audio Synthesis via One-Step Generation
von: Yang, Xiaoran, et al.
Veröffentlicht: (2025)
von: Yang, Xiaoran, et al.
Veröffentlicht: (2025)
Intelligent Text-Conditioned Music Generation
von: Xie, Zhouyao, et al.
Veröffentlicht: (2024)
von: Xie, Zhouyao, et al.
Veröffentlicht: (2024)
Activation Patching for Interpretable Steering in Music Generation
von: Facchiano, Simone, et al.
Veröffentlicht: (2025)
von: Facchiano, Simone, et al.
Veröffentlicht: (2025)
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
MusicMamba: A Dual-Feature Modeling Approach for Generating Chinese Traditional Music with Modal Precision
von: Chen, Jiatao, et al.
Veröffentlicht: (2024)
von: Chen, Jiatao, et al.
Veröffentlicht: (2024)
YuE: Scaling Open Foundation Models for Long-Form Music Generation
von: Yuan, Ruibin, et al.
Veröffentlicht: (2025)
von: Yuan, Ruibin, et al.
Veröffentlicht: (2025)
Musical Attention Transformer: Music Generation Using a Music-Specific Attention Model
von: Taksuka, Shinnosuke, et al.
Veröffentlicht: (2026)
von: Taksuka, Shinnosuke, et al.
Veröffentlicht: (2026)
TMD-Bench: A Multi-Level Evaluation Paradigm for Music-Dance Co-Generation
von: Yang, Xiaoda, et al.
Veröffentlicht: (2026)
von: Yang, Xiaoda, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ACE-Step: A Step Towards Music Generation Foundation Model
von: Gong, Junmin, et al.
Veröffentlicht: (2025) -
ConSinger: Efficient High-Fidelity Singing Voice Generation with Minimal Steps
von: Song, Yulin, et al.
Veröffentlicht: (2024) -
HeartMuLa: A Family of Open Sourced Music Foundation Models
von: Yang, Dongchao, et al.
Veröffentlicht: (2026) -
Amphion: An Open-Source Audio, Music and Speech Generation Toolkit
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023) -
U3-xi: Pushing the Boundaries of Speaker Recognition by Incorporating Uncertainty
von: Li, Junjie, et al.
Veröffentlicht: (2026)