LaDA-Band: Language Diffusion Models for Vocal-to-Accompaniment Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Qi, Shen, Zhexu, Chen, Meng, Yu, Guoxin, Pang, Chaoxu, Zhao, Weifeng, Zhou, Wenjiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Streaming Generation for Music Accompaniment
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling
von: Geng, Chen, et al.
Veröffentlicht: (2026)
von: Geng, Chen, et al.
Veröffentlicht: (2026)
TQCodec: Towards neural audio codec for high-fidelity music streaming
von: He, Lixing, et al.
Veröffentlicht: (2026)
von: He, Lixing, et al.
Veröffentlicht: (2026)
HAFM: Hierarchical Autoregressive Foundation Model for Music Accompaniment Generation
von: Zhu, Jian, et al.
Veröffentlicht: (2026)
von: Zhu, Jian, et al.
Veröffentlicht: (2026)
Bass Accompaniment Generation via Latent Diffusion
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
Structure-Aware Piano Accompaniment via Style Planning and Dataset-Aligned Pattern Retrieval
von: Zang, Wanyu, et al.
Veröffentlicht: (2026)
von: Zang, Wanyu, et al.
Veröffentlicht: (2026)
Generating Separated Singing Vocals Using a Diffusion Model Conditioned on Music Mixtures
von: Plaja-Roglans, Genís, et al.
Veröffentlicht: (2025)
von: Plaja-Roglans, Genís, et al.
Veröffentlicht: (2025)
VocalNet-MDM: Accelerating Streaming Speech LLM via Self-Distilled Masked Diffusion Modeling
von: Cheng, Ziyang, et al.
Veröffentlicht: (2026)
von: Cheng, Ziyang, et al.
Veröffentlicht: (2026)
Towards Real-Time Human-AI Musical Co-Performance: Accompaniment Generation with Latent Diffusion Models and MAX/MSP
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2026)
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2026)
D3PIA: A Discrete Denoising Diffusion Model for Piano Accompaniment Generation From Lead sheet
von: Choi, Eunjin, et al.
Veröffentlicht: (2026)
von: Choi, Eunjin, et al.
Veröffentlicht: (2026)
Smule Renaissance Small: Efficient General-Purpose Vocal Restoration
von: Zang, Yongyi, et al.
Veröffentlicht: (2025)
von: Zang, Yongyi, et al.
Veröffentlicht: (2025)
Mel-RoFormer for Vocal Separation and Vocal Melody Transcription
von: Wang, Ju-Chiang, et al.
Veröffentlicht: (2024)
von: Wang, Ju-Chiang, et al.
Veröffentlicht: (2024)
VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
Improving Musical Accompaniment Co-creation via Diffusion Transformers
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
JoyTTS: LLM-based Spoken Chatbot With Voice Cloning
von: Zhou, Fangru, et al.
Veröffentlicht: (2025)
von: Zhou, Fangru, et al.
Veröffentlicht: (2025)
Drop the beat! Freestyler for Accompaniment Conditioned Rapping Voice Generation
von: Ning, Ziqian, et al.
Veröffentlicht: (2024)
von: Ning, Ziqian, et al.
Veröffentlicht: (2024)
Accompaniment Prompt Adherence: A Measure for Evaluating Music Accompaniment Systems
von: Grachten, Maarten, et al.
Veröffentlicht: (2025)
von: Grachten, Maarten, et al.
Veröffentlicht: (2025)
VocalAgent: Large Language Models for Vocal Health Diagnostics with Safety-Aware Evaluation
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
UniVocal: Unified Speech-Singing Code-Switching Synthesis
von: Shi, Yufei, et al.
Veröffentlicht: (2026)
von: Shi, Yufei, et al.
Veröffentlicht: (2026)
Adaptive Accompaniment with ReaLchords
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
Structured Multi-Track Accompaniment Arrangement via Style Prior Modelling
von: Zhao, Jingwei, et al.
Veröffentlicht: (2023)
von: Zhao, Jingwei, et al.
Veröffentlicht: (2023)
SAMUeL: Efficient Vocal-Conditioned Music Generation via Soft Alignment Attention and Latent Diffusion
von: Cheung, Hei Shing, et al.
Veröffentlicht: (2025)
von: Cheung, Hei Shing, et al.
Veröffentlicht: (2025)
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
AudioRAG+: Feedback-driven Retrieval-augmented Audio Generation with Large Audio Language Models
von: Zhao, Junqi, et al.
Veröffentlicht: (2025)
von: Zhao, Junqi, et al.
Veröffentlicht: (2025)
AnyAccomp: Generalizable Accompaniment Generation via Quantized Melodic Bottleneck
von: Zhang, Junan, et al.
Veröffentlicht: (2025)
von: Zhang, Junan, et al.
Veröffentlicht: (2025)
Accompanied Singing Voice Synthesis with Fully Text-controlled Melody
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
Extract and Diffuse: Latent Integration for Improved Diffusion-based Speech and Vocal Enhancement
von: Yang, Yudong, et al.
Veröffentlicht: (2024)
von: Yang, Yudong, et al.
Veröffentlicht: (2024)
The Florence Price Art Song Dataset and Piano Accompaniment Generator
von: He, Tao-Tao, et al.
Veröffentlicht: (2025)
von: He, Tao-Tao, et al.
Veröffentlicht: (2025)
STAGE: Stemmed Accompaniment Generation through Prefix-Based Conditioning
von: Strano, Giorgio, et al.
Veröffentlicht: (2025)
von: Strano, Giorgio, et al.
Veröffentlicht: (2025)
Diff-A-Riff: Musical Accompaniment Co-creation via Latent Diffusion Models
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
NVBench: A Benchmark for Speech Synthesis with Non-Verbal Vocalizations
von: Xue, Liumeng, et al.
Veröffentlicht: (2026)
von: Xue, Liumeng, et al.
Veröffentlicht: (2026)
DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
von: Zhou, Jiaming, et al.
Veröffentlicht: (2026)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2026)
Tutti: Expressive Multi-Singer Synthesis via Structure-Level Timbre Control and Vocal Texture Modeling
von: Chen, Jiatao, et al.
Veröffentlicht: (2026)
von: Chen, Jiatao, et al.
Veröffentlicht: (2026)
CVSM: Contrastive Vocal Similarity Modeling
von: Garoufis, Christos, et al.
Veröffentlicht: (2025)
von: Garoufis, Christos, et al.
Veröffentlicht: (2025)
Fed-PISA: Federated Voice Cloning via Personalized Identity-Style Adaptation
von: Wang, Qi, et al.
Veröffentlicht: (2025)
von: Wang, Qi, et al.
Veröffentlicht: (2025)
WhAM: Towards A Translative Model of Sperm Whale Vocalization
von: Paradise, Orr, et al.
Veröffentlicht: (2025)
von: Paradise, Orr, et al.
Veröffentlicht: (2025)
AIBA: Attention-based Instrument Band Alignment for Text-to-Audio Diffusion
von: Koh, Junyoung, et al.
Veröffentlicht: (2025)
von: Koh, Junyoung, et al.
Veröffentlicht: (2025)
RMVPE: A Robust Model for Vocal Pitch Estimation in Polyphonic Music
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
Diffusion-based Symbolic Music Generation with Structured State Space Models
von: Yuan, Shenghua, et al.
Veröffentlicht: (2025)
von: Yuan, Shenghua, et al.
Veröffentlicht: (2025)
VocalNet-M2: Advancing Low-Latency Spoken Language Modeling via Integrated Multi-Codebook Tokenization and Multi-Token Prediction
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Streaming Generation for Music Accompaniment
von: Wu, Yusong, et al.
Veröffentlicht: (2025) -
Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling
von: Geng, Chen, et al.
Veröffentlicht: (2026) -
TQCodec: Towards neural audio codec for high-fidelity music streaming
von: He, Lixing, et al.
Veröffentlicht: (2026) -
HAFM: Hierarchical Autoregressive Foundation Model for Music Accompaniment Generation
von: Zhu, Jian, et al.
Veröffentlicht: (2026) -
Bass Accompaniment Generation via Latent Diffusion
von: Pasini, Marco, et al.
Veröffentlicht: (2024)