STAGE: Stemmed Accompaniment Generation through Prefix-Based Conditioning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Strano, Giorgio, Ballanti, Chiara, Crisostomi, Donato, Mancusi, Michele, Cosmo, Luca, Rodolà, Emanuele |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-Source Diffusion Models for Simultaneous Music Generation and Separation
von: Mariani, Giorgio, et al.
Veröffentlicht: (2023)
von: Mariani, Giorgio, et al.
Veröffentlicht: (2023)
EuleroDec: A Complex-Valued RVQ-VAE for Efficient and Robust Audio Coding
von: Cerovaz, Luca, et al.
Veröffentlicht: (2026)
von: Cerovaz, Luca, et al.
Veröffentlicht: (2026)
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
Generalized Multi-Source Inference for Text Conditioned Music Diffusion Models
von: Postolache, Emilian, et al.
Veröffentlicht: (2024)
von: Postolache, Emilian, et al.
Veröffentlicht: (2024)
Naturalistic Music Decoding from EEG Data via Latent Diffusion Models
von: Postolache, Emilian, et al.
Veröffentlicht: (2024)
von: Postolache, Emilian, et al.
Veröffentlicht: (2024)
Drop the beat! Freestyler for Accompaniment Conditioned Rapping Voice Generation
von: Ning, Ziqian, et al.
Veröffentlicht: (2024)
von: Ning, Ziqian, et al.
Veröffentlicht: (2024)
ITO-Master: Inference-Time Optimization for Audio Effects Modeling of Music Mastering Processors
von: Koo, Junghyun, et al.
Veröffentlicht: (2025)
von: Koo, Junghyun, et al.
Veröffentlicht: (2025)
Zero-Shot Duet Singing Voices Separation with Diffusion Models
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
Accompaniment Prompt Adherence: A Measure for Evaluating Music Accompaniment Systems
von: Grachten, Maarten, et al.
Veröffentlicht: (2025)
von: Grachten, Maarten, et al.
Veröffentlicht: (2025)
The Florence Price Art Song Dataset and Piano Accompaniment Generator
von: He, Tao-Tao, et al.
Veröffentlicht: (2025)
von: He, Tao-Tao, et al.
Veröffentlicht: (2025)
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
Improving Musical Accompaniment Co-creation via Diffusion Transformers
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
von: Nistal, Javier, et al.
Veröffentlicht: (2024)
A Neural Score Follower for Computer Accompaniment of Polyphonic Musical Instruments
von: Pillay, Ashwin
Veröffentlicht: (2025)
von: Pillay, Ashwin
Veröffentlicht: (2025)
Bass Accompaniment Generation via Latent Diffusion
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
MusicGen-Stem: Multi-stem music generation and edition through autoregressive modeling
von: Rouard, Simon, et al.
Veröffentlicht: (2025)
von: Rouard, Simon, et al.
Veröffentlicht: (2025)
SteerMusic: Enhanced Musical Consistency for Zero-shot Text-guided and Personalized Music Editing
von: Niu, Xinlei, et al.
Veröffentlicht: (2025)
von: Niu, Xinlei, et al.
Veröffentlicht: (2025)
High-Resolution Speech Restoration with Latent Diffusion Model
von: Dhyani, Tushar, et al.
Veröffentlicht: (2024)
von: Dhyani, Tushar, et al.
Veröffentlicht: (2024)
Sing-On-Your-Beat: Simple Text-Controllable Accompaniment Generations
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2024)
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2024)
AudSemThinker: Enhancing Audio-Language Models through Reasoning over Semantics of Sound
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2025)
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2025)
Room Impulse Response Generation Conditioned on Acoustic Parameters
von: Arellano, Silvia, et al.
Veröffentlicht: (2025)
von: Arellano, Silvia, et al.
Veröffentlicht: (2025)
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
von: Chen, Wei, et al.
Veröffentlicht: (2025)
von: Chen, Wei, et al.
Veröffentlicht: (2025)
LipDiffuser: Lip-to-Speech Generation with Conditional Diffusion Models
von: Richter, Julius, et al.
Veröffentlicht: (2025)
von: Richter, Julius, et al.
Veröffentlicht: (2025)
Audio Conditioning for Music Generation via Discrete Bottleneck Features
von: Rouard, Simon, et al.
Veröffentlicht: (2024)
von: Rouard, Simon, et al.
Veröffentlicht: (2024)
Enhanced Automatic Drum Transcription via Drum Stem Source Separation
von: Riley, Xavier, et al.
Veröffentlicht: (2025)
von: Riley, Xavier, et al.
Veröffentlicht: (2025)
LL-SDR: Low-Latency Speech enhancement through Discrete Representations
von: Li, Jingyi, et al.
Veröffentlicht: (2026)
von: Li, Jingyi, et al.
Veröffentlicht: (2026)
Joint Audio and Symbolic Conditioning for Temporally Controlled Text-to-Music Generation
von: Tal, Or, et al.
Veröffentlicht: (2024)
von: Tal, Or, et al.
Veröffentlicht: (2024)
Diffusion based Text-to-Music Generation with Global and Local Text based Conditioning
von: Zhang, Jisi, et al.
Veröffentlicht: (2025)
von: Zhang, Jisi, et al.
Veröffentlicht: (2025)
ACMID: Automatic Curation of Musical Instrument Dataset for 7-Stem Music Source Separation
von: Yu, Ji, et al.
Veröffentlicht: (2025)
von: Yu, Ji, et al.
Veröffentlicht: (2025)
Conditional Latent Diffusion-Based Speech Enhancement Via Dual Context Learning
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
Audible Networks: Deconstructing and Manipulating Sounds with Deep Non-Negative Autoencoders
von: Burred, Juan José, et al.
Veröffentlicht: (2025)
von: Burred, Juan José, et al.
Veröffentlicht: (2025)
MaskBeat: Loopable Drum Beat Generation
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
Improving Anomalous Sound Detection through Pseudo-anomalous Set Selection and Pseudo-label Utilization under Unlabeled Conditions
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
FakeMusicCaps: a Dataset for Detection and Attribution of Synthetic Music Generated via Text-to-Music Models
von: Comanducci, Luca, et al.
Veröffentlicht: (2024)
von: Comanducci, Luca, et al.
Veröffentlicht: (2024)
SNC: A Stem-Native Codec for Efficient Lossless Audio Storage with Adaptive Playback Capabilities
von: Sufi, Shaad
Veröffentlicht: (2026)
von: Sufi, Shaad
Veröffentlicht: (2026)
Combining Deterministic Enhanced Conditions with Dual-Streaming Encoding for Diffusion-Based Speech Enhancement
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
Improving Real-Time Music Accompaniment Separation with MMDenseNet
von: Wang, Chun-Hsiang, et al.
Veröffentlicht: (2024)
von: Wang, Chun-Hsiang, et al.
Veröffentlicht: (2024)
Accompanied Singing Voice Synthesis with Fully Text-controlled Melody
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
MoodLoopGP: Generating Emotion-Conditioned Loop Tablature Music with Multi-Granular Features
von: Cui, Wenqian, et al.
Veröffentlicht: (2024)
von: Cui, Wenqian, et al.
Veröffentlicht: (2024)
Small Tunes Transformer: Exploring Macro & Micro-Level Hierarchies for Skeleton-Conditioned Melody Generation
von: Lv, Yishan, et al.
Veröffentlicht: (2024)
von: Lv, Yishan, et al.
Veröffentlicht: (2024)
MambaFoley: Foley Sound Generation using Selective State-Space Models
von: Colombo, Marco Furio, et al.
Veröffentlicht: (2024)
von: Colombo, Marco Furio, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Multi-Source Diffusion Models for Simultaneous Music Generation and Separation
von: Mariani, Giorgio, et al.
Veröffentlicht: (2023) -
EuleroDec: A Complex-Valued RVQ-VAE for Efficient and Robust Audio Coding
von: Cerovaz, Luca, et al.
Veröffentlicht: (2026) -
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024) -
Generalized Multi-Source Inference for Text Conditioned Music Diffusion Models
von: Postolache, Emilian, et al.
Veröffentlicht: (2024) -
Naturalistic Music Decoding from EEG Data via Latent Diffusion Models
von: Postolache, Emilian, et al.
Veröffentlicht: (2024)