MaskBeat: Loopable Drum Beat Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lanzendörfer, Luca A., Grötschla, Florian, Galal, Karim, Wattenhofer, Roger |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Leveraging Contrastively Pretrained Neural Audio Embeddings for Recommender Tasks
von: Grötschla, Florian, et al.
Veröffentlicht: (2024)
von: Grötschla, Florian, et al.
Veröffentlicht: (2024)
Parametric Neural Amp Modeling with Active Learning
von: Grötschla, Florian, et al.
Veröffentlicht: (2025)
von: Grötschla, Florian, et al.
Veröffentlicht: (2025)
Multi-bit Audio Watermarking
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
Audio Atlas: Visualizing and Exploring Audio Datasets
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2024)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2024)
SNAC: Multi-Scale Neural Audio Codec
von: Siuzdak, Hubert, et al.
Veröffentlicht: (2024)
von: Siuzdak, Hubert, et al.
Veröffentlicht: (2024)
High-Fidelity Speech Enhancement via Discrete Audio Tokens
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
Beat-It: Beat-Synchronized Multi-Condition 3D Dance Generation
von: Huang, Zikai, et al.
Veröffentlicht: (2024)
von: Huang, Zikai, et al.
Veröffentlicht: (2024)
The Rhythm In Anything: Audio-Prompted Drums Generation with Masked Language Modeling
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2025)
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2025)
BEAST: Online Joint Beat and Downbeat Tracking Based on Streaming Transformer
von: Chang, Chih-Cheng, et al.
Veröffentlicht: (2023)
von: Chang, Chih-Cheng, et al.
Veröffentlicht: (2023)
HingeNet: A Harmonic-Aware Fine-Tuning Approach for Beat Tracking
von: Ru, Ganghui, et al.
Veröffentlicht: (2025)
von: Ru, Ganghui, et al.
Veröffentlicht: (2025)
The SMC Blind Spot: A Failure Mode Analysis of State-of-the-Art Beat Tracking
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2026)
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2026)
Sing-On-Your-Beat: Simple Text-Controllable Accompaniment Generations
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2024)
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2024)
Controlling Contrastive Self-Supervised Learning with Knowledge-Driven Multiple Hypothesis: Application to Beat Tracking
von: Gagnere, Antonin, et al.
Veröffentlicht: (2025)
von: Gagnere, Antonin, et al.
Veröffentlicht: (2025)
Enhanced Automatic Drum Transcription via Drum Stem Source Separation
von: Riley, Xavier, et al.
Veröffentlicht: (2025)
von: Riley, Xavier, et al.
Veröffentlicht: (2025)
Transformer-Based Rhythm Quantization of Performance MIDI Using Beat Annotations
von: Wachter, Maximilian, et al.
Veröffentlicht: (2026)
von: Wachter, Maximilian, et al.
Veröffentlicht: (2026)
Relationships between Keywords and Strong Beats in Lyrical Music
von: Liao, Callie C., et al.
Veröffentlicht: (2024)
von: Liao, Callie C., et al.
Veröffentlicht: (2024)
Beat this! Accurate beat tracking without DBN postprocessing
von: Foscarin, Francesco, et al.
Veröffentlicht: (2024)
von: Foscarin, Francesco, et al.
Veröffentlicht: (2024)
VCNAC: A Variable-Channel Neural Audio Codec for Mono, Stereo, and Surround Sound
von: Grötschla, Florian, et al.
Veröffentlicht: (2026)
von: Grötschla, Florian, et al.
Veröffentlicht: (2026)
Schrodinger Bridges Beat Diffusion Models on Text-to-Speech Synthesis
von: Chen, Zehua, et al.
Veröffentlicht: (2023)
von: Chen, Zehua, et al.
Veröffentlicht: (2023)
Beat-Based Rhythm Quantization of MIDI Performances
von: Wachter, Maximilian, et al.
Veröffentlicht: (2025)
von: Wachter, Maximilian, et al.
Veröffentlicht: (2025)
Efficient Adapter Tuning for Joint Singing Voice Beat and Downbeat Tracking with Self-supervised Learning Features
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
DOSE : Drum One-Shot Extraction from Music Mixture
von: Hwang, Suntae, et al.
Veröffentlicht: (2025)
von: Hwang, Suntae, et al.
Veröffentlicht: (2025)
Drum-to-Vocal Percussion Sound Conversion and Its Evaluation Methodology
von: Nobukawa, Rinka, et al.
Veröffentlicht: (2025)
von: Nobukawa, Rinka, et al.
Veröffentlicht: (2025)
Dance Any Beat: Blending Beats with Visuals in Dance Video Generation
von: Wang, Xuanchen, et al.
Veröffentlicht: (2024)
von: Wang, Xuanchen, et al.
Veröffentlicht: (2024)
DanceAnyWay: Synthesizing Beat-Guided 3D Dances with Randomized Temporal Contrastive Learning
von: Bhattacharya, Aneesh, et al.
Veröffentlicht: (2023)
von: Bhattacharya, Aneesh, et al.
Veröffentlicht: (2023)
Beat and Downbeat Tracking in Performance MIDI Using an End-to-End Transformer Architecture
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
SmoothSync: Dual-Stream Diffusion Transformers for Jitter-Robust Beat-Synchronized Gesture Generation from Quantized Audio
von: Jiang, Yujiao, et al.
Veröffentlicht: (2026)
von: Jiang, Yujiao, et al.
Veröffentlicht: (2026)
Toward Deep Drum Source Separation
von: Mezza, Alessandro Ilic, et al.
Veröffentlicht: (2023)
von: Mezza, Alessandro Ilic, et al.
Veröffentlicht: (2023)
Noise-to-Notes: Diffusion-based Generation and Refinement for Automatic Drum Transcription
von: Yeung, Michael, et al.
Veröffentlicht: (2025)
von: Yeung, Michael, et al.
Veröffentlicht: (2025)
SpecMaskGIT: Masked Generative Modeling of Audio Spectrograms for Efficient Audio Synthesis and Beyond
von: Comunità, Marco, et al.
Veröffentlicht: (2024)
von: Comunità, Marco, et al.
Veröffentlicht: (2024)
Long-Term, Store-Front Robotics: Interactive Music for Robotic Arm, Caxixi and Frame Drums
von: Savery, Richard, et al.
Veröffentlicht: (2024)
von: Savery, Richard, et al.
Veröffentlicht: (2024)
MAGE: A Coarse-to-Fine Speech Enhancer with Masked Generative Model
von: Pham, The Hieu, et al.
Veröffentlicht: (2025)
von: Pham, The Hieu, et al.
Veröffentlicht: (2025)
DARC: Drum accompaniment generation with fine-grained rhythm control
von: Brosnan, Trey
Veröffentlicht: (2026)
von: Brosnan, Trey
Veröffentlicht: (2026)
IMPACT: Iterative Mask-based Parallel Decoding for Text-to-Audio Generation with Diffusion Modeling
von: Huang, Kuan-Po, et al.
Veröffentlicht: (2025)
von: Huang, Kuan-Po, et al.
Veröffentlicht: (2025)
Genuine-Focused Learning using Mask AutoEncoder for Generalized Fake Audio Detection
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
Vocoder-Free Non-Parallel Conversion of Whispered Speech With Masked Cycle-Consistent Generative Adversarial Networks
von: Wagner, Dominik, et al.
Veröffentlicht: (2023)
von: Wagner, Dominik, et al.
Veröffentlicht: (2023)
Masked Audio Modeling with CLAP and Multi-Objective Learning
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
Ambisonics Binaural Rendering via Masked Magnitude Least Squares
von: Berebi, Or, et al.
Veröffentlicht: (2025)
von: Berebi, Or, et al.
Veröffentlicht: (2025)
Disentangling Dual-Encoder Masked Autoencoder for Respiratory Sound Classification
von: Wei, Peidong, et al.
Veröffentlicht: (2025)
von: Wei, Peidong, et al.
Veröffentlicht: (2025)
ChunkFormer: Masked Chunking Conformer For Long-Form Speech Transcription
von: Le, Khanh, et al.
Veröffentlicht: (2025)
von: Le, Khanh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Leveraging Contrastively Pretrained Neural Audio Embeddings for Recommender Tasks
von: Grötschla, Florian, et al.
Veröffentlicht: (2024) -
Parametric Neural Amp Modeling with Active Learning
von: Grötschla, Florian, et al.
Veröffentlicht: (2025) -
Multi-bit Audio Watermarking
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025) -
Audio Atlas: Visualizing and Exploring Audio Datasets
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2024) -
SNAC: Multi-Scale Neural Audio Codec
von: Siuzdak, Hubert, et al.
Veröffentlicht: (2024)