Keep the beat going: Automatic drum transcription with momentum
Fuente:
arXiv
Guardado en:
| Autores principales: | Foster, Alisha L., Webber, Robert J. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Personal Sound Zones and Shielded Localized Communication through Active Acoustic Control
por: Egarguin, Neil Jerome A., et al.
Publicado: (2024)
por: Egarguin, Neil Jerome A., et al.
Publicado: (2024)
Drop the beat! Freestyler for Accompaniment Conditioned Rapping Voice Generation
por: Ning, Ziqian, et al.
Publicado: (2024)
por: Ning, Ziqian, et al.
Publicado: (2024)
Improved symbolic drum style classification with grammar-based hierarchical representations
por: Géré, Léo, et al.
Publicado: (2024)
por: Géré, Léo, et al.
Publicado: (2024)
A Contrastive Self-Supervised Learning scheme for beat tracking amenable to few-shot learning
por: Gagnere, Antonin, et al.
Publicado: (2024)
por: Gagnere, Antonin, et al.
Publicado: (2024)
Voice Conversion-based Privacy through Adversarial Information Hiding
por: Webber, Jacob J, et al.
Publicado: (2024)
por: Webber, Jacob J, et al.
Publicado: (2024)
Analyzing and reducing the synthetic-to-real transfer gap in Music Information Retrieval: the task of automatic drum transcription
por: Zehren, Mickaël, et al.
Publicado: (2024)
por: Zehren, Mickaël, et al.
Publicado: (2024)
SongTrans: An unified song transcription and alignment method for lyrics and notes
por: Wu, Siwei, et al.
Publicado: (2024)
por: Wu, Siwei, et al.
Publicado: (2024)
Reconstructing the Charlie Parker Omnibook using an audio-to-score automatic transcription pipeline
por: Riley, Xavier, et al.
Publicado: (2024)
por: Riley, Xavier, et al.
Publicado: (2024)
Inversion of Arctic dual-channel sound speed profile based on random airgun signal
por: Weng, Jinbao, et al.
Publicado: (2025)
por: Weng, Jinbao, et al.
Publicado: (2025)
Acoustic source depth estimation method based on a single hydrophone in Arctic underwater
por: Weng, Jinbao, et al.
Publicado: (2025)
por: Weng, Jinbao, et al.
Publicado: (2025)
Beat this! Accurate beat tracking without DBN postprocessing
por: Foscarin, Francesco, et al.
Publicado: (2024)
por: Foscarin, Francesco, et al.
Publicado: (2024)
Comparator Loss: An Ordinal Contrastive Loss to Derive a Severity Score for Speech-based Health Monitoring
por: Webber, Jacob J, et al.
Publicado: (2025)
por: Webber, Jacob J, et al.
Publicado: (2025)
In-context learning capabilities of Large Language Models to detect suicide risk among adolescents from speech transcripts
por: Roquefort, Filomene, et al.
Publicado: (2025)
por: Roquefort, Filomene, et al.
Publicado: (2025)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
por: Tian, Jingguang, et al.
Publicado: (2024)
por: Tian, Jingguang, et al.
Publicado: (2024)
CAFA: a Controllable Automatic Foley Artist
por: Benita, Roi, et al.
Publicado: (2025)
por: Benita, Roi, et al.
Publicado: (2025)
Automatic Detection and Annotation of Sperm Whale Codas
por: Gubnitsky, Guy, et al.
Publicado: (2024)
por: Gubnitsky, Guy, et al.
Publicado: (2024)
Phonetic Richness for Improved Automatic Speaker Verification
por: Klein, Nicholas, et al.
Publicado: (2024)
por: Klein, Nicholas, et al.
Publicado: (2024)
BWSNet: Automatic Perceptual Assessment of Audio Signals
por: Veillon, Clément Le Moine, et al.
Publicado: (2023)
por: Veillon, Clément Le Moine, et al.
Publicado: (2023)
Automatic Melody Reduction via Shortest Path Finding
por: Wang, Ziyu, et al.
Publicado: (2025)
por: Wang, Ziyu, et al.
Publicado: (2025)
A Dataset for Automatic Assessment of TTS Quality in Spanish
por: Welford, Alejandro Sosa, et al.
Publicado: (2025)
por: Welford, Alejandro Sosa, et al.
Publicado: (2025)
The ICASSP 2026 Automatic Song Aesthetics Evaluation Challenge
por: Ma, Guobin, et al.
Publicado: (2026)
por: Ma, Guobin, et al.
Publicado: (2026)
Exploiting Music Source Separation for Automatic Lyrics Transcription with Whisper
por: Syed, Jaza, et al.
Publicado: (2025)
por: Syed, Jaza, et al.
Publicado: (2025)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
por: Bondaruk, Łukasz, et al.
Publicado: (2024)
por: Bondaruk, Łukasz, et al.
Publicado: (2024)
Applying Automatic Differentiation to Optimize Differential Microphone Array Designs
por: Galougah, Siminfar Samakoush, et al.
Publicado: (2024)
por: Galougah, Siminfar Samakoush, et al.
Publicado: (2024)
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
por: Farhadipour, Aref, et al.
Publicado: (2024)
por: Farhadipour, Aref, et al.
Publicado: (2024)
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2023)
por: Eeckt, Steven Vander, et al.
Publicado: (2023)
On the Relevance of Clinical Assessment Tasks for the Automatic Detection of Parkinson's Disease Medication State from Speech
por: Gimeno-Gómez, David, et al.
Publicado: (2025)
por: Gimeno-Gómez, David, et al.
Publicado: (2025)
Enhancing Automatic Chord Recognition through LLM Chain-of-Thought Reasoning
por: Chang, Chih-Cheng, et al.
Publicado: (2025)
por: Chang, Chih-Cheng, et al.
Publicado: (2025)
Automatic Music Mixing using a Generative Model of Effect Embeddings
por: Moliner, Eloi, et al.
Publicado: (2025)
por: Moliner, Eloi, et al.
Publicado: (2025)
Enhanced Automatic Drum Transcription via Drum Stem Source Separation
por: Riley, Xavier, et al.
Publicado: (2025)
por: Riley, Xavier, et al.
Publicado: (2025)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
por: Poncelet, Jakob, et al.
Publicado: (2025)
por: Poncelet, Jakob, et al.
Publicado: (2025)
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
por: He, Zhanhong, et al.
Publicado: (2025)
por: He, Zhanhong, et al.
Publicado: (2025)
Latent-Level Enhancement with Flow Matching for Robust Automatic Speech Recognition
por: Yang, Da-Hee, et al.
Publicado: (2026)
por: Yang, Da-Hee, et al.
Publicado: (2026)
Towards Automatic Assessment of Self-Supervised Speech Models using Rank
por: Aldeneh, Zakaria, et al.
Publicado: (2024)
por: Aldeneh, Zakaria, et al.
Publicado: (2024)
CLAP-Based Automatic Word Naming Recognition in Post-Stroke Aphasia
por: Kaloga, Yacouba, et al.
Publicado: (2026)
por: Kaloga, Yacouba, et al.
Publicado: (2026)
Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
por: Amiri, Mahdi, et al.
Publicado: (2024)
por: Amiri, Mahdi, et al.
Publicado: (2024)
The DKU System for Multi-Speaker Automatic Speech Recognition in MLC-SLM Challenge
por: Lin, Yuke, et al.
Publicado: (2025)
por: Lin, Yuke, et al.
Publicado: (2025)
A Survey on 30+ Years of Automatic Singing Assessment and Singing Information Processing
por: Santos, Arthur N. dos, et al.
Publicado: (2026)
por: Santos, Arthur N. dos, et al.
Publicado: (2026)
A Comparison of Differential Performance Metrics for the Evaluation of Automatic Speaker Verification Fairness
por: Chouchane, Oubaida, et al.
Publicado: (2024)
por: Chouchane, Oubaida, et al.
Publicado: (2024)
Fusion of Discrete Representations and Self-Augmented Representations for Multilingual Automatic Speech Recognition
por: Wang, Shih-heng, et al.
Publicado: (2024)
por: Wang, Shih-heng, et al.
Publicado: (2024)
Ejemplares similares
-
Personal Sound Zones and Shielded Localized Communication through Active Acoustic Control
por: Egarguin, Neil Jerome A., et al.
Publicado: (2024) -
Drop the beat! Freestyler for Accompaniment Conditioned Rapping Voice Generation
por: Ning, Ziqian, et al.
Publicado: (2024) -
Improved symbolic drum style classification with grammar-based hierarchical representations
por: Géré, Léo, et al.
Publicado: (2024) -
A Contrastive Self-Supervised Learning scheme for beat tracking amenable to few-shot learning
por: Gagnere, Antonin, et al.
Publicado: (2024) -
Voice Conversion-based Privacy through Adversarial Information Hiding
por: Webber, Jacob J, et al.
Publicado: (2024)