Exploring Procedural Data Generation for Automatic Acoustic Guitar Fingerpicking Transcription

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Murgul, Sebastian, Heizmann, Michael
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915439810969600
author Murgul, Sebastian
Heizmann, Michael
author_facet Murgul, Sebastian
Heizmann, Michael
contents Automatic transcription of acoustic guitar fingerpicking performances remains a challenging task due to the scarcity of labeled training data and legal constraints connected with musical recordings. This work investigates a procedural data generation pipeline as an alternative to real audio recordings for training transcription models. Our approach synthesizes training data through four stages: knowledge-based fingerpicking tablature composition, MIDI performance rendering, physical modeling using an extended Karplus-Strong algorithm, and audio augmentation including reverb and distortion. We train and evaluate a CRNN-based note-tracking model on both real and synthetic datasets, demonstrating that procedural data can be used to achieve reasonable note-tracking results. Finetuning with a small amount of real data further enhances transcription accuracy, improving over models trained exclusively on real recordings. These results highlight the potential of procedurally generated audio for data-scarce music information retrieval tasks.
format Preprint
id arxiv_https___arxiv_org_abs_2508_07987
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Exploring Procedural Data Generation for Automatic Acoustic Guitar Fingerpicking Transcription
Murgul, Sebastian
Heizmann, Michael
Sound
Computation and Language
Audio and Speech Processing
Automatic transcription of acoustic guitar fingerpicking performances remains a challenging task due to the scarcity of labeled training data and legal constraints connected with musical recordings. This work investigates a procedural data generation pipeline as an alternative to real audio recordings for training transcription models. Our approach synthesizes training data through four stages: knowledge-based fingerpicking tablature composition, MIDI performance rendering, physical modeling using an extended Karplus-Strong algorithm, and audio augmentation including reverb and distortion. We train and evaluate a CRNN-based note-tracking model on both real and synthetic datasets, demonstrating that procedural data can be used to achieve reasonable note-tracking results. Finetuning with a small amount of real data further enhances transcription accuracy, improving over models trained exclusively on real recordings. These results highlight the potential of procedurally generated audio for data-scarce music information retrieval tasks.
title Exploring Procedural Data Generation for Automatic Acoustic Guitar Fingerpicking Transcription
topic Sound
Computation and Language
Audio and Speech Processing
url https://arxiv.org/abs/2508.07987