CHORDONOMICON: A Dataset of 666,000 Songs and their Chord Progressions

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Kantarelis, Spyridon, Thomas, Konstantinos, Lyberatos, Vassilis, Dervakos, Edmund, Stamou, Giorgos
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913607038533632
author Kantarelis, Spyridon
Thomas, Konstantinos
Lyberatos, Vassilis
Dervakos, Edmund
Stamou, Giorgos
author_facet Kantarelis, Spyridon
Thomas, Konstantinos
Lyberatos, Vassilis
Dervakos, Edmund
Stamou, Giorgos
contents Chord progressions encapsulate important information about music, pertaining to its structure and conveyed emotions. They serve as the backbone of musical composition, and in many cases, they are the sole information required for a musician to play along and follow the music. Despite their importance, chord progressions as a data domain remain underexplored. There is a lack of large-scale datasets suitable for deep learning applications, and limited research exploring chord progressions as an input modality. In this work, we present Chordonomicon, a dataset of over 666,000 songs and their chord progressions, annotated with structural parts, genre, and release date - created by scraping various sources of user-generated progressions and associated metadata. We demonstrate the practical utility of the Chordonomicon dataset for classification and generation tasks, and discuss its potential to provide valuable insights to the research community. Chord progressions are unique in their ability to be represented in multiple formats (e.g. text, graph) and the wealth of information chords convey in given contexts, such as their harmonic function . These characteristics make the Chordonomicon an ideal testbed for exploring advanced machine learning techniques, including transformers, graph machine learning, and hybrid systems that combine knowledge representation and machine learning.
format Preprint
id arxiv_https___arxiv_org_abs_2410_22046
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle CHORDONOMICON: A Dataset of 666,000 Songs and their Chord Progressions
Kantarelis, Spyridon
Thomas, Konstantinos
Lyberatos, Vassilis
Dervakos, Edmund
Stamou, Giorgos
Sound
Machine Learning
Multimedia
Audio and Speech Processing
Chord progressions encapsulate important information about music, pertaining to its structure and conveyed emotions. They serve as the backbone of musical composition, and in many cases, they are the sole information required for a musician to play along and follow the music. Despite their importance, chord progressions as a data domain remain underexplored. There is a lack of large-scale datasets suitable for deep learning applications, and limited research exploring chord progressions as an input modality. In this work, we present Chordonomicon, a dataset of over 666,000 songs and their chord progressions, annotated with structural parts, genre, and release date - created by scraping various sources of user-generated progressions and associated metadata. We demonstrate the practical utility of the Chordonomicon dataset for classification and generation tasks, and discuss its potential to provide valuable insights to the research community. Chord progressions are unique in their ability to be represented in multiple formats (e.g. text, graph) and the wealth of information chords convey in given contexts, such as their harmonic function . These characteristics make the Chordonomicon an ideal testbed for exploring advanced machine learning techniques, including transformers, graph machine learning, and hybrid systems that combine knowledge representation and machine learning.
title CHORDONOMICON: A Dataset of 666,000 Songs and their Chord Progressions
topic Sound
Machine Learning
Multimedia
Audio and Speech Processing
url https://arxiv.org/abs/2410.22046