Convolution-based classification of audio and symbolic representations of music

Gissel Velarde, Carlos Cancino Chacón, David Meredith, Tillman Weyde, Maarten Grachten

Publikation: Bidrag til tidsskriftTidsskriftartikelForskningpeer review

10 Citationer (Scopus)
121 Downloads (Pure)

Abstract

We present a novel convolution-based method for classification of audio and symbolic representations of music, which we apply to classification of music by style. Pieces of music are first sampled to pitch–time representations (piano-rolls or spectrograms) and then convolved with a Gaussian filter, before being classified by a support vector machine or by k-nearest neighbours in an ensemble of classifiers. On the well-studied task of discriminating between string quartet movements by Haydn and Mozart, we obtain accuracies that equal the state of the art on two data-sets. However, in multi-class composer identification, methods specialised for classifying symbolic representations of music are more effective. We also performed experiments on symbolic representations, synthetic audio and two different recordings of The Well-Tempered Clavier by J. S. Bach to study the method’s capacity to distinguish preludes from fugues. Our experimental results show that our approach performs similarly on symbolic representations, synthetic audio and audio recordings, setting our method apart from most previous studies that have been designed for use with either audio or symbolic data, but not both.
OriginalsprogEngelsk
TidsskriftJournal of New Music Research
Vol/bind47
Udgave nummer3
Sider (fra-til)191-205
Antal sider15
ISSN0929-8215
DOI
StatusUdgivet - 27 maj 2018

Fingeraftryk

Dyk ned i forskningsemnerne om 'Convolution-based classification of audio and symbolic representations of music'. Sammen danner de et unikt fingeraftryk.

Citationsformater