An RNN-based Music Language Model for Improving Automatic Music Transcription

Sigtia, S.; Benetos, E.; Cherla, S.; Weyde, T.; Garcez, A.; Dixon, S.

research

An RNN-based Music Language Model for Improving Automatic Music Transcription

Authors: S. Sigtia
E. Benetos
S. Cherla
T. Weyde
A. Garcez
S. Dixon
Publication date: 1 January 2014
Publisher: International Society for Music Information Retrieval

Abstract

In this paper, we investigate the use of Music Language Models (MLMs) for improving Automatic Music Transcription performance. The MLMs are trained on sequences of symbolic polyphonic music from the Nottingham dataset. We train Recurrent Neural Network (RNN)-based models, as they are capable of capturing complex temporal structure present in symbolic music data. Similar to the function of language models in automatic speech recognition, we use the MLMs to generate a prior probability for the occurrence of a sequence. The acoustic AMT model is based on probabilistic latent component analysis, and prior information from the MLM is incorporated into the transcription framework using Dirichlet priors. We test our hybrid models on a dataset of multiple-instrument polyphonic music and report a significant 3% improvement in terms of F-measure, when compared to using an acoustic-only model

Similar works

Full text

Open in the Core reader

Download PDF

Available Versions

Sustaining member

City Research Online

oai:openaccess.city.ac.uk:4529

Last time updated on 02/08/2016