1 research outputs found

    Discrete Wavelet Packet Transform and Ensembles of Lazy and Eager Learners for Music Genre Classification

    No full text
    This paper presents a process for determining the music genre of an item using a new set of descriptors. A discrete wavelet packet transform is applied to obtain the signal representation at two different resolutions, a frequency resolution and a time resolution tuned to encode music notes and their onset and offset. These features are tested on a number of datasets as descriptors for music genre classification. Lazy learning classifiers (k-nearest neighbor) and eager learners (neural networks and support vector machines) are applied in order to assess the classification power of the proposed features. Different feature selection techniques and ensemble methods are explored to maximize the accuracy of the classifiers and stabilize their behavior. Our evaluation shows that these frequency descriptors perform better than a standard approach based on Mel-Frequency Cepstral Coefficients and on the Short Time Fourier Transform in music genre classification. Moreover, our work confirms that a parameterization of the music rythm based on the beat-histogram provides some meaningfull information in the context of music classification by genre. Finally, our evaluation suggests that multiclass support vector machines with a linear kernel and round-robin binarization are the simplest and more effective process for music genre classification.
    corecore