Porovnání metod pro normalizaci skóre aplikovaných na úlohu multi-label klasifikace

Müller, Luděk; Skorkovská, Lucie; Zajíc, Zbyněk

Porovnání metod pro normalizaci skóre aplikovaných na úlohu multi-label klasifikace

Authors: Luděk Müller
Lucie Skorkovská
Zbyněk Zajíc
Publication date: 1 January 2014
Publisher: IEEE Press

Abstract

Our paper deals with the multi-label text classification of the newspaper articles, where the classifier must decide if a document does or does not belong to each topic from the predefined topic set. A generative classifier is used to tackle this task and the problem with finding a threshold for the positive classification is mainly addressed. This threshold can vary for each document depending on the content of the document (words used, length of the document, etc.). An extensive comparison of the score normalization methods, primary proposed in the speaker identification/verification task, for robustly finding the threshold defining the boundary between the "correct'' and the "incorrect'' topics of a document is presented. Score normalization methods (based on World Model and Unconstrained Cohort Normalization) applied to the topic identification task has shown an improvement of results in our former experiments, therefore in this paper an in-depth experiments with more score normalization techniques applied to the multi-label classification were performed. Thorough analysis of the effects of the various parameters setting is presented

Similar works

Full text

Available Versions

DSpace at University of West Bohemia

oai:dspace5.zcu.cz:11025/17048

Last time updated on 09/04/2020

University of West Bohemia Digital Library

oai:dspace5.zcu.cz:11025/17048

Last time updated on 03/12/2017