2,878 research outputs found
Credibility Adjusted Term Frequency: A Supervised Term Weighting Scheme for Sentiment Analysis and Text Classification
We provide a simple but novel supervised weighting scheme for adjusting term
frequency in tf-idf for sentiment analysis and text classification. We compare
our method to baseline weighting schemes and find that it outperforms them on
multiple benchmarks. The method is robust and works well on both snippets and
longer documents
A Supervised Term-Weighting Method and its Application to Variable Extraction from Digital Media
Successful modeling and prediction depend on effective methods for the extraction of domain-relevant variables. This paper proposes a methodology for identifying domain-specific terms. The proposed methodology relies on a collection of documents labeled as relevant or irrelevant to the domain under analysis. Based on the labeled document collection, we propose a supervised technique that weights terms based on their descriptive and discriminating power. Finally, the descriptive and discriminating values are combined into a general measure that, through the use of an adjustable parameter, allows to independently favor different aspects of retrieval such as maximizing precision or recall, or achieving a balance between both of them. The proposed technique is applied to the economic domain and is empirically evaluated through a human-subject experiment involving experts and non-experts in Economy. It is also evaluated as a term-weighting technique for query-term selection showing promising results. We finally illustrate the potential of the proposal as a first step for identifying different types of associations between words.Fil: Maisonnave, Mariano. Consejo Nacional de Investigaciones CientÃficas y Técnicas. Centro CientÃfico Tecnológico Conicet - BahÃa Blanca. Instituto de Ciencias e IngenierÃa de la Computación. Universidad Nacional del Sur. Departamento de Ciencias e IngenierÃa de la Computación. Instituto de Ciencias e IngenierÃa de la Computación; Argentina. Universidad Nacional del Sur. Departamento de Ciencias e IngenierÃa de la Computación; ArgentinaFil: Delbianco, Fernando Andrés. Consejo Nacional de Investigaciones CientÃficas y Técnicas. Centro CientÃfico Tecnológico Conicet - BahÃa Blanca. Instituto de Matemática BahÃa Blanca. Universidad Nacional del Sur. Departamento de Matemática. Instituto de Matemática BahÃa Blanca; Argentina. Universidad Nacional del Sur. Departamento de EconomÃa; ArgentinaFil: Tohmé, Fernando Abel. Consejo Nacional de Investigaciones CientÃficas y Técnicas. Centro CientÃfico Tecnológico Conicet - BahÃa Blanca. Instituto de Matemática BahÃa Blanca. Universidad Nacional del Sur. Departamento de Matemática. Instituto de Matemática BahÃa Blanca; Argentina. Universidad Nacional del Sur. Departamento de EconomÃa; ArgentinaFil: Maguitman, Ana Gabriela. Consejo Nacional de Investigaciones CientÃficas y Técnicas. Centro CientÃfico Tecnológico Conicet - BahÃa Blanca. Instituto de Ciencias e IngenierÃa de la Computación. Universidad Nacional del Sur. Departamento de Ciencias e IngenierÃa de la Computación. Instituto de Ciencias e IngenierÃa de la Computación; Argentina. Universidad Nacional del Sur. Departamento de Ciencias e IngenierÃa de la Computación; ArgentinaXIX Simposio Argentino de Inteligencia ArtificialBuenos AiresArgentinaUniversidad de Palermo. Facultad de IngenierÃa. Asociación Argentina de Inteligencia Artificial. Sociedad Argentina de Informátic
- …