Search CORE

453 research outputs found

SentiBench - a benchmark comparison of state-of-the-practice sentiment analysis methods

In the last few years thousands of scientific papers have investigated sentiment analysis, several startups that measure opinions on real data have emerged and a number of innovative products related to this theme have been developed. There are multiple methods for measuring sentiments, including lexical-based and supervised machine learning methods. Despite the vast interest on the theme and wide popularity of some methods, it is unclear which one is better for identifying the polarity (i.e., positive or negative) of a message. Accordingly, there is a strong need to conduct a thorough apple-to-apple comparison of sentiment analysis methods, \textit{as they are used in practice}, across multiple datasets originated from different data sources. Such a comparison is key for understanding the potential limitations, advantages, and disadvantages of popular methods. This article aims at filling this gap by presenting a benchmark comparison of twenty-four popular sentiment analysis methods (which we call the state-of-the-practice methods). Our evaluation is based on a benchmark of eighteen labeled datasets, covering messages posted on social networks, movie and product reviews, as well as opinions and comments in news articles. Our results highlight the extent to which the prediction performance of these methods varies considerably across datasets. Aiming at boosting the development of this research area, we open the methods' codes and datasets used in this article, deploying them in a benchmark system, which provides an open API for accessing and comparing sentence-level sentiment analysis methods

arXiv.org e-Print Archive

Crossref

Springer - Publisher Connector

REPOSITORIO INSTITUCIONAL DA UFOP

Improving Spanish Polarity Classification Combining Different Linguistic Resources

Author: Cruz Mata Fermín
Martín Valdivia M. Teresa
Martínez Cámara Eugenio
Molina González M. Dolores
Ortega Rodríguez Francisco Javier
Ureña López L. Alfonso
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2015
Field of study

Sentiment analysis is a challenging task which is attracting the attention of researchers. However, most of work is only focused on English documents, perhaps due to the lack of linguistic resources for other languages. In this paper, we present several Spanish opinion mining resources in order to develop a polarity classification system. In addition, we propose the combination of different features extracted from each resource in order to train a classifier over two different opinion corpora. We prove that the integration of knowledge from several resources can improve the final Spanish polarity classification system. The good results encourage us to continue developing sentiment resources for Spanish, and studying the combination of features extracted from different resourcesMinisterio de Economía y Competitividad TIN2012-38536-C03-0Junta de Andalucía P11-TIC-7684Universidad de Jaén CEATIC-2013-0

idUS. Depósito de Investigación Universidad de Sevilla

BCS SGAI SMA 2013: the BCS SGAI workshop on social media analysis

Author
Publication venue: M. Jeusfeld
Publication date: 01/01/2013
Field of study

Portsmouth University Research Portal (Pure)

Attentional Encoder Network for Targeted Sentiment Classification

Author: Jiang Tao
Liu Zhiyue
Rao Yanghui
Song Youwei
Wang Jiahai
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/04/2019
Field of study

Targeted sentiment classification aims at determining the sentimental tendency towards specific targets. Most of the previous approaches model context and target words with RNN and attention. However, RNNs are difficult to parallelize and truncated backpropagation through time brings difficulty in remembering long-term patterns. To address this issue, this paper proposes an Attentional Encoder Network (AEN) which eschews recurrence and employs attention based encoders for the modeling between context and target. We raise the label unreliability issue and introduce label smoothing regularization. We also apply pre-trained BERT to this task and obtain new state-of-the-art results. Experiments and analysis demonstrate the effectiveness and lightweight of our model.Comment: 7 page

arXiv.org e-Print Archive

Crossref

European Central Bank Speeches: A Sentiment Analysis Case Study

Author: Gustavo de Oliveira Vital
Publication venue
Publication date: 20/10/2022
Field of study

Repositório Aberto da Universidade do Porto

Attention-based word embeddings using Artificial Bee Colony algorithm for aspect-level sentiment classification

Author: Ji Zhicheng
Palade Vasile
Wang Yan
Zhang Ming
Publication venue: 'Elsevier BV'
Publication date: 04/02/2021
Field of study

Coventry University Pure Portal