login

Sentiment Lexicon Creation from Lexical Resources

Lecture notes in business information processingPublished 1 January 2011
Bas Heerschop, Alexander Hogenboom, Flavius Frăsincar
Citations34
SJR quartileQ3
SJR score0.26
SNIP0.52

TL;DR

A corpus-based evaluation of several automated methods for creating lexicons, exploiting vast lexical resources, and considers propagating the sentiment of a seed set of words through semantic relations or through PageRank-based similarities, which turns out to outperform the others.

Abstract

Today's business information systems face the challenge of analyzing sentiment in massive data sets for supporting, e.g., reputation management. Many approaches rely on lexical resources containing words and their associated sentiment. We perform a corpus-based evaluation of several automated methods for creating such lexicons, exploiting vast lexical resources. We consider propagating the sentiment of a seed set of words through semantic relations or through PageRank-based similarities. We also consider a machine learning approach using an ensemble of classifiers. The latter approach turns out to outperform the others. However, PageRank-based propagation appears to yield a more robust sentiment classifier.

Keywords

Computer Science