login

Sentiment classification: a lexical similarity based approach for extracting subjectivity in documents

Information RetrievalPublished 11 February 2011Open access
Kiran Sarvabhotla, Prasad Pingali, Vasudeva Varma
Citations34
View PDF

TL;DR

This work proposes a simple and statistical methodology called review summary (RSUMM) and uses it in combination with well-known feature selection methods to extract subjectivity and the experimental results prove the effectiveness of the proposed methodology.

Abstract

With the growth of social media, document sentiment classification has become an active area of research in this decade. It can be viewed as a special case of topical classification applied only to subjective portions of a document (sources of sentiment). Hence, the key task in document sentiment classification is extracting subjectivity. Existing approaches to extract subjectivity rely heavily on linguistic resources such as sentiment lexicons and complex supervised patterns based on part-of-speech (POS) information. This makes the task of subjective feature extraction complex and resource dependent. In this work, we try to minimize the dependency on linguistic resources in sentiment classification. We propose a simple and statistical methodology called review summary (RSUMM) and use it in combination with well-known feature selection methods to extract subjectivity. Our experimental results on a movie review dataset prove the effectiveness of the proposed methodology.

Keywords

Computer Science