Employing Personal/Impersonal Views in Supervised and Semi-Supervised Sentiment Classification
PolyU Institutional Research Archive (Hong Kong Polytechnic University)Published 11 July 2010
Shoushan Li, Chu‐Ren Huang, Guodong Zhou, Sophia Yat Mei Lee
Citations124
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
An ensemble method and a co-training algorithm are explored to employ the two views, personal and impersonal views, in both supervised and semi-supervised sentiment classification respectively.
Abstract
48th Annual Meeting of the Association for Computational Linguistics, ACL 2010, Uppsala, 11-16 July 2010
Keywords
Computer Science
Thumbs up?
6,987 Citations2002Bo Pang, Lillian Lee +1 more
This work considers the problem of classifying documents not by topic, but by overall sentiment, e.g., determining whether a review is positive or negative, and concludes by examining factors that make the sentiment classification problem more challenging.
Combining labeled and unlabeled data with co-training
5,604 Citations1998Avrim Blum, Tom M. Mitchell
IEEE Transactions on Pattern Analysis and Machine IntelligenceOn combining classifiers
5,227 Citations1998Josef Kittler, M. Hatef +2 more
Minds at UW (University of Wisconsin)Active Learning Literature Survey
4,776 Citations2009Burr Settles
This report provides a general introduction to active learning and a survey of the literature, including a discussion of the scenarios in which queries can be formulated, and an overview of the query strategy frameworks proposed in the literature to date.
Minds at UW (University of Wisconsin)Semi-Supervised Learning Literature Survey
3,868 Citations2005Xiaojin Zhu
The study clearly indicates that the common practice of stripwise precommercial thinning is unjustified, and the justification of heavy 'chessboard' thinning (with pruning) depends on whether the potential reduction in rotation length and the improvement in wood quality outweigh the discounted costs of pre-commercial thinning and selection and pruning of crop trees.
Meeting of the Association for Computational LinguisticsThumbs Up or Thumbs Down? Semantic Orientation Applied to Unsupervised Classification of Reviews
3,654 Citations2002Peter Peter, Turney
A sentimental education
3,343 Citations2004Bo Pang, Lillian Lee
A novel machine-learning method is proposed that applies text-categorization techniques to just the subjective portions of the document, which greatly facilitates incorporation of cross-sentence contextual constraints.
Transductive Inference for Text Classification using Support Vector Machines
2,717 Citations1999Thorsten Joachims
An analysis of why TSVMs are well suited for text classi(cid:12)cation is presented, and an algorithm for training TSVMs e(cid:14)-ciently, handling 10,000 examples and more is proposed.
A re-examination of text categorization methods
2,660 Citations1999Yiming Yang, Xin Liu
The results show that SVM, kNN and LLSF signi cantly outperform NNet and NB when the number of positive training instances per category are small, and that all the methods perform comparably when the categories are over 300 instances.
Biographies, Bollywood, Boom-boxes and Blenders: Domain Adaptation for Sentiment Classification
2,026 Citations2007John Blitzer, Mark Dredze +1 more
This work extends to sentiment classification the recently-proposed structural correspondence learning (SCL) algorithm, reducing the relative error due to adaptation between domains by an average of 30% over the original SCL algorithm and 46% over a supervised baseline.
Opinion observer
1,604 Citations2005Bing Liu, Minqing Hu +1 more
A novel framework for analyzing and comparing consumer opinions of competing products is proposed, and a new technique based on language pattern mining is proposed to extract product features from Pros and Cons in a particular type of reviews.
Determining the sentiment of opinions
1,485 Citations2004Soo-Min Kim, Eduard Hovy
A system that, given a topic, automatically finds the people who hold opinions about that topic and the sentiment of each opinion and another module for determining word sentiment and another for combining sentiments within a sentence is presented.
Cambridge University Press eBooksThe Cambridge Encyclopedia of the English Language
1,167 Citations2018David Crystal
Artificial Intelligence ReviewA Perspective View and Survey of Meta-Learning
1,114 Citations2002Ricardo Vilalta, Youssef Drissi
This paper provides its own perspective view in which the goal is to build self-adaptive learners that improve their bias dynamically through experience by accumulating meta-knowledge, and provides a survey of meta-learning as reported by the machine-learning literature.
Machine LearningIs Combining Classifiers with Stacking Better than Selecting the Best One?
891 Citations2004Sašo Džeroski, Bernard Ženko
This work empirically evaluates several state-of-the-art methods for constructing ensembles of heterogeneous classifiers with stacking and shows that they perform (at best) comparably to selecting the best classifier from the ensemble by cross validation and proposes two extensions of this method using an extended set of meta-level features and multi-response model trees to learn at the meta- level.
Computational IntelligenceSENTIMENT CLASSIFICATION of MOVIE REVIEWS USING CONTEXTUAL VALENCE SHIFTERS
750 Citations2006Alistair Kennedy, Diana Inkpen
It is shown that extending the term‐counting method with contextual valence shifters improves the accuracy of the classification, and combining the two methods achieves better results than either method alone.
Computational LinguisticsRecognizing Contextual Polarity: An Exploration of Features for Phrase-Level Sentiment Analysis
731 Citations2009Theresa Wilson, Janyce Wiebe +1 more
The goal of this work is to automatically distinguish between prior and contextual polarity, with a focus on understanding which features are important for this task, and it is shown that the presence of neutral instances greatly degrades the performance of features for distinguishing between positive and negative polarity.
Co-training for cross-lingual sentiment classification
499 Citations2009Xiaojun Wan
A cotraining approach is proposed to making use of unlabeled Chinese data for cross-lingual sentiment classification, which leverages an available English corpus for Chinese sentiment classification by using the English corpus as training data.
Determining the semantic orientation of terms through gloss classification
388 Citations2005Andrea Esuli, Fabrizio Sebastiani
This paper presents a new method for determining the orientation of subjective terms based on the quantitative analysis of the glosses of such terms given in on-line dictionaries, and on the use of the resulting term representations for semi-supervised term classification.
The combining classifier: to train or not to train?
328 Citations2003Robert P. W. Duin
An intuitive discussion on the use of trained combiners is presented, relating the question of the choice of the combining classifier to a similar choice in the area of dissimilarity based pattern recognition.
IEEE Transactions on Pattern Analysis and Machine IntelligenceA theoretical and experimental analysis of linear combiners for multiple classifier systems
309 Citations2005Giorgio Fumera, Fabio Roli
This analysis focuses on the simplest and most widely used implementation of linear combiners, which consists of assigning a nonnegative weight to each individual classifier, and considers the ideal performance of this combining rule, i.e., that achievable when the optimal values of the weights are used.
Structured Models for Fine-to-Coarse Sentiment Analysis
286 Citations2007Ryan McDonald, Kerry Hannan +3 more
Experiments show that this structured model for jointly classifying the sentiment of text at varying levels of granularity can significantly reduce classification error relative to models trained in isolation.
Feature subsumption for opinion analysis
263 Citations2006Ellen Riloff, Siddharth Patwardhan +1 more
It is shown that reducing the feature set improves performance on three opinion classification tasks, especially when combined with traditional feature selection.
Mine the easy, classify the hard
165 Citations2009Sajib Dasgupta, Vincent Ng
This work proposes a semi-supervised approach to sentiment classification where it first mine the unambiguous reviews using spectral techniques and then exploit them to classify the ambiguous reviews via a novel combination of active learning, transductive learning, and ensemble learning.
Automatic seed word selection for unsupervised sentiment classification of Chinese text
139 Citations2008Taras Zagibalov, John M. Carroll
A new method of automatic seed word selection for un-supervised sentiment classification of product reviews in Chinese that is close to those of supervised classifiers and sometimes better, up to an F1 of 92%.
Lecture notes in computer sciencePredicting the Political Sentiment of Web Log Posts Using Supervised Machine Learning Techniques Coupled with Feature Selection
62 Citations2007Kathleen T. Durant, Michael D. Smith
It is shown that a Naive Bayes classifier coupled with a forward feature selection technique can on average correctly predict a posting's sentiment 89.77% of the time with a standard deviation of 3.01.
