A random walk on the red carpet
Published 26 October 2008
Derry Wijaya, Stéphane Bressan
Citations28
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
The approach is illustrated and evaluated using reviews of box office movies by users of a popular movie review site, showing that the approach is very effective and that the ranking it computes is comparable to the ranking obtained from the box office figures.
Abstract
10.1145/1458082.1458207
Keywords
Computer Science
Computer Networks and ISDN SystemsThe anatomy of a large-scale hypertextual Web search engine
15,828 Citations1998Sergey Brin, Lawrence M. Page
This paper provides an in-depth description of Google, a prototype of a large-scale search engine which makes heavy use of the structure present in hypertext and looks at the problem of how to effectively deal with uncontrolled hypertext collections where anyone can publish anything they want.
Meeting of the Association for Computational LinguisticsThumbs Up or Thumbs Down? Semantic Orientation Applied to Unsupervised Classification of Reviews
3,654 Citations2002Peter Peter, Turney
Recognizing contextual polarity in phrase-level sentiment analysis
3,379 Citations2005Theresa Wilson, Janyce Wiebe +1 more
A new approach to phrase-level sentiment analysis is presented that first determines whether an expression is neutral or polar and then disambiguates the polarity of the polar expressions.
SENTIWORDNET: A Publicly Available Lexical Resource for Opinion Mining
2,489 Citations2006Andrea Esuli, Fabrizio Sebastiani
SENTIWORDNET is a lexical resource in which each WORDNET synset is associated to three numerical scores Obj, Pos and Neg, describing how objective, positive, and negative the terms contained in the synset are.
Mining the peanut gallery
1,908 Citations2003Kushal Dave, Steve Lawrence +1 more
This work develops a method for automatically distinguishing between positive and negative reviews and draws on information retrieval techniques for feature extraction and scoring, and the results for various metrics and heuristics vary depending on the testing situation.
Opinion observer
1,604 Citations2005Bing Liu, Minqing Hu +1 more
A novel framework for analyzing and comparing consumer opinions of competing products is proposed, and a new technique based on language pattern mining is proposed to extract product features from Pros and Cons in a particular type of reviews.
Transformation-based error-driven learning and natural language processing: a case study in part-of-speech tagging
1,531 Citations1995Eric Brill
This paper describes a simple rule-based approach to automated learning of linguistic knowledge that has been shown for a number of tasks to capture information in a clearer and more direct fashion without a compromise in performance.
Predicting the semantic orientation of adjectives
1,439 Citations1997Vasileios Hatzivassiloglou, Kathleen McKeown
A log-linear regression model uses constraints from conjunctions to predict whether conjoined adjectives are of same or different orientations, achieving 82% accuracy in this task when each conjunction is considered independently.
Mining opinion features in customer reviews
1,273 Citations2004Minqing Hu, Bing Liu
This project aims to summarize all the customer reviews of a product by mining opinion/product features that the reviewers have commented on and a number of techniques are presented to mine such features.
Graphs And Algorithms
621 Citations1984Michel Gondran, Michel Minoux +1 more
Presents a review of graph theory, analyzing the existing links between abstract theoretical results and their practical implications using graph theoretical models and combinatorial algorithms.
A maximum entropy approach to identifying sentence boundaries
393 Citations1997Jeffrey C. Reynar, Adwait Ratnaparkhi
A trainable model for identifying sentence boundaries in raw text that can be trained easily on any genre of English, and should be trainable on any other Romanalphabet language.
Lecture notes in computer scienceEfficient k-Anonymization Using Clustering Techniques
329 Citations2007Ji-Won Byun, Ashish Kamra +2 more
An approach that uses the idea of clustering to minimize information loss and thus ensure good data quality is proposed, and a suitable metric to estimate the information loss introduced by generalizations is developed, which works for both numeric and categorical data.
The utility of linguistic rules in opinion mining
209 Citations2007Xiaowen Ding, Bing Liu
This paper proposes to use some linguistic rules to deal with the problem of determining the semantic orientations (positive or negative) of opinions expressed on product features in reviews together with a new opinion aggregation function.
Opinion Mining using Econometrics: A Case Study on Reputation Systems
150 Citations2007Anindya Ghose, Panagiotis G. Ipeirotis +1 more
Econometrics is used to identify the “economic value of text” and assign a “dollar value” to each opinion phrase, measuring sentiment effectively and without the need for manual labeling, and argues that by interpreting opinions using econometricrics, it has the first objective, quantifiable, and contextsensitive evaluation of opinions.
PageRanking WordNet Synsets: An Application to Opinion Mining
148 Citations2007Andrea Esuli, Fabrizio Sebastiani
This paper presents an application of PageRank, a random-walk model originally devised for ranking Web search results, to ranking WordNet synsets in terms of how strongly they possess a given semantic property.
Computer NetworksMethods for comparing rankings of search engine results
144 Citations2005Judit Bar‐Ilan, Mazlita Mat‐Hassan +1 more
A number of measures that compare rankings of search engine results are presented, which apply to five queries that were monitored daily for two periods of 14 or 21 days each.
Identifying Collocations for Recognizing Opinions
135 Citations2001Jens Wiebe
Promising results are shown for a straightforward method of identifying collocational clues of subjectivity, as well as evidence of the usefulness of these clues for recognizing opinionated documents.
Using Appraisal Taxonomies for Sentiment Analysis
73 Citations2005Casey Whitelaw
Non-topical text analysis, in which characterizations are sought of the opinions, feelings, and attitudes expressed in a text, rather than just the facts, is seen as a growing interest.
Visual and Multimedia Information Management
6 Citations2002Xiaofang Zhou, Pearl Pu
The following topics are dealt with: hyperdatabases; federated information systems; relevance feedback in CBIR; Oracle Chart Builder and MapViewer; SOM-based k-nearest neighbors search; partial image retrieval; high-dimensional image indexing; design and visualization of active capability; multimedia displays.
