Semantic Sentiment Analyses Based on Reputations of Web Information Sources
Auerbach Publications eBooksPublished 10 August 2011
Donato Barbagallo, Cinzia Cappiello, Chiara Francalanci, Maristella Mate
Citations6
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
Abstract
Contents 13.1 Introduction ............................................................................................ 348 13.2 Semantic Representations .........................................................................350 13.3 Layered Model of Semantic Search...........................................................352 13.4 Index Matching Approaches .....................................................................353
Keywords
Computer SciencePhysics and Astronomy
Computer Networks and ISDN SystemsThe anatomy of a large-scale hypertextual Web search engine
15,828 Citations1998Sergey Brin, Lawrence M. Page
This paper provides an in-depth description of Google, a prototype of a large-scale search engine which makes heavy use of the structure present in hypertext and looks at the problem of how to effectively deal with uncontrolled hypertext collections where anyone can publish anything they want.
Language<b>WordNet: An electronic lexical database</b> . Ed. by Christiane Fellbaum. Cambridge, MA: MIT Press, 1998. Pp. xxii, 423.
11,687 Citations2000Adam Kilgarriff
The lexical database: nouns in WordNet, George A. Miller modifiers in WordNet, Katherine J. Miller a semantic network of English verbs, and applications of WordNet: building semantic concordances are presented.
SENTIWORDNET: A Publicly Available Lexical Resource for Opinion Mining
2,489 Citations2006Andrea Esuli, Fabrizio Sebastiani
SENTIWORDNET is a lexical resource in which each WORDNET synset is associated to three numerical scores Obj, Pos and Neg, describing how objective, positive, and negative the terms contained in the synset are.
ACM Computing SurveysMethodologies for data quality assessment and improvement
1,251 Citations2009Carlo Batini, Cinzia Cappiello +2 more
Methodologies are compared along several dimensions, including the methodological phases and steps, the strategies and techniques, the data quality dimensions, the types of data, and, finally, thetypes of information systems addressed by each methodology.
Topic sentiment mixture
813 Citations2007Qiaozhu Mei, Xu Ling +3 more
The proposed Topic-Sentiment Mixture (TSM) model can reveal the latent topical facets in a Weblog collection, the subtopics in the results of an ad hoc query, and their associated sentiments and could also provide general sentiment models that are applicable to any ad hoc topics.
Journal of Web SemanticsA survey of trust in computer science and the Semantic Web
674 Citations2007Donovan Artz, Yolanda Gil
This paper gives an overview of existing trust research in computer science and the Semantic Web.
Information Technology and ManagementDo online reviews affect product sales? The role of reviewer characteristics and temporal effects
602 Citations2008Nan Hu, Ling Liu +1 more
International Conference on Weblogs and Social MediaLarge-Scale Sentiment Analysis for News and Blogs
558 Citations2007Namrata Godbole, Manjunath Srinivasaiah +1 more
ACM Computing SurveysHubs, authorities, and communities
504 Citations1999Jon Kleinberg
The Web has become the most visible manifestation of a new medium: a global, populist hypertext, where researchers now release their results to the Web before they appear in print, and corporations list their URLs alongside their toll-free numbers.
IEEE Internet ComputingUnderstanding Mashup Development
460 Citations2008Jin Yu, Boualem Benatallah +2 more
Current tools, frameworks, and trends that aim to facilitate mashup development are overviewed and a set of characteristic dimensions are used to highlight the strengths and weaknesses of some representative approaches.
Using SentiWordNet for multilingual sentiment analysis
362 Citations2008Kerstin Denecke
The results show that working with standard technology and existing sentiment analysis approaches is a viable approach to sentiment analysis within a multilingual framework.
ARSA
309 Citations2007Yang Liu, Xiangji Huang +2 more
ARSA is presented, an autoregressive sentiment-aware model, to utilize the sentiment information captured by S-PLSA for predicting product sales performance and is compared with alternative models that do not take into account the sentiment Information.
Conference of the European Chapter of the Association for Computational LinguisticsDetermining Term Subjectivity and Term Orientation for Opinion Mining
306 Citations2006Andrea Esuli, Fabrizio Sebastiani
The task of deciding whether a given term has a positive connotations, or a negative connotation, or has no subjective connotation at all is confronted, and it is shown that determining subjectivity and orientation is a much harder problem than determining orientation alone.
ACM Transactions on Information SystemsRepeatable evaluation of search services in dynamic environments
271 Citations2007Eric C. Jensen, Steven M. Beitzel +2 more
The bootstrap estimate of the reproducibility probability of hypothesis tests is leveraged in determining the query sample sizes required to ensure this, finding they are much larger than those required for static collections.
The importance of stop word removal on recall values in text categorization
205 Citations2004Catarina Silva, Bernardete Ribeiro
The purpose is to determine the importance of several basic reduction techniques on Support Vector Machines, by comparing their relative performance improvement when applied on the standard REUTERS-21578 benchmark.
Very Large Data BasesEnterprise information mashups: integrating information, simply
183 Citations2006Anant Jhingran
This talk describes the fundamental transformation that is taking place on the web around information composition through mashups, and asserts that this will also affect enterprise architectures and call it an enterprise information mashup fabric.
FreeLing 1.3: Syntactic and semantic services in an open-source NLP library
175 Citations2006Jordi Atserias, Bernardino Casas +4 more
This paper describes version 1.3 of the FreeLing suite of NLP tools, which has been improved and enlarged to cover more languages and offer more services: Named entity recognition and classification, chunking, dependency parsing, and WordNet based semantic annotation.
More than words
172 Citations2009Stephan Greene, Philip Resnik
A strong predictive connection between linguistically well motivated features and implicit sentiment is established, and it is shown how computational approximations of these features can be used to improve on existing state-of-the-art sentiment classification results.
Deriving marketing intelligence from online discussion
171 Citations2005Natalie Glance, Matthew Hurst +4 more
It is argued that applications for mining large volumes of textual data for marketing intelligence should provide two key elements: a suite of powerful mining and visualization technologies and an interactive analysis environment which allows for rapid generation and testing of hypotheses.
A framework for rapid integration of presentation components
155 Citations2007Jin Yu, Boualem Benatallah +4 more
This paper proposes a framework for the integration of stand-alone modules or applications, where integration occurs at the presentation layer, and provides an abstract component model to specify characteristics and behaviors of presentation components and an event-based composition model to specifying the composition logic.
MashMaker
151 Citations2007Robert Ennals, Minos Garofalakis
Natural Language EngineeringThe role of domain information in Word Sense Disambiguation
142 Citations2002Bernardo Magnini, Carlo Strapparava +2 more
Results obtained at the SENSEVAL-2 initiative confirm that for a significant subset of words domain information can be used to disambiguate with a very high level of precision.
Web Science 2.0: Identifying Trends through Semantic Social Network Analysis
129 Citations2009Peter A. Gloor, Jonas Krauß +3 more
A novel set of social network analysis based algorithms for mining the Web, blogs, and online forums to identify trends and find the people launching these new trends to predict long-term trends on the popularity of relevant concepts such as brands, movies, and politicians are introduced.
Topic-dependent sentiment analysis of financial blogs
110 Citations2009Neil O’Hare, Michael Davy +5 more
This work develops a corpus of financial blogs, annotated with polarity of sentiment with respect to a number of companies, and proposes text extraction techniques to create topic-specific sub-documents, which are used to train a sentiment classifier.
Lecture notes in computer scienceA Quality Model for Mashup Components
60 Citations2009Cinzia Cappiello, Florian Daniel +1 more
The quality properties of mashup components (APIs), the building blocks of any mashup application, are analyzed and a quality model is defined, which is claimed represents a valuable instrument in the hands of both component developers and mashup composers.
Lecture notes in computer scienceTrust and Reputation Mining in Professional Virtual Communities
41 Citations2009Florian Skopik, Hong‐Linh Truong +1 more
A system which determines trust relationships between community members automatically and objectively by mining communication data is proposed, and by applying natural language processing on log files, a new approach to make contributions visible is followed.
Using syntactic and contextual information for sentiment polarity analysis
23 Citations2009Shaishav Agrawal, Tanveer J. Siddiqui
A new method for sentiment polarity analysis that first assigns scores to a sentence using SentiWordNet and then uses heuristics to handle context dependent sentiment expressions shows significant improvement on movie-review dataset over the baseline.
Lecture notes in computer scienceWeb Site Evaluation: Methodology and Case Study
18 Citations2002Paolo Atzeni, Paolo Merialdo +1 more
A hierarchical model is proposed, which comprises several quality attributes of Web sites, and is designed for use by independent analysts who may have no knowledge of the technology underlying the site nor any contact with the site managers.
Empirical Analysis of the Rank Distribution of Relevant Documents in Web Search
7 Citations2008Jiang Shen, Sandra Zilles +1 more
An empirical approach for analyzing the rank distribution of relevant documents in Web search is proposed and a new transaction log analysis method is proposed; the relevance of documents is studied over transaction sessions rather than single transactions.
Virtual Community of Pathological Anatomy (University of Castilla La Mancha)A Broker for Selecting and Provisioning High Quality Syndicated Data.
6 Citations2005Danilo Ardagna, Cinzia Cappiello +3 more
This paper proposes a broker architecture that works as an intermediary between users and syndicated data providers on the basis of data quality and cost requirements, and builds the most suitable data set by integrating data from different providers.
Virtual Community of Pathological Anatomy (University of Castilla La Mancha)A Reputation-based DSS: the INTEREST Approach
4 Citations2010Donato Barbagallo, Stefania Bruno +4 more
This paper proposes a platform for the construction of self-service environments for the selection and composition of trustworthy services for information access.
Lecture notes in computer scienceBrokering Multisource Data with Quality Constraints
1 Citations2006Danilo Ardagna, Cinzia Cappiello +2 more
A data quality perspective on data brokering is taken and data accuracy is considered, which assumes that actual data are transparent to the broker and results comparing the delta between the data visibility and transparency approaches to data Brokering are presented.
