Learning to disambiguate potentially subjective expressions
Published 1 January 2002Open access
Janyce Wiebe, Theresa Wilson
Citations54
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
This paper focuses on disambiguating potentially subjective expressions in context, based on the density of other clues in the surrounding text.
Abstract
The goal of this work is recognizing opinionated and evaluative (subjective) language in text. The ability to recognize such language would be beneficial for many NLP applications such as question answering, information extraction, summarization, and genre detection. This paper focuses on disambiguating potentially subjective expressions in context, based on the density of other clues in the surrounding text.
Keywords
Computer Science
Automatic retrieval and clustering of similar words
1,555 Citations1998Dekang Lin
A word similarity measure based on the distributional pattern of words allows the automatically constructed thesaurus to be significantly closer to WordNet than Roget Thesaurus is.
Learning dictionaries for information extraction by multi-level bootstrapping
688 Citations1999Ellen Riloff, Rosie Jones
A multilevel bootstrapping algorithm is presented that generates both the semantic lexicon and extraction patterns simultaneously simultaneously and produces high-quality dictionaries for several semantic categories.
Building a discourse-tagged corpus in the framework of Rhetorical Structure Theory
600 Citations2001Lynn Carlson, Daniel Marcu +1 more
Working in the framework of Rhetorical Structure Theory, a large annotated resource with very high consistency is created, using a well-defined methodology and protocol to enable researchers to develop empirically grounded, discourse-specific applications.
Learning Subjective Adjectives from Corpora
520 Citations2000Janyce Wiebe
This paper identifies strong clues of subjectivity using the results of a method for clustering words according to distributional similarity (Lin 1998), seeded by a small amount of detailed manual annotation.
Development and use of a gold-standard data set for subjectivity classifications
498 Citations1999Janyce Wiebe, Rebecca Bruce +1 more
Bias-corrected tags are formulated and successfully used to guide a revision of the coding manual and develop an automatic classifier.
Automatic detection of text genre
358 Citations1997Brett Kessler, Geoffrey Numberg +1 more
A theory of genres as bundles of facets, which correlate with various surface cues, are proposed, and it is argued that genre detection based on surface cues is as successful as Detection based on deeper structural properties.
Journal of PragmaticsGenerating natural language under pragmatic constraints
340 Citations1987Eduard Hovy
arXiv (Cornell University)Tracking point of view in narrative
307 Citations1994Janyce Wiebe
This paper presents an algorithm to develop an algorithm that tracks point of view on the basis of the regularities found in naturally occurring narrative, and describes the results of some preliminary empirical studies, which lend support to the algorithm.
Identifying Collocations for Recognizing Opinions
135 Citations2001Jens Wiebe
Promising results are shown for a straightforward method of identifying collocational clues of subjectivity, as well as evidence of the usefulness of these clues for recognizing opinionated documents.
A corpus study of evaluative and speculative language
59 Citations2001Janyce Wiebe, Rebecca Bruce +3 more
This study yields knowledge needed to design effective machine learning systems for identifying subjective language, and analyses of annotator agreement and of characteristics of subjective language are performed.
What's yours and what's mine
42 Citations2000Simone Teufel, Marc Moens
The algorithm and a systematic evaluation of a system which can recognize the most salient textual properties that contribute to the global argumentative structure of a text are presented.
