WordNet: similarity - measuring the relatedness of concepts
Published 25 July 2004
Ted Pedersen, Siddharth Patwardhan, Jason Michelizzi
Citations866
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
Abstract
WordNet::Similarity is a freely available software package that makes it possible to measure the semantic similarity or relatedness between a pair of concepts (or word senses). It provides six measures of similarity, and three measures of relatedness, all of which are based on the lexical database WordNet. These measures are implemented as Perl modules which take as input two concepts, and return a numeric value that represents the degree to which they are similar or related.
Keywords
Computer Science
An Information-Theoretic Definition of Similarity
3,662 Citations1998Dekang Lin
This work presents an informationtheoretic definition of similarity that is applicable as long as there is a probabilistic model and demonstrates how this definition can be used to measure the similarity in a number of different domains.
Verbs semantics and lexical selection
3,086 Citations1994Zhibiao Wu, Martha Palmer
This paper will focus on the semantic representation of verbs in computer systems and its impact on lexical selection problems in machine translation (MT), and sees the approach as closely aligned with knowledge-based MT approaches (KBMT), and as a separate component that could be incorporated into existing systems.
arXiv (Cornell University)Semantic Similarity Based on Corpus Statistics and Lexical Taxonomy
2,224 Citations1997Jay J. Jiang, David W. Conrath
arXiv (Cornell University)Using Information Content to Evaluate Semantic Similarity in a Taxonomy
2,147 Citations1995Philip Resnik
The MIT Press eBooksCombining Local Context and WordNet Similarity for Word Sense Identification
1,638 Citations1998Claudia Leacock
This chapter contains sections titled: Introducfion, Training and Testing Data, Experiment 1: The Local Context Classifier, Experiment 2: Measuring Word Similarity In Wordnet, and Combining Local Context and Wordnet Similarity Measures.
Lecture notes in computer scienceAn Adapted Lesk Algorithm for Word Sense Disambiguation Using WordNet
850 Citations2002Satanjeev Banerjee, Ted Pedersen
This paper presents an adaptation of Lesk's dictionary-based word sense disambiguation algorithm that uses the lexical database WordNet as the source of glosses for this approach, and attains an overall accuracy of 32%.
The MIT Press eBooksLexical Chains as Representations of Context for the Detection and Correction of Malapropisms
840 Citations1998Graeme Hirst
How lexical chains can be constructed by means of WordNet, and how they can be applied in one particularlinguistic task: the detection and correction of malapropisms is shown.
Extended gloss overlaps as a measure of semantic relatedness
719 Citations2003Satanjeev Banerjee, Ted Pedersen
A new measure of semantic relatedness between concepts that is based on the number of shared words (overlaps) in their definitions (glosses) and reasonably correlates to human judgments is presented.
Lecture notes in computer scienceUsing Measures of Semantic Relatedness for Word Sense Disambiguation
490 Citations2003Siddharth Patwardhan, Satanjeev Banerjee +1 more
This paper generalizes the Adapted Lesk Algorithm to a method of word sense disambiguation based on semantic relatedness and finds that the gloss overlaps of AdaptedLesk and the semantic distance measure of Jiang and Conrath (1997) result in the highest accuracy.
An empirical model of multiword expression decomposability
233 Citations2003Timothy Baldwin, Colin Bannard +2 more
A construction-inspecific model of multiword expression decomposability based on latent semantic analysis is presented, and evidence is furnished for the calculated similarities being correlated with the semantic relational content of WordNet.
Learning cross-document structural relationships using boosting
54 Citations2003Zhu Zhang, Jahna Otterbacher +1 more
An empirical study that uses boosting to classify CST relationships between sentence pairs extracted from topically related documents shows that the binary classifier for determining existence of structural relationships significantly outperforms the baseline.
WORD SENSE DISAMBIGUATION WITHIN A MULTILINGUAL FRAMEWORK
34 Citations2003Mona Diab
A practical and functional implementation of a basic idea common to research interest in defining word meanings in cross-linguistic terms, SALAAM for Sense Assignment Leveraging Alignment And Multilinguality, is proposed as a means of improving its performance on sense disambiguation for verbs in particular and other word types in general.
Ranking WordNet Senses Automatically
15 Citations2004Csrp, Diana McCarthy +1 more
This work presents work on the use of the WordNet similarity package and a thesaurus automatically acquired from raw textual corpora to rank WordNet noun senses automatically, and shows that the automatic ranking can be used to filter senses which are unseen or infrequent in the gold-standard.
