login

A domain-independent semantic tagger for the study of meaning associations in English text

White Rose Research Online (University of Leeds, The University of Sheffield, University of York)Published 1 January 2001
George Demetriou, Eric Atwell
Citations9

TL;DR

It is proposed that a domain-independent semantic tagger for English corpora should not aim to annotate each word with an atomic 'sem-tag', but instead that a semantic tagging should attach to each word a set of semantic primitive attributes or features.

Abstract

A comparison of semantic tagging with syntactic Part-of-Speech tagging leads us to propose that a domain-independent semantic tagger for English corpora should not aim to annotate each word with an atomic `sem-tag', but instead that a semantic tagging should attach to each word a set of semantic primitive attributes or features. These features should include: - lemma or root, grouping together inflected and derived forms of the same lexical item; - broad subject categories where - selectional restrictions where - a meaning definition, stated in terms of a restricted Defining Vocabulary, and processed to remove stoplist-words and repetitions.

Keywords

PsychologyComputer ScienceArts and Humanities