login

A second-order Hidden Markov Model for part-of-speech tagging

Published 1 January 1999Open access
Scott M. Thede, Mary P. Harper
Citations141
View PDF

TL;DR

An extension to the hidden Markov model for part-of-speech tagging using second-order approximations for both contextual and lexical probabilities that increases the accuracy of the tagger to state of the art levels is described.

Abstract

This paper describes an extension to the hidden Markov model for part-of-speech tagging using second-order approximations for both contextual and lexical probabilities. This model increases the accuracy of the tagger to state of the art levels. These approximations make use of more contextual information than standard statistical systems. New methods of smoothing the estimated probabilities are also introduced to address the sparse data problem.

Keywords

Computer Science