login

Topic identification in natural language dialogues using neural networks

Published 1 January 2002Open access
Krista Lagus, Jukka Kuusisto
Citations21
View PDF

TL;DR

A probabilistic model is defined and different methods for model parameter estimation on a corpus of 189 dialogues are compared and the utilization of information regarding the position of the word in the utterance is found to improve the results.

Abstract

In human-computer interaction systems using natural language, the recognition of the topic from user's utterances is an important task. We examine two different perspectives to the problem of topic analysis needed for carrying out a successful dialogue. First, we apply self-organized document maps for modeling the broader subject of discourse based on the occurrence of content words in the dialogue context. On a Finnish corpus of 57 dialogues the method is shown to work well for recognizing subjects of longer dialogue segments, whereas for individual utterances the subject recognition history should perhaps be taken into account. Second, we attempt to identify topically relevant words in the utterances and thus locate the old information ('topic words') and new information ('focus words'). For this we define a probabilistic model and compare different methods for model parameter estimation on a corpus of 189 dialogues. Moreover, the utilization of information regarding the position of the word in the utterance is found to improve the results.

Keywords

Computer Science