login

Multi-document summarization using sentence clustering

Published 1 December 2012
V.K. Gupta, Tanveer J. Siddiqui
Citations48

TL;DR

This paper presents an approach to query focused multi document summarization by combining single document summary using sentence clustering, and observed an average F-measure on DUC 2002 multi-document dataset, which is comparable to three best performing systems reported on the same dataset.

Abstract

This paper presents an approach to query focused multi document summarization by combining single document summary using sentence clustering. Both syntactic and semantic similarity between sentences is used for clustering. Single document summary is generated using document feature, sentence reference index feature, location feature and concept similarity feature. Sentences from single document summaries are clustered and top most sentences from each cluster are used for creating multi-document summary. We observed an average F-measure of 0.33774 on DUC 2002 multi-document dataset, which is comparable to three best performing systems reported on the same dataset.

Keywords

Computer Science