Message Understanding Conference (MUC) tests of discourse processing
Generate an AI Snapshot to get a quick, structured summary of this paper.
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
An early attempt to use the results from an information extraction evaluation to provide insight notice relationship between the difficulty of discourse processing and performance on the information extraction task and an upcoming noun phrase coreference evaluation is described.
Abstract
sundheim @ nose.rail Performance evaluations of NLP systems have been designed and conducted that require systems to extract cer-tain prespecified information about events and entities. A single text may describe multiple events and entities, and the evaluation task requires the system to resolve references to produce the expected output.We describe an early attempt to use the results from an information extraction evaluation to provide insight notice relationship between the difficulty of discourse processing and performance on the information extraction task. We then discuss an upcoming noun phrase coreference evaluation that has been designed indepen-dently of any other evaluation task in order to establish a clear performance benchmark on a small set of discourse phenomena. Background on the MUC Evaluations Five Message Understanding Conferences have been held since 1987 (Sundheim and Chinchor 1993) and a sixth one
