Automatic acquisition of subcategorization frames from tagged text
Generate an AI Snapshot to get a quick, structured summary of this paper.
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
An implemented program that takes a tagged text corpus and generates a partial list of the subcategorization frames in which each verb occurs and expects to provide a large subc categorization dictionary to the NLP community and to train dictionaries for specific corpora.
Abstract
This paper describes an implemented program that takes a tagged text corpus and generates a partial list of the subcategorization frames in which each verb occurs. The completeness of the output list increases monotonically with the total occurrences of each verb in the training corpus. False positive rates are one to three percent. Five subcategorization frames are currently detected and we foresee no impediment to detecting many more. Ultimately, we expect to provide a large subcategorization dictionary to the NLP community and to train dictionaries for specific corpora.
