OntoNotes
Published 1 January 2006Open access
Eduard Hovy, Mitchell P. Marcus, Martha Palmer, Lance Ramshaw, Ralph Weischedel
Citations810
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
It is described the OntoNotes methodology and its result, a large multilingual richly-annotated corpus constructed at 90% interannotator agreement, which will be made available to the community during 2007.
Abstract
We describe the OntoNotes methodology and its result, a large multilingual richly-annotated corpus constructed at 90% interannotator agreement. An initial portion (300K words of English newswire and 250K words of Chinese newswire) will be made available to the community during 2007.
Keywords
Computer ScienceBiochemistry, Genetics and Molecular Biology
Building a Large Annotated Corpus of English: The Penn Treebank
7,528 Citations1993Mitchell P. Marcus
The Berkeley FrameNet Project
2,564 Citations1998Collin F. Baker, Charles J. Fillmore +1 more
This report will present the project's goals and workflow, and information about the computational tools that have been adapted or created in-house for this work.
Computational LinguisticsThe Proposition Bank: An Annotated Corpus of Semantic Roles
2,302 Citations2005Martha Palmer, Daniel Gildea +1 more
An automatic system for semantic role tagging trained on the corpus is described and the effect on its performance of various types of information is discussed, including a comparison of full syntactic parsing with a flat representation and the contribution of the empty trace categories of the treebank.
Towards a standard upper ontology
1,601 Citations2001Ian Niles, Adam Pease
The strategy used to create the current version of the SUMO is outlined, some of the challenges that were faced in constructing the ontology are discussed, and its most general concepts and the relations between them are described.
Lecture notes in computer scienceSweetening Ontologies with DOLCE
1,018 Citations2002Aldo Gangemi, Nicola Guarino +3 more
This paper introduces the DOLCE upper level ontology, the first module of a Foundational Ontologies Library being developed within the WonderWeb project, and suggests that such analysis could hopefully lead to an ?
The NomBank Project: An Interim Report
315 Citations2004Adam Meyers, Ruth Reeves +5 more
The NomBank project is described, a project that will provide argument structure for instances of common nouns in the Penn Treebank II corpus, including its speci(cid:2)cations and the process involved in creating the resource.
Natural Language EngineeringMaking fine-grained and coarse-grained sense distinctions, both manually and automatically
155 Citations2006Martha Palmer, Hoa Trang Dang +1 more
A persistent problem arising from polysemy is discussed: namely the difficulty of finding consistent criteria for making fine-grained sense distinctions, either manually or automatically, and well-defined sense groups can be of value in improving word sense disambiguation by both humans and machines.
Semantic role labeling using different syntactic views
136 Citations2005Sameer Pradhan, Wayne Ward +3 more
A state-of-the-art baseline semantic role labeling system based on Support Vector Machine classifiers is presented and improvements on this system are shown by adding new features including features extracted from dependency parses, performing feature selection and calibration and combining parses obtained from semantic parsers trained using different syntactic views.
Fully parsing the Penn Treebank
79 Citations2006Ryan Gabbard, Mitchell P. Marcus +1 more
A two stage parser that recovers Penn Treebank style syntactic analyses of new sentences including skeletal syntactic structure, and, for the first time, both function tags and empty categories is presented.
Cambridge University Press eBooksThe Omega ontology
57 Citations2010Andrew Philpot, Eduard Hovy +1 more
The Omega ontology, a large terminological ontology obtained by remerging WordNet and Mikrokosmos, adding information from various other sources, and subordinating the result to a newly designed feature-oriented upper model is presented.
Lecture notes in computer scienceTowards Robust High Performance Word Sense Disambiguation of English Verbs Using Rich Linguistic Features
28 Citations2005Jinying Chen, Martha Palmer
This paper shows that the WSD system using rich linguistic features achieved high accuracy in the classification of English SENSEVAL2 verbs for both fine-grained (64.6%) and coarse- grained (73.7%) senses.
Proposition Bank II: Delving Deeper
16 Citations2004Olga Babko-Malaya, Martha Palmer +3 more
An overview of the second phase of PropBank Annotation, PropBank II, which is being applied to English and Chinese, and includes (Neodavidsonian) eventuality variables, nominal references, sense tagging, and connections to the Penn Discourse Treebank (PDTB), a project for annotating discourse connectives and their arguments.
Lecture notes in computer scienceInterlingual Annotation for MT Development
8 Citations2004Florence Reeder, Bonnie J. Dorr +9 more
This paper describes the creation of an interlingua and the development of a corpus of semantically annotated text, to be validated in six languages and evaluated in several ways.
