login

Exploiting Linked Open Data to Uncover Entity Types

Communications in computer and information sciencePublished 1 January 2015
Jie Gao, Suvodeep Mazumdar
Citations7
SJR quartileQ4
SJR score0.18
SNIP0.24

TL;DR

It is argued that a rich feature space can improve extraction accuracy and it is proposed to exploit Linked Open Data (LOD) for feature enrichment, which improves the accuracy of the supervised induction method and enables easy matching with the Dolce+DnS Ultra Lite ontology classes.

Abstract

Extracting structured information from text plays a crucial role in automatic knowledge acquisition and is at the core of any knowledge representation and reasoning system. Traditional methods rely on hand-crafted rules and are restricted by the performance of various linguistic pre-processing tools. More recent approaches rely on supervised learning of relations trained on labelled examples, which can be manually created or sometimes automatically generated (referred as distant supervision). We propose a supervised method for entity typing and alignment. We argue that a rich feature space can improve extraction accuracy and we propose to exploit Linked Open Data (LOD) for feature enrichment. Our approach is tested on task-2 of the Open Knowledge Extraction challenge, including automatic entity typing and alignment. Our approach demonstrate that by combining evidences derived from LOD (e.g. DBpedia) and conventional lexical resources (e.g. WordNet) (i) improves the accuracy of the supervised induction method and (ii) enables easy matching with the Dolce+DnS Ultra Lite ontology classes.

Keywords

Computer ScienceBiochemistry, Genetics and Molecular Biology