2008 OntologyBasedIEandIntegFromHeterDataSources

From GM-RKB
Jump to navigation Jump to search

Subject Headings: Ontology-based Information Extraction Task, SOBA System. Ontology-based natural language processing; Information extraction; Knowledge integration; Question answering.

Notes

  • This is a journal paper of their paper at the LREC 2006 conference.

Cited by

Quotes

Abstract

In this paper we present the design, implementation and evaluation of SOBA, a system for ontology-based information extraction from heterogeneous data resources, including plain text, tables and image captions. SOBA is capable of processing structured information, text and image captions to extract information and integrate it into a coherent knowledge base. To establish coherence, SOBA interlinks the information extracted from different sources and detects duplicate information. The knowledge base produced by SOBA can then be used to query for information contained in the different sources in an integrated and seamless manner. Overall, this allows for advanced retrieval functionality by which questions can be answered precisely. A further distinguishing feature of the SOBA system is that it straightforwardly integrates deep and shallow natural language processing to increase robustness and accuracy. We discuss the implementation and application of the SOBA system within the SmartWeb multimodal dialog system. In addition, we present a thorough evaluation of the different components of the system. However, an end-to-end evaluation of the whole SmartWeb system is out of the scope of this paper and has been presented elsewhere by the SmartWeb consortium.


References


,

 AuthorvolumeDate ValuetitletypejournaltitleUrldoinoteyear
2008 OntologyBasedIEandIntegFromHeterDataSourcesPaul Buitelaar
Philipp Cimiano
Anette Frank
Matthias Hartung
Stefania Racioppa
Ontology-based Information Extraction and Integration from Heterogeneous Data Sourceshttp://www.cimiano.de/Publications/2006/soba/soba.pdf10.1016/j.ijhcs.2008.07.007