Early steps toward web-scale information extraction with lodie

7Citations
Citations of this article
19Readers
Mendeley users who have this article in their library.

Abstract

Information extraction (IE) is the technique for transforming unstructured textual data into a structured representation that can be understood by machines. The exponential growth of the web generates an exceptional quantity of data for which automatic knowledge capture is essential. This work describes the methodology for web-scale information extraction in the linked open data information-extraction (LODIE) project and highlights results from the early experiments carried out in the initial phase of the project. LODIE aims to develop informationextraction techniques able to scale at web level and adapt to user information needs. The core idea behind LODIE is the usage of linked open data, a very large-scale information resource, as a ground-breaking solution for IE, which provides invaluable annotated data on a growing number of domains. This article has two objectives, first, describing the LODIE project as a whole and depicting its general challenges and directions; and second, describing some initial steps taken toward the general solution, focusing on a specific IE subtask, wrapper induction.

Cite

CITATION STYLE

APA

Gentile, A. L., Zhang, Z., & Ciravegna, F. (2015). Early steps toward web-scale information extraction with lodie. AI Magazine, 36(1), 55–64. https://doi.org/10.1609/aimag.v36i1.2567

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free