DBpedia commons: Structured multimedia metadata from the wikimedia commons

Gaurav Vaidya; Dimitris Kontokostas; Magnus Knuth; Jens Lehmann; Sebastian Hellmann

Conference ProceedingsOPEN ACCESS

DBpedia commons: Structured multimedia metadata from the wikimedia commons

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2015) 9367 281-289

DOI: 10.1007/978-3-319-25010-6_17

14Citations

16Readers

Abstract

The Wikimedia Commons is an online repository of over twenty-five million freely usable audio, video and still image files, including scanned books, historically significant photographs, animal recordings, illustrative figures and maps. Being volunteer-contributed, these media files have different amounts of descriptive metadata with varying degrees of accuracy. The DBpedia Information Extraction Framework is capable of parsing unstructured text into semi-structured data from Wikipedia and transforming it into RDF for general use, but so far it has only been used to extract encyclopedia-like content. In this paper, we describe the creation of the DBpedia Commons (DBc) dataset, which was achieved by an extension of the Extraction Framework to support knowledge extraction from Wikimedia Commons as amedia repository. To our knowledge, this is the first complete RDFization of the Wikimedia Commons and the largest media metadata RDF database in the LOD cloud.

Author supplied keywords

Cite

CITATION STYLE

APA

Vaidya, G., Kontokostas, D., Knuth, M., Lehmann, J., & Hellmann, S. (2015). DBpedia commons: Structured multimedia metadata from the wikimedia commons. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 9367, pp. 281–289). Springer Verlag. https://doi.org/10.1007/978-3-319-25010-6_17

DBpedia commons: Structured multimedia metadata from the wikimedia commons

Abstract

Author supplied keywords

Cite

Register to see more suggestions