Advertisement

DBpedia Commons: Structured Multimedia Metadata from the Wikimedia Commons

  • Gaurav Vaidya
  • Dimitris Kontokostas
  • Magnus Knuth
  • Jens Lehmann
  • Sebastian Hellmann
Conference paper
Part of the Lecture Notes in Computer Science book series (LNCS, volume 9367)

Abstract

The Wikimedia Commons is an online repository of over twenty-five million freely usable audio, video and still image files, including scanned books, historically significant photographs, animal recordings, illustrative figures and maps. Being volunteer-contributed, these media files have different amounts of descriptive metadata with varying degrees of accuracy. The DBpedia Information Extraction Framework is capable of parsing unstructured text into semi-structured data from Wikipedia and transforming it into RDF for general use, but so far it has only been used to extract encyclopedia-like content. In this paper, we describe the creation of the DBpedia Commons (DBc) dataset, which was achieved by an extension of the Extraction Framework to support knowledge extraction from Wikimedia Commons as a media repository. To our knowledge, this is the first complete RDFization of the Wikimedia Commons and the largest media metadata RDF database in the LOD cloud.

Keywords

Wikimedia commons DBpedia Multimedia RDF 

Preview

Unable to display preview. Download preview PDF.

Unable to display preview. Download preview PDF.

References

  1. 1.
    Brümmer, M., Baron, C., Ermilov, I., Freudenberg, M., Kontokostas, D., Hellmann, S.: DataID: towards semantically rich metadata for complex datasets. In: Proc. of the 10th International Conference on Semantic Systems, pp. 84–91. ACM (2014)Google Scholar
  2. 2.
    Kontokostas, D., Bratsas, C., Auer, S., Hellmann, S., Antoniou, I., Metakides, G.: Internationalization of Linked Data: The case of the Greek DBpedia edition. Web Semantics: Science, Services and Agents on the WWW 15, 51–61 (2012)CrossRefGoogle Scholar
  3. 3.
    Lehmann, J., Isele, R., Jakob, M., Jentzsch, A., Kontokostas, D., Mendes, P.N., Hellmann, S., Morsey, M., van Kleef, P., Auer, S., Bizer, C.: DBpedia - a large-scale, multilingual knowledge base extracted from wikipedia. Semantic Web Journal (2014)Google Scholar
  4. 4.
    Lukovnikov, D., Stadler, C., Kontokostas, D., Hellmann, S., Lehmann, J.: DBpedia viewer - an integrative interface for DBpedia leveraging the DBpedia service eco system. In: Proc. of the Linked Data on the Web 2014 Workshop (2014)Google Scholar
  5. 5.
    Troncy, R., Mannens, E., Pfeiffer, S., Deursen, D.V.: Media fragments URI. W3C Recommendation, September 2012. http://www.w3.org/TR/media-frags/

Copyright information

© Springer International Publishing Switzerland 2015

Authors and Affiliations

  • Gaurav Vaidya
    • 1
  • Dimitris Kontokostas
    • 2
  • Magnus Knuth
    • 3
  • Jens Lehmann
    • 2
  • Sebastian Hellmann
    • 2
  1. 1.University of Colorado BoulderColoradoUSA
  2. 2.University of Leipzig, Computer Science, AKSWLeipzigGermany
  3. 3.Hasso Plattner InstituteUniversity of PotsdamPotsdamGermany

Personalised recommendations