Journal on Data Semantics

, Volume 1, Issue 3, pp 187–201

Linked Open Piracy: A Story about e-Science, Linked Data, and Statistics

  • Willem Robert van Hage
  • Marieke van Erp
  • Véronique Malaisé
Open Access
Original Article

DOI: 10.1007/s13740-012-0009-6

Cite this article as:
van Hage, W.R., van Erp, M. & Malaisé, V. J Data Semant (2012) 1: 187. doi:10.1007/s13740-012-0009-6


There is an abundance of semi-structured reports on events being written and made available on the World Wide Web on a daily basis. These reports are primarily meant for human use. A recent movement is the addition of RDF metadata to make automatic processing by computers easier. A fine example of this movement is the open government data initiative which, by representing data from spreadsheets and textual reports in RDF, strives to speed up the creation of geographical mashups and visual analytic applications. In this paper, we present a newly linked dataset and the method we used to automatically translate semi-structured reports on the Web to an RDF event model. We demonstrate how the semantic representation layer makes it possible to easily analyze and visualize the aggregated reports to answer domain questions through a SPARQL client for the R statistical programming language. We showcase our method on piracy attack reports issued by the International Chamber of Commerce (ICC-CCS). Our pipeline includes conversion of the reports to RDF, linking their parts to external resources from the linked open data cloud and exposing them to the Web.


Information extraction Metadata enrichment Linked data 
Download to read the full article text

Copyright information

© The Author(s) 2012

Authors and Affiliations

  • Willem Robert van Hage
    • 1
  • Marieke van Erp
    • 1
  • Véronique Malaisé
    • 2
  1. 1.Department of Computer ScienceVU University AmsterdamAmsterdamThe Netherlands
  2. 2.Elsevier Content Enrichment Center (CEC)AmsterdamThe Netherlands

Personalised recommendations