Statistical Analysis of the owl:sameAs Network for Aligning Concepts in the Linking Open Data Cloud [chapter]

Gianluca Correndo, Antonio Penta, Nicholas Gibbins, Nigel Shadbolt
2012 Lecture Notes in Computer Science  
The massively distributed publication of linked data has brought to the attention of scientific community the limitations of classic methods for achieving data integration and the opportunities of pushing the boundaries of the field by experimenting this collective enterprise that is the linking open data cloud. While reusing existing ontologies is the choice of preference, the exploitation of ontology alignments still is a required step for easing the burden of integrating heterogeneous data
more » ... ts. Alignments, even between the most used vocabularies, is still poorly supported in systems nowadays whereas links between instances are the most widely used means for bridging the gap between different data sets. We provide in this paper an account of our statistical and qualitative analysis of the network of instance level equivalences in the Linking Open Data Cloud (i.e. the sameAs network) in order to automatically compute alignments at the conceptual level. Moreover, we explore the effect of ontological information when adopting classical Jaccard methods to the ontology alignment task. Automating such task will allow in fact to achieve a clearer conceptual description of the data at the cloud level, while improving the level of integration between datasets.
doi:10.1007/978-3-642-32597-7_20 fatcat:u5szzm2nyzhnzkswj35fceomze