Finding the achilles heel of the web of data: Using network analysis for link-recommendation

Christophe Guéret, Paul Groth, Frank Van Harmelen, Stefan Schlobach

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

24 Scopus citations

Abstract

The Web of Data is increasingly becoming an important infrastructure for such diverse sectors as entertainment, government, e-commerce and science. As a result, the robustness of this Web of Data is now crucial. Prior studies show that the Web of Data is strongly dependent on a small number of central hubs, making it highly vulnerable to single points of failure. In this paper, we present concepts and algorithms to analyse and repair the brittleness of the Web of Data. We apply these on a substantial subset of it, the 2010 Billion Triple Challenge dataset. We first distinguish the physical structure of the Web of Data from its semantic structure. For both of these structures, we then calculate their robustness, taking betweenness centrality as a robustness-measure. To the best of our knowledge, this is the first time that such robustness-indicators have been calculated for the Web of Data. Finally, we determine which links should be added to the Web of Data in order to improve its robustness most effectively. We are able to determine such links by interpreting the question as a very large optimisation problem and deploying an evolutionary algorithm to solve this problem. We believe that with this work, we offer an effective method to analyse and improve the most important structure that the Semantic Web community has constructed to date.

Original languageEnglish
Title of host publicationThe Semantic Web, ISWC 2010 - 9th International Semantic Web Conference, ISWC 2010, Revised Selected Papers
PublisherSpringer Verlag
Pages289-304
Number of pages16
EditionPART 1
ISBN (Print)364217745X, 9783642177453
DOIs
StatePublished - 2010
Event9th International Semantic Web Conference, ISWC 2010 - Shanghai, China
Duration: Nov 7 2010Nov 11 2010

Publication series

NameLecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
NumberPART 1
Volume6496 LNCS
ISSN (Print)0302-9743
ISSN (Electronic)1611-3349

Conference

Conference9th International Semantic Web Conference, ISWC 2010
Country/TerritoryChina
CityShanghai
Period11/7/1011/11/10

Fingerprint

Dive into the research topics of 'Finding the achilles heel of the web of data: Using network analysis for link-recommendation'. Together they form a unique fingerprint.

Cite this