000123859 001__ 123859
000123859 005__ 20230706131755.0
000123859 0247_ $$2doi$$a10.3233/SW-200368
000123859 0248_ $$2sideral$$a131855
000123859 037__ $$aART-2020-131855
000123859 041__ $$aeng
000123859 100__ $$0(orcid)0000-0003-4239-8785$$aBobed, C.$$uUniversidad de Zaragoza
000123859 245__ $$aData-driven assessment of structural evolution of RDF graphs
000123859 260__ $$c2020
000123859 5060_ $$aAccess copy available to the general public$$fUnrestricted
000123859 5203_ $$aSince the birth of the Semantic Web, numerous knowledge bases have appeared. The applications that exploit them rely on the quality of their data through time. In this regard, one of the main dimensions of data quality is conformance to the expected usage of the vocabulary. However, the vocabulary usage (i.e., how classes and properties are actually populated) can vary from one base to another. Moreover, through time, such usage can evolve within a base and diverges from the previous practices. Methods have been proposed to follow the evolution of a knowledge base by the observation of the changes of their intentional schema (or ontology); however, they do not capture the evolution of their actual data, which can vary greatly in practice. In this paper, we propose a data-driven approach to assess the global evolution of vocabulary usage in large RDF graphs. Our proposal relies on two structural measures defined at different granularities (dataset vs update), which are based on pattern mining techniques. We have performed a thorough experimentation which shows that our approach is scalable, and can capture structural evolution through time of both synthetic (LUBM) and real knowledge bases (different snapshots and updates of DBpedia).
000123859 540__ $$9info:eu-repo/semantics/openAccess$$aAll rights reserved$$uhttp://www.europeana.eu/rights/rr-f/
000123859 655_4 $$ainfo:eu-repo/semantics/article$$vinfo:eu-repo/semantics/acceptedVersion
000123859 700__ $$aMaillot, P.
000123859 700__ $$aCellier, P.
000123859 700__ $$aFerré, S.
000123859 7102_ $$15007$$2570$$aUniversidad de Zaragoza$$bDpto. Informát.Ingenie.Sistms.$$cÁrea Lenguajes y Sistemas Inf.
000123859 773__ $$g11, 5 (2020), 831-853$$tSemantic Web$$x2210-4968
000123859 8564_ $$s704675$$uhttps://zaguan.unizar.es/record/123859/files/texto_completo.pdf$$yPostprint
000123859 8564_ $$s2139420$$uhttps://zaguan.unizar.es/record/123859/files/texto_completo.jpg?subformat=icon$$xicon$$yPostprint
000123859 909CO $$ooai:zaguan.unizar.es:123859$$particulos$$pdriver
000123859 951__ $$a2023-07-06-12:25:27
000123859 980__ $$aARTICLE