Summarization of Massive RDF Graphs Using Identifier Classification

被引:0
|
作者
dos Santos, Andre Fernandes [1 ]
Leal, Jose Paulo
机构
[1] Univ Porto, CRAGS, Porto, Portugal
关键词
knowledge graphs; graph summarization; namespaces; RDF;
D O I
10.1007/978-3-031-40960-8_8
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
The size of massive knowledge graphs (KGs) and the lack of prior information regarding the schemas, ontologies and vocabularies they use frequently makes them hard to understand and visualize. Graph summarization techniques can help by abstracting details of the original graph to produce a reduced summary that can more easily be explored. Identifiers often carry latent information which could be used for classification of the entities they represent. Particularly, IRI namespaces can be used to classify RDF resources. Namespaces, used in some RDF serialization formats as a shortening mechanism for resource IRIs, have no role in the semantics of RDF. Nevertheless, there is often a hidden meaning behind the decision of grouping resources under a common prefix and assigning an alias to it. We improved on previous work on a namespace-based approach to KG summarization that classifies resources using their namespaces. Producing the summary graph is fast, light on computing resources and requires no previous domain knowledge. The summary graph can be used to analyze the namespace interdependencies of the original graph. We also present chilon, a tool for calculating namespace-based KG summaries. Namespaces are gathered from explicit declarations in the graph serialization, community contributions or resource IRI prefix analysis. We applied chilon to publicly available KGs, used it to generate interactive visualizations of the summaries, and discuss the results obtained.
引用
收藏
页码:89 / 103
页数:15
相关论文
共 50 条
  • [31] Controlling Access to RDF Graphs
    Flouris, Giorgos
    Fundulaki, Irini
    Michou, Maria
    Antoniou, Grigoris
    [J]. FUTURE INTERNET-FIS 2010, 2010, 6369 : 107 - 117
  • [32] Theme identification in RDF graphs
    [J]. Ouksili, Hanane (Hanane@prism.uvsq.fr), 1600, Springer Verlag (8748):
  • [33] Theme Identification in RDF Graphs
    Ouksili, Hanane
    Kedad, Zoubida
    Lopes, Stephane
    [J]. MODEL AND DATA ENGINEERING, MEDI 2014, 2014, 8748 : 321 - 329
  • [34] Text Summarization Based on Classification Using ANFIS
    Kumar, Yogan Jaya
    Kang, Fong Jia
    Goh, Ong Sing
    Khan, Atif
    [J]. ADVANCED TOPICS IN INTELLIGENT INFORMATION AND DATABASE SYSTEMS, 2017, 710 : 405 - 417
  • [35] Video Summarization using Text Subjectivity Classification
    Moraes, Leonardo
    Marcacini, Ricardo Marcondes
    Goularte, Rudinei
    [J]. PROCEEDINGS OF THE 28TH BRAZILIAN SYMPOSIUM ON MULTIMEDIA AND THE WEB, WEBMEDIA 2022, 2022, : 133 - 141
  • [36] RDF Graph Summarization Based on Node Characteristic and Centrality
    Guo, Jimao
    Wang, Yi
    [J]. JOURNAL OF WEB ENGINEERING, 2022, 21 (07): : 2073 - 2094
  • [37] PageRank and Generic Entity Summarization for RDF Knowledge Bases
    Diefenbach, Dennis
    Thalhammer, Andreas
    [J]. SEMANTIC WEB (ESWC 2018), 2018, 10843 : 145 - 160
  • [38] Using Provenance to Efficiently Propagate SPARQL Updates on RDF Source Graphs
    Naja, Iman
    Gibbins, Nicholas
    [J]. PROVENANCE AND ANNOTATION OF DATA AND PROCESSES, IPAW 2018, 2018, 11017 : 158 - 170
  • [39] Improving query-based summarization using document graphs
    Mohamed, Ahmed A.
    Rajasekaran, Sanguthevar
    [J]. 2006 IEEE INTERNATIONAL SYMPOSIUM ON SIGNAL PROCESSING AND INFORMATION TECHNOLOGY, VOLS 1 AND 2, 2006, : 408 - +
  • [40] Flexible querying of fuzzy RDF annotations using fuzzy conceptual graphs
    Buche, Patrice
    Dibie-Barthelemy, Juliette
    Hignette, Gaelle
    [J]. CONCEPTUAL STRUCTURES: KNOWLEDGE VISUALIZATION AND REASONING, 2008, 5113 : 133 - 146