KaBOB: ontology-based semantic integration of biomedical databases

被引:49
|
作者
Livingston, Kevin M. [1 ]
Bada, Michael [1 ]
Baumgartner, William A., Jr. [1 ]
Hunter, Lawrence E. [1 ]
机构
[1] Univ Colorado, Computat Biosci Program, Aurora, CO 80045 USA
来源
BMC BIOINFORMATICS | 2015年 / 16卷
关键词
Knowledge representation and reasoning; Semantic data integration; Biomedical; Databases; Open biomedical ontologies; Semantic web; OWL; RDF; COMMUNITY STANDARD; KNOWLEDGE; ACCESS; FORMAT;
D O I
10.1186/s12859-015-0559-3
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Background: The ability to query many independent biological databases using a common ontology-based semantic model would facilitate deeper integration and more effective utilization of these diverse and rapidly growing resources. Despite ongoing work moving toward shared data formats and linked identifiers, significant problems persist in semantic data integration in order to establish shared identity and shared meaning across heterogeneous biomedical data sources. Results: We present five processes for semantic data integration that, when applied collectively, solve seven key problems. These processes include making explicit the differences between biomedical concepts and database records, aggregating sets of identifiers denoting the same biomedical concepts across data sources, and using declaratively represented forward-chaining rules to take information that is variably represented in source databases and integrating it into a consistent biomedical representation. We demonstrate these processes and solutions by presenting KaBOB (the Knowledge Base Of Biomedicine), a knowledge base of semantically integrated data from 18 prominent biomedical databases using common representations grounded in Open Biomedical Ontologies. An instance of KaBOB with data about humans and seven major model organisms can be built using on the order of 500 million RDF triples. All source code for building KaBOB is available under an open-source license. Conclusions: KaBOB is an integrated knowledge base of biomedical data representationally based in prominent, actively maintained Open Biomedical Ontologies, thus enabling queries of the underlying data in terms of biomedical concepts (e. g., genes and gene products, interactions and processes) rather than features of source specific data schemas or file formats. KaBOB resolves many of the issues that routinely plague biomedical researchers intending to work with data from multiple data sources and provides a platform for ongoing data integration and development and for formal reasoning over a wealth of integrated biomedical data.
引用
收藏
页数:21
相关论文
共 50 条
  • [1] KaBOB: ontology-based semantic integration of biomedical databases
    Kevin M Livingston
    Michael Bada
    William A Baumgartner
    Lawrence E Hunter
    [J]. BMC Bioinformatics, 16
  • [2] DartWiki: A Semantic Wiki for Ontology-Based Knowledge Integration in the Biomedical Domain
    Yu, Tong
    Chen, Huajun
    Mi, Jinhua
    Gu, Peiqin
    Wu, Ting
    Pan, Jeff Z.
    [J]. CURRENT BIOINFORMATICS, 2012, 7 (03) : 278 - 288
  • [3] Institutionalising ontology-based semantic integration
    Schorlemmer, Marco
    Kalfoglou, Yannis
    [J]. APPLIED ONTOLOGY, 2008, 3 (03) : 131 - 150
  • [4] Using XML technology for the ontology-based semantic integration of life science Databases
    Philippi, S
    Köhler, J
    [J]. IEEE TRANSACTIONS ON INFORMATION TECHNOLOGY IN BIOMEDICINE, 2004, 8 (02): : 154 - 160
  • [5] An Ontology-based Approach for Semantic Integration
    Calhau, Rodrigo Fernandes
    Falbo, Ricardo de Almeida
    [J]. 2010 14TH IEEE INTERNATIONAL ENTERPRISE DISTRIBUTED OBJECT COMPUTING CONFERENCE (EDOC 2010), 2010, : 111 - 120
  • [6] An ontology-based semantic integration for digital museums
    Bao, H
    Liu, HZ
    Yu, JH
    Xu, HW
    [J]. ADVANCES IN WEB-AGE INFORMATION MANAGEMENT, PROCEEDINGS, 2005, 3739 : 626 - 631
  • [7] An ontology-based framework for XML semantic integration
    Cruz, IF
    Xiao, HY
    Hsu, FH
    [J]. INTERNATIONAL DATABASE ENGINEERING AND APPLICATIONS SYMPOSIUM, PROCEEDINGS, 2004, : 217 - 226
  • [8] Semantic integration: A survey of ontology-based approaches
    Noy, NF
    [J]. SIGMOD RECORD, 2004, 33 (04) : 65 - 70
  • [9] ONTOFUSION:: Ontology-based integration of genomic and clinical databases
    Perez-Rey, D.
    Maojo, V.
    Garcia-Remesal, M.
    Alonso-Calvo, R.
    Billhardt, H.
    Martin-Sanchez, F.
    Sousa, A.
    [J]. COMPUTERS IN BIOLOGY AND MEDICINE, 2006, 36 (7-8) : 712 - 730
  • [10] ONTOGRATE: TOWARDS AUTOMATIC INTEGRATION FOR RELATIONAL DATABASES AND THE SEMANTIC WEB THROUGH AN ONTOLOGY-BASED FRAMEWORK
    Dou, Dejing
    Qin, Han
    Lependu, Paea
    [J]. INTERNATIONAL JOURNAL OF SEMANTIC COMPUTING, 2010, 4 (01) : 123 - 151