Ten quick tips for avoiding pitfalls in multi-omics data integration analyses

被引:5
|
作者
Chicco, Davide [1 ]
Cumbo, Fabio [2 ]
Angione, Claudio [3 ]
机构
[1] Univ Toronto, Inst Hlth Policy Management & Evaluat, Toronto, ON, Canada
[2] Cleveland Clin, Genom Med Inst, Lerner Res Inst, Cleveland, OH USA
[3] Teesside Univ, Sch Comp Engn & Digital Technol, Middlesbrough, England
关键词
STANDARDS;
D O I
10.1371/journal.pcbi.1011224
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Data are the most important elements of bioinformatics: Computational analysis of bioinformatics data, in fact, can help researchers infer new knowledge about biology, chemistry, biophysics, and sometimes even medicine, influencing treatments and therapies for patients. Bioinformatics and high-throughput biological data coming from different sources can even be more helpful, because each of these different data chunks can provide alternative, complementary information about a specific biological phenomenon, similar to multiple photos of the same subject taken from different angles. In this context, the integration of bioinformatics and high-throughput biological data gets a pivotal role in running a successful bioinformatics study. In the last decades, data originating from proteomics, metabolomics, metagenomics, phenomics, transcriptomics, and epigenomics have been labelled -omics data, as a unique name to refer to them, and the integration of these omics data has gained importance in all biological areas. Even if this omics data integration is useful and relevant, due to its heterogeneity, it is not uncommon to make mistakes during the integration phases. We therefore decided to present these ten quick tips to perform an omics data integration correctly, avoiding common mistakes we experienced or noticed in published studies in the past. Even if we designed our ten guidelines for beginners, by using a simple language that (we hope) can be understood by anyone, we believe our ten recommendations should be taken into account by all the bioinformaticians performing omics data integration, including experts.
引用
收藏
页数:15
相关论文
共 50 条
  • [1] Survey on Multi-omics, and Multi-omics Data Analysis, Integration and Application
    Shahrajabian, Mohamad Hesam
    Sun, Wenli
    CURRENT PHARMACEUTICAL ANALYSIS, 2023, 19 (04) : 267 - 281
  • [2] Towards multi-omics synthetic data integration
    Selvarajoo, Kumar
    Maurer-Stroh, Sebastian
    BRIEFINGS IN BIOINFORMATICS, 2024, 25 (03)
  • [3] A cloud solution for multi-omics data integration
    Tordini, Fabio
    2016 INT IEEE CONFERENCES ON UBIQUITOUS INTELLIGENCE & COMPUTING, ADVANCED & TRUSTED COMPUTING, SCALABLE COMPUTING AND COMMUNICATIONS, CLOUD AND BIG DATA COMPUTING, INTERNET OF PEOPLE, AND SMART WORLD CONGRESS (UIC/ATC/SCALCOM/CBDCOM/IOP/SMARTWORLD), 2016, : 559 - 566
  • [4] Multi-Omics Factor Analysis-a framework for unsupervised integration of multi-omics data sets
    Argelaguet, Ricard
    Velten, Britta
    Arnol, Damien
    Dietrich, Sascha
    Zenz, Thorsten
    Marioni, John C.
    Buettner, Florian
    Huber, Wolfgang
    Stegle, Oliver
    MOLECULAR SYSTEMS BIOLOGY, 2018, 14 (06)
  • [5] Multi-omics Data Integration, Interpretation, and Its Application
    Subramanian, Indhupriya
    Verma, Srikant
    Kumar, Shiva
    Jere, Abhay
    Anamika, Krishanpal
    BIOINFORMATICS AND BIOLOGY INSIGHTS, 2020, 14
  • [6] Optimizing network propagation for multi-omics data integration
    Charmpi, Konstantina
    Chokkalingam, Manopriya
    Johnen, Ronja
    Beyer, Andreas
    PLOS COMPUTATIONAL BIOLOGY, 2021, 17 (11)
  • [7] ‘Multi-omics’ data integration: applications in probiotics studies
    Iliya Dauda Kwoji
    Olayinka Ayobami Aiyegoro
    Moses Okpeku
    Matthew Adekunle Adeleke
    npj Science of Food, 7
  • [8] Methods for the integration of multi-omics data: mathematical aspects
    Bersanelli, Matteo
    Mosca, Ettore
    Remondini, Daniel
    Giampieri, Enrico
    Sala, Claudia
    Castellani, Gastone
    Milanesi, Luciano
    BMC BIOINFORMATICS, 2016, 17
  • [9] Multi-omics data integration by generative adversarial network
    Ahmed, Khandakar Tanvir
    Sun, Jiao
    Cheng, Sze
    Yong, Jeongsik
    Zhang, Wei
    BIOINFORMATICS, 2022, 38 (01) : 179 - 186
  • [10] A survey on data integration for multi-omics sample clustering
    Lovino, Marta
    Randazzo, Vincenzo
    Ciravegna, Gabriele
    Barbiero, Pietro
    Ficarra, Elisa
    Cirrincione, Giansalvo
    NEUROCOMPUTING, 2022, 488 : 494 - 508