Decoding the protein–ligand interactions using parallel graph neural networks

被引:0
|
作者
Carter Knutson
Mridula Bontha
Jenna A. Bilbrey
Neeraj Kumar
机构
[1] Pacific Northwest National Laboratory,
来源
关键词
D O I
暂无
中图分类号
学科分类号
摘要
Protein–ligand interactions (PLIs) are essential for biochemical functionality and their identification is crucial for estimating biophysical properties for rational therapeutic design. Currently, experimental characterization of these properties is the most accurate method, however, this is very time-consuming and labor-intensive. A number of computational methods have been developed in this context but most of the existing PLI prediction heavily depends on 2D protein sequence data. Here, we present a novel parallel graph neural network (GNN) to integrate knowledge representation and reasoning for PLI prediction to perform deep learning guided by expert knowledge and informed by 3D structural data. We develop two distinct GNN architectures: GNNF\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {GNN}_{\mathrm{F}}$$\end{document} is the base implementation that employs distinct featurization to enhance domain-awareness, while GNNP\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {GNN}_{\mathrm{P}}$$\end{document} is a novel implementation that can predict with no prior knowledge of the intermolecular interactions. The comprehensive evaluation demonstrated that GNN can successfully capture the binary interactions between ligand and protein’s 3D structure with 0.979 test accuracy for GNNF\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {GNN}_{\mathrm{F}}$$\end{document} and 0.958 for GNNP\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {GNN}_{\mathrm{P}}$$\end{document} for predicting activity of a protein–ligand complex. These models are further adapted for regression tasks to predict experimental binding affinities and pIC50\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {pIC}_{\mathrm{50}}$$\end{document} crucial for compound’s potency and efficacy. We achieve a Pearson correlation coefficient of 0.66 and 0.65 on experimental affinity and 0.50 and 0.51 on pIC50\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {pIC}_{\mathrm{50}}$$\end{document} with GNNF\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {GNN}_{\mathrm{F}}$$\end{document} and GNNP\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {GNN}_{\mathrm{P}}$$\end{document}, respectively, outperforming similar 2D sequence based models. Our method can serve as an interpretable and explainable artificial intelligence (AI) tool for predicted activity, potency, and biophysical properties of lead candidates. To this end, we show the utility of GNNP\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\hbox {GNN}_{\mathrm{P}}$$\end{document} on SARS-Cov-2 protein targets by screening a large compound library and comparing the prediction with the experimentally measured data.
引用
下载
收藏
相关论文
共 50 条
  • [31] Fast and effective protein model refinement using deep graph neural networks
    Xiaoyang Jing
    Jinbo Xu
    Nature Computational Science, 2021, 1 : 462 - 469
  • [32] PANDA2: protein function prediction using graph neural networks
    Zhao, Chenguang
    Liu, Tong
    Wang, Zheng
    NAR GENOMICS AND BIOINFORMATICS, 2022, 4 (01)
  • [33] GNNfam: Utilizing Sparsity in Protein Family Predictions using Graph Neural Networks
    Godase, Anuj
    Rahman, Md Khaledur
    Azad, Ariful
    12TH ACM CONFERENCE ON BIOINFORMATICS, COMPUTATIONAL BIOLOGY, AND HEALTH INFORMATICS (ACM-BCB 2021), 2021,
  • [34] Fast and effective protein model refinement using deep graph neural networks
    Jing, Xiaoyang
    Xu, Jinbo
    NATURE COMPUTATIONAL SCIENCE, 2021, 1 (07): : 462 - +
  • [35] Structure-aware Interactive Graph Neural Networks for the Prediction of Protein-Ligand Binding Affinity
    Li, Shuangli
    Zhou, Jingbo
    Xu, Tong
    Huang, Liang
    Wang, Fan
    Xiong, Haoyi
    Huang, Weili
    Dou, Dejing
    Xiong, Hui
    KDD '21: PROCEEDINGS OF THE 27TH ACM SIGKDD CONFERENCE ON KNOWLEDGE DISCOVERY & DATA MINING, 2021, : 975 - 985
  • [36] Predicting cancer drug response using parallel heterogeneous graph convolutional networks with neighborhood interactions
    Peng, Wei
    Liu, Hancheng
    Dai, Wei
    Yu, Ning
    Wang, Jianxin
    BIOINFORMATICS, 2022, 38 (19) : 4546 - 4553
  • [37] Decoding Gestures in Electromyography: Spatiotemporal Graph Neural Networks for Generalizable and Interpretable Classification
    University of Minnesota, College of Engineering and Science, Department of Computer Science, Minneapolis
    MN
    55455, United States
    不详
    MN
    55455, United States
    IEEE Trans. Neural Syst. Rehabil. Eng., 1600, 404-419 (2025):
  • [38] Predicting residue-specific qualities of individual protein models using residual neural networks and graph neural networks
    Zhao, Chenguang
    Liu, Tong
    Wang, Zheng
    PROTEINS-STRUCTURE FUNCTION AND BIOINFORMATICS, 2022, 90 (12) : 2091 - 2102
  • [39] Using Deep Neural Networks to Improve the Performance of Protein-Protein Interactions Prediction
    Gui, Yuan-Miao
    Wang, Ru-Jing
    Wang, Xue
    Wei, Yuan-Yuan
    INTERNATIONAL JOURNAL OF PATTERN RECOGNITION AND ARTIFICIAL INTELLIGENCE, 2020, 34 (13)
  • [40] Analysis of Protein-Ligand Interactions of SARS-CoV-2 Against Selective Drug Using Deep Neural Networks
    Natarajan Yuvaraj
    Kannan Srihari
    Selvaraj Chandragandhi
    Rajan Arshath Raja
    Gaurav Dhiman
    Amandeep Kaur
    Big Data Mining and Analytics, 2021, 4 (02) : 76 - 83