DeepPL: A deep-learning-based tool for the prediction of bacteriophage lifecycle

被引:0
|
作者
Zhang, Yujie [1 ]
Mao, Mark [2 ]
Zhang, Robert [2 ]
Liao, Yen-Te [1 ]
Wu, Vivian C. H. [1 ]
机构
[1] US Dept Agr, Agr Res Serv, Western Reg Res Ctr, Produce Safety & Microbiol Res Unit, Albany, CA 94710 USA
[2] Clowit LLC, Burlingame, CA USA
关键词
46;
D O I
10.1371/journal.pcbi.1012525
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Bacteriophages (phages) are viruses that infect bacteria and can be classified into two different lifecycles. Virulent phages (or lytic phages) have a lytic cycle that can lyse the bacteria host after their infection. Temperate phages (or lysogenic phages) can integrate their phage genomes into bacterial chromosomes and replicate with bacterial hosts via the lysogenic cycle. Identifying phage lifecycles is a crucial step in developing suitable applications for phages. Compared to the complicated traditional biological experiments, several tools have been designed for predicting phage lifecycle using different algorithms, such as random forest (RF), linear support-vector classifier (SVC), and convolutional neural network (CNN). In this study, we developed a natural language processing (NLP)-based tool-DeepPL-for predicting phage lifecycles via nucleotide sequences. The test results showed that our DeepPL had an accuracy of 94.65% with a sensitivity of 92.24% and a specificity of 95.91%. Moreover, DeepPL had 100% accuracy in lifecycle prediction on the phages we isolated and biologically verified previously in the lab. Additionally, a mock phage community metagenomic dataset was used to test the potential usage of DeepPL in viral metagenomic research. DeepPL displayed a 100% accuracy for individual phage complete genomes and high accuracies ranging from 71.14% to 100% on phage contigs produced by various next-generation sequencing technologies. Overall, our study indicates that DeepPL has a reliable performance on phage lifecycle prediction using the most fundamental nucleotide sequences and can be applied to future phage and metagenomic research. Bacteriophages are viruses that infect bacteria and play a critical role in the microbial community within different environments via phage-bacterial evolutionary interactions. The classification of phage lifecycles is of great importance in deploying the potential applications of phages and better understanding complex microbial interactions. However, the traditional biological methods for phage lifecycle identification are complicated and time-consuming. In this study, we proposed a deep learning-based tool-DeepPL-for predicting phage lifecycles using the phage nucleotide genome. Compared with other bioinformatic tools, DeepPL was developed from the pre-trained transformers model-DNABERT-designed for fundamental nucleotide language combined with representative phage complete genomes. Our in-house biological results were further used to verify the output of DeepPL. Overall, DeepPL performs with high precision for phage lifecycle prediction and could contribute to the genomic data-driven direction of phage research and applications.
引用
收藏
页数:13
相关论文
共 50 条
  • [1] Deep-Learning-Based Approach for Prediction of Algal Blooms
    Zhang, Feng
    Wang, Yuanyuan
    Cao, Minjie
    Sun, Xiaoxiao
    Du, Zhenhong
    Liu, Renyi
    Ye, Xinyue
    SUSTAINABILITY, 2016, 8 (10)
  • [2] Deep-learning-based vehicle trajectory prediction: A review
    Yin, Chenhui
    Cecotti, Marco
    Auger, Daniel J.
    Fotouhi, Abbas
    Jiang, Haobin
    IET INTELLIGENT TRANSPORT SYSTEMS, 2025, 19 (01)
  • [3] A Deep-Learning-Based Visualization Tool for Air Pollution Forecasting
    Nguyen, Huynh A. D.
    Le, Hoang T.
    Barthelemy, Xavier
    Azzi, Merched
    Duc, Hiep
    Jiang, Ningbo
    Riley, Matthew
    Ha, Quang P.
    IEEE SOFTWARE, 2025, 42 (02) : 47 - 56
  • [4] A Study of Deep-Learning-based Prediction Methods for Lossless Coding
    Schiopu, Ionut
    Huang, Hongyue
    Munteanu, Adrian
    28TH EUROPEAN SIGNAL PROCESSING CONFERENCE (EUSIPCO 2020), 2021, : 521 - 525
  • [5] DeepSipred: A deep-learning-based approach on siRNA inhibition prediction
    Liu, Bin
    Huang, Huiya
    Liao, Weixi
    Pan, Xiaoyong
    Jin, Cheng
    Yuan, Ye
    PROCEEDINGS OF 2024 4TH INTERNATIONAL CONFERENCE ON BIOINFORMATICS AND INTELLIGENT COMPUTING, BIC 2024, 2024, : 430 - 436
  • [6] Deep-Learning-Based Drug-Target Interaction Prediction
    Wen, Ming
    Zhang, Zhimin
    Niu, Shaoyu
    Sha, Haozhi
    Yang, Ruihan
    Yun, Yonghuan
    Lu, Hongmei
    JOURNAL OF PROTEOME RESEARCH, 2017, 16 (04) : 1401 - 1409
  • [7] Deep-Learning-Based Prediction of the Tetragonal → Cubic Transition in Davemaoite
    Wu, Fulun
    Sun, Yang
    Wan, Tianqi
    Wu, Shunqing
    Wentzcovitch, Renata M.
    GEOPHYSICAL RESEARCH LETTERS, 2024, 51 (12)
  • [8] The recent progress of deep-learning-based in silico prediction of drug combination
    Liu, Haoyang
    Fan, Zhiguang
    Lin, Jie
    Yang, Yuedong
    Ran, Ting
    Chen, Hongming
    DRUG DISCOVERY TODAY, 2023, 28 (07)
  • [9] Deep-learning-based model for prediction of crowding in a public transit system
    Shrivastava, Arpit
    Rawat, Nishtha
    Agarwal, Amit
    PUBLIC TRANSPORT, 2024, 16 (02) : 449 - 484
  • [10] Deep-learning-based survival prediction of patients with lower limb melanoma
    Zhang, Jinrong
    Yu, Hai
    Zheng, Xinkai
    Ming, Wai-kit
    Lak, Yau Sun
    Tom, Kong Ching
    Lee, Alice
    Huang, Hui
    Chen, Wenhui
    Lyu, Jun
    Deng, Liehua
    DISCOVER ONCOLOGY, 2023, 14 (01)