Scalable de novo classification of antibiotic resistance of Mycobacterium tuberculosis

被引:0
|
作者
Serajian, Mohammadali [1 ]
Marini, Simone [2 ]
Alanko, Jarno N. [3 ]
Noyes, Noelle R. [4 ]
Prosperi, Mattia [2 ]
Boucher, Christina [1 ]
机构
[1] Univ Florida, Dept Comp & Informat Sci & Engn, 1889 Museum Rd, Gainesville, FL 32611 USA
[2] Univ Florida, Dept Epidemiol, POB 100231, Gainesville, FL 32601 USA
[3] Univ Helsinki, Dept Comp Sci, POB 4, Helsinki 00014, Finland
[4] Univ Minnesota, Dept Vet Populat Med, 1365 Gortner Ave, St Paul, MN 55108 USA
关键词
READ ALIGNMENT; GENOME; TOOL;
D O I
10.1093/bioinformatics/btae243
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Motivation: World Health Organization estimates that there were over 10 million cases of tuberculosis (TB) worldwide in 2019, resulting in over 1.4 million deaths, with a worrisome increasing trend yearly. The disease is caused by Mycobacterium tuberculosis (MTB) through airborne transmission. Treatment of TB is estimated to be 85% successful, however, this drops to 57% if MTB exhibits multiple antimicrobial resistance (AMR), for which fewer treatment options are available. Results: We develop a robust machine-learning classifier using both linear and nonlinear models (i.e. LASSO logistic regression (LR) and random forests (RF)) to predict the phenotypic resistance of Mycobacterium tuberculosis (MTB) for a broad range of antibiotic drugs. We use data from the CRyPTIC consortium to train our classifier, which consists of whole genome sequencing and antibiotic susceptibility testing (AST) phenotypic data for 13 different antibiotics. To train our model, we assemble the sequence data into genomic contigs, identify all unique 31-mers in the set of contigs, and build a feature matrix M, where M[i, j] is equal to the number of times the ith 31-mer occurs in the jth genome. Due to the size of this feature matrix (over 350 million unique 31-mers), we build and use a sparse matrix representation. Our method, which we refer to as MTB++, leverages compact data structures and iterative methods to allow for the screening of all the 31-mers in the development of both LASSO LR and RF. MTB++ is able to achieve high discrimination (F-1 >80%) for the first-line antibiotics. Moreover, MTB++ had the highest F-1 score in all but three classes and was the most comprehensive since it had an F-1 score >75% in all but four (rare) antibiotic drugs. We use our feature selection to contextualize the 31-mers that are used for the prediction of phenotypic resistance, leading to some insights about sequence similarity to genes in MEGARes. Lastly, we give an estimate of the amount of data that is needed in order to provide accurate predictions.
引用
收藏
页码:i39 / i47
页数:9
相关论文
共 50 条
  • [1] Ancestral antibiotic resistance in Mycobacterium tuberculosis
    Morris, RP
    Nguyen, L
    Gatfield, J
    Visconti, K
    Nguyen, K
    Schnappinger, D
    Ehrt, S
    Liu, Y
    Heifets, L
    Pieters, J
    Schoolnik, G
    Thompson, CJ
    PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 2005, 102 (34) : 12200 - 12205
  • [2] The complex evolution of antibiotic resistance in Mycobacterium tuberculosis
    Fonseca, J. D.
    Knight, G. M.
    McHugh, H. D.
    INTERNATIONAL JOURNAL OF INFECTIOUS DISEASES, 2015, 32 : 94 - 100
  • [3] The competitive cost of antibiotic resistance in Mycobacterium tuberculosis
    Gagneux, Sebastien
    Long, Clara Davis
    Small, Peter M.
    Van, Tran
    Schoolnik, Gary K.
    Bohannan, Brendan J. M.
    SCIENCE, 2006, 312 (5782) : 1944 - 1946
  • [4] A comparative study of antibiotic resistance patterns in Mycobacterium tuberculosis
    Mohammadali Serajian
    Conrad Testagrose
    Mattia Prosperi
    Christina Boucher
    Scientific Reports, 15 (1)
  • [5] De Novo Emergence of Peptides That Confer Antibiotic Resistance
    Knopp, Michael
    Gudmundsdottir, Jonina S.
    Nilsson, Tobias
    Konig, Finja
    Warsi, Omar
    Rajer, Fredrika
    Adelroth, Pia
    Andersson, Dan I.
    MBIO, 2019, 10 (03):
  • [6] The Epistatic Landscape of Antibiotic Resistance of Different Clades of Mycobacterium tuberculosis
    Muzondiwa, Dillon
    Hlanze, Hleliwe
    Reva, Oleg N.
    ANTIBIOTICS-BASEL, 2021, 10 (07):
  • [7] Transition bias influences the evolution of antibiotic resistance in Mycobacterium tuberculosis
    Payne, Joshua L.
    Menardo, Fabrizio
    Trauner, Andrej
    Borrell, Sonia
    Gygli, Sebastian M.
    Loiseau, Chloe
    Gagneux, Sebastien
    Hall, Alex R.
    PLOS BIOLOGY, 2019, 17 (05)
  • [8] ANTIBIOTIC RESISTANCE OF MYCOBACTERIUM TUBERCULOSIS; MECHANISMS AND SPECIFIC THERAPEUTIC RESPONSE
    Popescu, Gilda Georgeta
    Arghir, Oana Cristina
    Fildan, Ariadna Petronela
    Spanu, Victor
    Cambrea, Simona Claudia
    Rafila, Alexandru
    Buicu, Florin Corneliu
    FARMACIA, 2020, 68 (02) : 197 - 205
  • [9] Phylogenetic polymorphisms in antibiotic resistance genes of the Mycobacterium tuberculosis complex
    Feuerriegel, Silke
    Koeser, Claudio U.
    Niemann, Stefan
    JOURNAL OF ANTIMICROBIAL CHEMOTHERAPY, 2014, 69 (05) : 1205 - 1210
  • [10] Overcoming aminoglycoside antibiotic resistance in Mycobacterium tuberculosis by targeting Eis protein
    Geethu S. Kumar
    Kuldeep Sharma
    Richa Mishra
    Esam Ibraheem Azhar
    Vivek Dhar Dwivedi
    Sharad Agrawal
    In Silico Pharmacology, 13 (1)