Cont-ID: detection of sample cross-contamination in viral metagenomic data

被引:2
|
作者
Rollin, Johan [1 ,2 ]
Rong, Wei [1 ]
Massart, Sebastien [1 ]
机构
[1] Univ Liege, Plant Pathol Lab, Gembloux Agrobio Tech, B-5030 Gembloux, Belgium
[2] DNAVision, B-6041 Gosselies, Belgium
关键词
Bioinformatic; Virus; Detection; Sequencing; Contamination; Metagenomic; PIPELINE; IMPACT;
D O I
10.1186/s12915-023-01708-w
中图分类号
Q [生物科学];
学科分类号
07 ; 0710 ; 09 ;
摘要
BackgroundHigh-throughput sequencing (HTS) technologies completed by the bioinformatic analysis of the generated data are becoming an important detection technique for virus diagnostics. They have the potential to replace or complement the current PCR-based methods thanks to their improved inclusivity and analytical sensitivity, as well as their overall good repeatability and reproducibility. Cross-contamination is a well-known phenomenon in molecular diagnostics and corresponds to the exchange of genetic material between samples. Cross-contamination management was a key drawback during the development of PCR-based detection and is now adequately monitored in routine diagnostics. HTS technologies are facing similar difficulties due to their very high analytical sensitivity. As a single viral read could be detected in millions of sequencing reads, it is mandatory to fix a detection threshold that will be informed by estimated cross-contamination. Cross-contamination monitoring should therefore be a priority when detecting viruses by HTS technologies.ResultsWe present Cont-ID, a bioinformatic tool designed to check for cross-contamination by analysing the relative abundance of virus sequencing reads identified in sequence metagenomic datasets and their duplication between samples. It can be applied when the samples in a sequencing batch have been processed in parallel in the laboratory and with at least one specific external control called Alien control. Using 273 real datasets, including 68 virus species from different hosts (fruit tree, plant, human) and several library preparation protocols (Ribodepleted total RNA, small RNA and double-stranded RNA), we demonstrated that Cont-ID classifies with high accuracy (91%) viral species detection into (true) infection or (cross) contamination. This classification raises confidence in the detection and facilitates the downstream interpretation and confirmation of the results by prioritising the virus detections that should be confirmed.ConclusionsCross-contamination between samples when detecting viruses using HTS (Illumina technology) can be monitored and highlighted by Cont-ID (provided an alien control is present). Cont-ID is based on a flexible methodology relying on the output of bioinformatics analyses of the sequencing reads and considering the contamination pattern specific to each batch of samples. The Cont-ID method is adaptable so that each laboratory can optimise it before its validation and routine use.
引用
收藏
页数:20
相关论文
共 50 条
  • [21] Detection of tuberculosis laboratory cross-contamination using whole-genome sequencing
    Wu, Jie
    Yang, Chongguang
    Lu, Liping
    Dai, Wanqin
    TUBERCULOSIS, 2019, 115 : 121 - 125
  • [22] Cross-contamination in the Molecular Detection of Bartonella from Paraffin-embedded Tissues
    Varanat, M.
    Maggi, R. G.
    Linder, K. E.
    Horton, S.
    Breitschwerdt, E. B.
    VETERINARY PATHOLOGY, 2009, 46 (05) : 940 - 944
  • [23] INVESTIGATION OF CROSS-CONTAMINATION IN A MYCOBACTERIUM TUBERCULOSIS LABORATORY USING EPIDEMIOLOGICAL DATA AND SPOLIGOTYPING
    Al-Otaibi, Fawzia
    El-Hazmi, Malak M.
    PAKISTAN JOURNAL OF MEDICAL SCIENCES, 2008, 24 (06) : 827 - 832
  • [24] Rapid detection of laboratory cross-contamination with Mycobacterium tuberculosis using multispacer sequence typing
    Zoheira Djelouadji
    Jean Orehek
    Michel Drancourt
    BMC Microbiology, 9
  • [25] Patterns of cross-contamination in a multispecies population genomic project: detection, quantification, impact, and solutions
    Marion Ballenghien
    Nicolas Faivre
    Nicolas Galtier
    BMC Biology, 15
  • [26] Rapid detection of laboratory cross-contamination with Mycobacterium tuberculosis using multispacer sequence typing
    Djelouadji, Zoheira
    Orehek, Jean
    Drancourt, Michel
    BMC MICROBIOLOGY, 2009, 9
  • [27] Patterns of cross-contamination in a multispecies population genomic project: detection, quantification, impact, and solutions
    Ballenghien, Marion
    Faivre, Nicolas
    Galtier, Nicolas
    BMC BIOLOGY, 2017, 15
  • [28] Rapid screening and detection of cellular cross-contamination in cell cultures: A new application of cytochemistry
    Markovic, O
    Hay, RJ
    Steenbergen, K
    JOURNAL OF HISTOCHEMISTRY & CYTOCHEMISTRY, 1996, 44 (01) : 75 - 76
  • [29] Potassium ethylenediaminetetraacetic acid (kEDTA) sample cross-contamination: prevalence, consequences, identification, mechanisms, prevention and mitigation
    Lorde, Nathan
    Mahapatra, Shivani
    Kalaria, Tejas
    Gama, Rousseau
    JOURNAL OF LABORATORY AND PRECISION MEDICINE, 2024, 9
  • [30] ContEst: estimating cross-contamination of human samples in next-generation sequencing data
    Cibulskis, Kristian
    McKenna, Aaron
    Fennell, Tim
    Banks, Eric
    DePristo, Mark
    Getz, Gad
    BIOINFORMATICS, 2011, 27 (18) : 2601 - 2602