Yang, Y;
Niehaus, KE;
Walker, TM;
Iqbal, Z;
Walker, AS;
Wilson, DJ;
Peto, TEA;
... Clifton, DA; + view all
(2018)
Machine learning for classifying tuberculosis drug-resistance from DNA sequencing data.
Bioinformatics
, 34
(10)
pp. 1666-1671.
10.1093/bioinformatics/btx801.
Preview |
Text
Walker_btx801.pdf - Published Version Download (392kB) | Preview |
Abstract
Motivation: Correct and rapid determination of Mycobacterium tuberculosis (MTB) resistance against available tuberculosis (TB) drugs is essential for the control and management of TB. Conventional molecular diagnostic test assumes that the presence of any well-studied single nucleotide polymorphisms is sufficient to cause resistance, which yields low sensitivity for resistance classification. Methods: Given the availability of DNA sequencing data from MTB, we developed machine learning models for a cohort of 1839 UK bacterial isolates to classify MTB resistance against eight anti-TB drugs (isoniazid, rifampicin, ethambutol, pyrazinamide, ciprofloxacin, moxifloxacin, ofloxacin, streptomycin) and to classify multi-drug resistance. Results: Compared to previous rules-based approach, the sensitivities from the best-performing models increased by 2-4% for isoniazid, rifampicin and ethambutol to 97% (p<0.01), respectively; for ciprofloxacin and multi-drug resistant TB, they increased to 96%. For moxifloxacin and ofloxacin, sensitivities increased by 12% and 15% from 83% and 81% based on existing known resistance alleles to 95% and 96% (p<0.01), respectively. Particularly, our models improved sensitivities compared to the previous rules-based approach by 15% and 24% to 84% and 87% for pyrazinamide and streptomycin (p<0.01), respectively. The best-performing models increase the area-under-the-ROC curve by 10% for pyrazinamide and streptomycin (p<0.01), and 4-8% for other drugs (p<0.01). Availability: The details of source code are provided at http://www.robots.ox.ac.uk/davidc/code.php
Type: | Article |
---|---|
Title: | Machine learning for classifying tuberculosis drug-resistance from DNA sequencing data |
Open access status: | An open access version is available from UCL Discovery |
DOI: | 10.1093/bioinformatics/btx801 |
Publisher version: | https://doi.org/10.1093/bioinformatics/btx801 |
Language: | English |
Additional information: | Copyright © The Author(s) 2017. Published by Oxford University Press. This is an Open Access article distributed under the terms of the Creative Commons Attribution Non-Commercial License (http://creativecommons.org/licenses/by-nc/4.0/), which permits non-commercial re-use, distribution, and reproduction in any medium, provided the original work is properly cited. For commercial re-use, please contact journals.permissions@oup.com. |
UCL classification: | UCL UCL > Provost and Vice Provost Offices > School of Life and Medical Sciences UCL > Provost and Vice Provost Offices > School of Life and Medical Sciences > Faculty of Population Health Sciences > Inst of Clinical Trials and Methodology UCL > Provost and Vice Provost Offices > School of Life and Medical Sciences > Faculty of Population Health Sciences > Inst of Clinical Trials and Methodology > MRC Clinical Trials Unit at UCL |
URI: | https://discovery-pp.ucl.ac.uk/id/eprint/10039954 |
Archive Staff Only
![]() |
View Item |