Refining dataset curation methods for deep learning-based automated tuberculosis screening.

Kim, Tae Kyung; Yi, Paul H; Hager, Gregory D; Lin, Cheng Ting

Kim, Tae Kyung; Yi, Paul H; Hager, Gregory D; Lin, Cheng Ting.

Afiliación

Kim TK; Russell H. Morgan Department of Radiology and Radiological Science, Johns Hopkins Hospital, Baltimore, MD, USA.
Yi PH; Radiology Artificial Intelligence Lab (RAIL), Johns Hopkins Malone Center for Engineering in Healthcare, Baltimore, MD, USA.
Hager GD; Russell H. Morgan Department of Radiology and Radiological Science, Johns Hopkins Hospital, Baltimore, MD, USA.
Lin CT; Radiology Artificial Intelligence Lab (RAIL), Johns Hopkins Malone Center for Engineering in Healthcare, Baltimore, MD, USA.

J Thorac Dis ; 12(9): 5078-5085, 2020 Sep.

Article en En | MEDLINE | ID: mdl-33145084

RESUMEN

BACKGROUND: The study objective was to determine whether unlabeled datasets can be used to further train and improve the accuracy of a deep learning system (DLS) for the detection of tuberculosis (TB) on chest radiographs (CXRs) using a two-stage semi-supervised approach. METHODS: A total of 111,622 CXRs from the National Institute of Health ChestX-ray14 database were collected. A cardiothoracic radiologist reviewed a subset of 11,000 CXRs and dichotomously labeled each for the presence or absence of potential TB findings; these interpretations were used to train a deep convolutional neural network (DCNN) to identify CXRs with possible TB (Phase I). The best performing algorithm was then used to label the remaining database consisting of 100,622 radiographs; subsequently, these newly-labeled images were used to train a second DCNN (phase II). The best-performing algorithm from phase II (TBNet) was then tested against CXRs obtained from 3 separate sites (2 from the USA, 1 from China) with clinically confirmed cases of TB. Receiver operating characteristic (ROC) curves were generated with area under the curve (AUC) calculated. RESULTS: The phase I algorithm trained using 11,000 expert-labelled radiographs achieved an AUC of 0.88. The phase II algorithm trained on images labeled by the phase I algorithm achieved an AUC of 0.91 testing against a TB dataset obtained from Shenzhen, China and Montgomery County, USA. The algorithm generalized well to radiographs obtained from a tertiary care hospital, achieving an AUC of 0.87; TBNet's sensitivity, specificity, positive predictive value, and negative predictive value were 85%, 76%, 0.64, and 0.9, respectively. When TBNet was used to arbitrate discrepancies between 2 radiologists, the overall sensitivity reached 94% and negative predictive value reached 0.96, demonstrating a synergistic effect between the algorithm's output and radiologists' interpretations. CONCLUSIONS: Using semi-supervised learning, we trained a deep learning algorithm that detected TB at a high accuracy and demonstrated value as a CAD tool by identifying relevant CXR findings, especially in cases that were misinterpreted by radiologists. When dataset labels are noisy or absent, the described methods can significantly reduce the required amount of curated data to build clinically-relevant deep learning models, which will play an important role in the era of precision medicine.

Palabras clave

Artificial intelligence (AI); chest radiography (CXR); deep learning system (DLS); tuberculosis (TB)

Texto completo

Añadir a Mi BVS

Imprimir

XML

PubMed Links

Buscar en Google

Texto completo: 1 Colección: 01-internacional Base de datos: MEDLINE Contexto en salud: 2_ODS3 / 3_ND Problema de salud: 2_enfermedades_transmissibles / 3_neglected_diseases / 3_tuberculosis Tipo de estudio: Diagnostic_studies / Prognostic_studies / Screening_studies Idioma: En Revista: J Thorac Dis Año: 2020 Tipo del documento: Article País de afiliación: Estados Unidos

Texto completo

Añadir a Mi BVS

Imprimir

XML

PubMed Links

Buscar en Google