Learning to diagnose common thorax diseases on chest radiographs from radiology reports in Vietnamese

Nguyen, Thao; Vo, Tam M.; Nguyen, Thang V; Pham, Hieu H.; Nguyen, Ha Q.

dc.contributor.author	Nguyen, Thao
dc.contributor.author	Vo, Tam M.
dc.contributor.author	Nguyen, Thang V
dc.contributor.author	Pham, Hieu H.
dc.contributor.author	Nguyen, Ha Q.
dc.date.accessioned	2024-08-16T04:08:44Z
dc.date.available	2024-08-16T04:08:44Z
dc.date.issued	2022-10-07
dc.identifier.uri	https://vinspace.edu.vn/handle/VIN/154
dc.description.abstract	Deep learning, in recent times, has made remarkable strides when it comes to impressive performance for many tasks, including medical image processing. One of the contributing factors to these advancements is the emergence of large medical image datasets. However, it is exceedingly expensive and time-consuming to construct a large and trustworthy medical dataset; hence, there has been multiple research leveraging medical reports to automatically extract labels for data. The majority of this labor, however, is performed in English. In this work, we propose a data collecting and annotation pipeline that extracts information from Vietnamese radiology reports to provide accurate labels for chest X-ray (CXR) images. This can benefit Vietnamese radiologists and clinicians by annotating data that closely match their endemic diagnosis categories which may vary from country to country. To assess the efficacy of the proposed labeling technique, we built a CXR dataset containing 9,752 studies and evaluated our pipeline using a subset of this dataset. With an F1-score of at least 0.9923, the evaluation demonstrates that our labeling tool performs precisely and consistently across all classes. After building the dataset, we train deep learning models that leverage knowledge transferred from large public CXR datasets. We employ a variety of loss functions to overcome the curse of imbalanced multi-label datasets and conduct experiments with various model architectures to select the one that delivers the best performance. Our best model (CheXpert-pretrained EfficientNet-B2) yields an F1-score of 0.6989 (95% CI 0.6740, 0.7240), AUC of 0.7912, sensitivity of 0.7064, and specificity of 0.8760 for the abnormal diagnosis in general. Finally, we demonstrate that our coarse classification (based on five specific locations of abnormalities) yields comparable results to fine classification (twelve pathologies) on the benchmark CheXpert dataset for general anomaly detection while delivering better performance in terms of the average performance of all classes.	en_US
dc.language.iso	en	en_US
dc.title	Learning to diagnose common thorax diseases on chest radiographs from radiology reports in Vietnamese	en_US
dc.type	Article	en_US

Files in this item

Name:: Learning to diagnose common ...
Size:: 1.266Mb
Format:: PDF

View/Open

This item appears in the following Collection(s)

Pham Huy Hieu, PhD. [36]
College of Engineering and Computer Science Associate Director, VinUni-Illinois Smart Health Center Assistant Professor, Computer Science program

Show simple item record