Papers › Interpreting chest X-rays via CNNs that exploit hierarchical disease dependencies and...

Interpreting chest X-rays via CNNs that exploit hierarchical disease dependencies and uncertainty labels

15 Nov 2019arXiv:1911.06475archive 2025-07-28

Hieu H. Pham, Tung T. Le, Dat Q. Tran, Dat T. Ngo, Ha Q. Nguyen

Chest radiography is one of the most common types of diagnostic radiology exams, which is critical for screening and diagnosis of many different thoracic diseases. Specialized algorithms have been developed to detect several specific pathologies such as lung nodule or lung cancer. However, accurately detecting the presence of multiple diseases from chest X-rays (CXRs) is still a challenging task. This paper presents a supervised multi-label classification framework based on deep convolutional neural networks (CNNs) for predicting the risk of 14 common thoracic diseases. We tackle this problem by training state-of-the-art CNNs that exploit dependencies among abnormality labels. We also propose to use the label smoothing technique for a better handling of uncertain samples, which occupy a significant portion of almost every CXR dataset. Our model is trained on over 200,000 CXRs of the recently released CheXpert dataset and achieves a mean area under the curve (AUC) of 0.940 in predicting 5 selected pathologies from the validation set. This is the highest AUC score yet reported to date. The proposed method is also evaluated on the independent test set of the CheXpert competition, which is composed of 500 CXR studies annotated by a panel of 5 experienced radiologists. The performance is on average better than 2.6 out of 3 other individual radiologists with a mean AUC of 0.930, which ranks first on the CheXpert leaderboard at the time of writing this paper.

PaperPDFCode

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

thanhtran98/chexpert_pytorch_base mentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

DiagnosticMUlTI-LABEL-ClASSIFICATIONMulti-Label Classification

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Multi-Label Classification CheXpert Hierarchical-Learning-V1 (ensemble) AVERAGE AUC ON 14 LABEL 0.930 #3 of 226 Archive leaderboard report
Multi-Label Classification CheXpert Hierarchical-Learning-V1 (ensemble) NUM RADS BELOW CURVE 2.600 #3 of 226 Archive leaderboard report
Multi-Label Classification CheXpert Hierarchical-Learning-V4 (ensemble) AVERAGE AUC ON 14 LABEL 0.929 #6 of 226 Archive leaderboard report
Multi-Label Classification CheXpert Hierarchical-Learning-V4 (ensemble) NUM RADS BELOW CURVE 2.600 #6 of 226 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

Label SmoothingTest

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections