Research Article

Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets

Volume: 7 Number: 2 March 27, 2026

Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets

Abstract

Aims: To systematically compare the candidate-level classification performance of four state-of-the-art deep learning architectures—U-Net, ResNet-50, Vision Transformer (ViT-B/16), and YOLOv8—for pulmonary nodule candidate classification using patch-based classification on public benchmark CT datasets, and to evaluate the trade-off between detection accuracy and computational efficiency. Methods: This computational study utilized publicly available, de-identified datasets including LUNA16 (Lung Nodule Analysis 2016) and LIDC-IDRI (Lung Image Database Consortium). Candidate nodule locations were generated using a multi-scale Laplacian of Gaussian (LoG) blob detector applied to full CT volumes. From these candidates, 64×64×64 voxel patches were extracted and classified as true nodules or false positives. The dataset was partitioned at the patient level: 70% training, 15% validation, and 15% held-out test. Stratified 5-fold cross-validation was conducted exclusively within the training set for hyperparameter optimization. Four deep learning architectures were trained under identical protocols: U-Net (encoder decoder), ResNet-50 (residual CNN), ViT-B/16 (self-attention transformer, adapted to 3D patch input via 3D patch embedding), and YOLOv8 (real-time detector, applied slice-by-slice with 3D aggregation). Primary performance metrics included sensitivity, specificity, F1-score, mAP@0.5, and area under the ROC curve (AUC). Free-response ROC (FROC) analysis was performed following LUNA16 challenge standards, reporting sensitivity at 0.125, 0.25, 0.5, 1, 2, 4, and 8 false positives per scan (FP/scan). Statistical comparisons focused on AUC using paired DeLong’s test with Bonferroni correction for multiple comparisons. Bootstrap confidence intervals (n=2,000 resamples) were computed for sensitivity, specificity, and F1-score. Results: Across 888 CT scans (1,186 annotated nodules; LUNA16 test set: 133 scans, 178 nodules), Vision Transformer achieved the highest candidate-level patch classification performance: sensitivity 94.2% (95% CI: 91.8–96.1%), specificity 92.8% (95% CI: 90.3–94.9%), F1-score 0.935, mAP@0.5 0.947, and AUC 0.971 (95% CI: 0.958–0.982). Pairwise AUC comparisons using DeLong’s test confirmed superior discrimination for ViT-B/16 relative to the comparator architectures. FROC analysis demonstrated ViT-B/16 achieved the highest mean sensitivity at 7 operating points (CPM=0.847), outperforming ResNet-50 (CPM=0.798), YOLOv8 (CPM=0.781), and U-Net (CPM=0.762). However, ViT-B/16 required 3.2× longer inference time (8.4 vs 2.6 seconds/scan) and 3.7× more trainable parameters than ResNet-50. YOLOv8 demonstrated superior computational efficiency with the shortest inference time (1.1 seconds/scan). Conclusion: The attention-based Vision Transformer architecture achieved superior candidate-level patch classification performance for pulmonary nodule candidate evaluation; however, this advantage must be weighed against substantial computational costs. Architecture selection should be guided by deployment context, with ResNet-50 offering optimal accuracy efficiency balance for clinical deployment and YOLOv8 for real-time screening applications.

Keywords

Supporting Institution

None

Ethical Statement

This computational study utilized exclusively publicly available, de-identified medical imaging datasets (LUNA16, LIDC-IDRI) that were originally collected and distributed under institutional review board oversight by their respective consortia. No new patient data were collected for this investigation. No human subjects were involved in this research as defined by 45 CFR 46.102(f). Consequently, this study did not require institutional review board approval or informed consent. This classification aligns with federal regulations governing human subjects research and established precedent for computational algorithm development using pre-existing public datasets. All dataset usage complied with the terms of use specified by the original data providers.

Thanks

Cemil Gürses, MD

References

  1. Sung H, Ferlay J, Siegel RL, et al. Global cancer statistics 2020: GLOBOCAN estimates of incidence and mortality worldwide for 36 cancers in 185 countries. CA Cancer J Clin. 2021;71(3):209-249. doi:10. 3322/caac.21660
  2. LeCun Y, Bengio Y, Hinton G. Deep learning. Nature. 2015;521(7553):436-444. doi:10.1038/nature14539
  3. Litjens G, Kooi T, Bejnordi BE, et al. A survey on deep learning in medical image analysis. Med Image Anal. 2017;42:60-88. doi:10.1016/j.media.2017.07.005
  4. Ronneberger O, Fischer P, Brox T. U-Net: convolutional networks for biomedical image segmentation. In: Medical Image Computing and Computer-Assisted Intervention. Springer; 2015:234-241.
  5. He K, Zhang X, Ren S, Sun J. Deep residual learning for image recognition. Proc IEEE CVPR. 2016:770-778. doi:10.1109/CVPR.2016.90
  6. Dosovitskiy A, Beyer L, Kolesnikov A, et al. An image is worth 16x16 words: transformers for image recognition at scale. ICLR Conf Proc. 2021. doi:10.48550/arXiv.2010.11929
  7. Terven J, Cordova-Esparza D. A comprehensive review of YOLO: from YOLOv1 to YOLOv8 and beyond. arXiv [Preprint]. 2023. doi:10.48550/arXiv.2304.00501
  8. Matsoukas C, Haslum JF, Söderberg M, Smith K. Is it time to replace CNNs with transformers for medical images? arXiv [Preprint]. 2021. doi:10.48550/arXiv.2108.09038

Details

Primary Language

English

Subjects

Radiology and Organ Imaging

Journal Section

Research Article

Publication Date

March 27, 2026

Submission Date

January 23, 2026

Acceptance Date

March 17, 2026

Published in Issue

Year 2026 Volume: 7 Number: 2

APA
Kılıç, K. K., Çavuşoğlu Yalçın, N., Yalçın, M., & Kahvecioğlu, N. (2026). Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets. Journal of Medicine and Palliative Care, 7(2), 363-368. https://doi.org/10.47582/jompac.1870165
AMA
1.Kılıç KK, Çavuşoğlu Yalçın N, Yalçın M, Kahvecioğlu N. Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets. J Med Palliat Care / JOMPAC / jompac. 2026;7(2):363-368. doi:10.47582/jompac.1870165
Chicago
Kılıç, Koray Kaya, Nilay Çavuşoğlu Yalçın, Mustafa Yalçın, and Nevfel Kahvecioğlu. 2026. “Comparative Performance Analysis of Deep Learning Architectures for Pulmonary Nodule Candidate Classification: A Computational Study Using Public Benchmark Datasets”. Journal of Medicine and Palliative Care 7 (2): 363-68. https://doi.org/10.47582/jompac.1870165.
EndNote
Kılıç KK, Çavuşoğlu Yalçın N, Yalçın M, Kahvecioğlu N (March 1, 2026) Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets. Journal of Medicine and Palliative Care 7 2 363–368.
IEEE
[1]K. K. Kılıç, N. Çavuşoğlu Yalçın, M. Yalçın, and N. Kahvecioğlu, “Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets”, J Med Palliat Care / JOMPAC / jompac, vol. 7, no. 2, pp. 363–368, Mar. 2026, doi: 10.47582/jompac.1870165.
ISNAD
Kılıç, Koray Kaya - Çavuşoğlu Yalçın, Nilay - Yalçın, Mustafa - Kahvecioğlu, Nevfel. “Comparative Performance Analysis of Deep Learning Architectures for Pulmonary Nodule Candidate Classification: A Computational Study Using Public Benchmark Datasets”. Journal of Medicine and Palliative Care 7/2 (March 1, 2026): 363-368. https://doi.org/10.47582/jompac.1870165.
JAMA
1.Kılıç KK, Çavuşoğlu Yalçın N, Yalçın M, Kahvecioğlu N. Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets. J Med Palliat Care / JOMPAC / jompac. 2026;7:363–368.
MLA
Kılıç, Koray Kaya, et al. “Comparative Performance Analysis of Deep Learning Architectures for Pulmonary Nodule Candidate Classification: A Computational Study Using Public Benchmark Datasets”. Journal of Medicine and Palliative Care, vol. 7, no. 2, Mar. 2026, pp. 363-8, doi:10.47582/jompac.1870165.
Vancouver
1.Koray Kaya Kılıç, Nilay Çavuşoğlu Yalçın, Mustafa Yalçın, Nevfel Kahvecioğlu. Comparative performance analysis of deep learning architectures for pulmonary nodule candidate classification: a computational study using public Benchmark datasets. J Med Palliat Care / JOMPAC / jompac. 2026 Mar. 1;7(2):363-8. doi:10.47582/jompac.1870165

TR DİZİN ULAKBİM and International Indexes (1d)

Interuniversity Board (UAK) Equivalency: Article published in Ulakbim TR Index journal [10 POINTS], and Article published in other (excuding 1a, b, c) international indexed journal (1d) [5 POINTS]
 


 

download?token=eyJhdXRoX3JvbGVzIjpbXSwiZW5kcG9pbnQiOiJqb3VybmFsIiwib3JpZ2luYWxuYW1lIjoiVHJfSW5kZXhfbG9nby5wbmciLCJwYXRoIjoiN2EzMC84NTVhL2UyMWMvNjlkZjRkZmVhNTUyNTYuNzg3NjU2ODgucG5nIiwiZXhwIjoxNzc2MjQ1Nzc0LCJub25jZSI6IjU0MDZkMWE2NmE1Y2QwZTJjNGYyNDA1OTM2MTE0YWIxIn0.Tt-WScFXTj5r2jji5eDMFApNzujLMjMPl8ivXRbozSI



f9ab67f.png
asos-index.png


 


download?token=eyJhdXRoX3JvbGVzIjpbXSwiZW5kcG9pbnQiOiJqb3VybmFsIiwib3JpZ2luYWxuYW1lIjoiQ3Jvc3NyZWYuanBnIiwicGF0aCI6IjAzMzEvMTdkZi8yN2ZkLzY5ZGY0ZThhMDZkMjg0LjQxMjAyNDg5LmpwZyIsImV4cCI6MTc3NjI0NTkxNCwibm9uY2UiOiI2NjM1Yjc5MWFiY2I1MDQ0NjkzMTAxMDhjY2Y2NzRlMCJ9.5jDQBEY-KErkDK1QjDmv9ichOkNIn5CWYibe1Wz1644
icmje_1_orig.png
 
cc.logo.large.png
 
ncbi.png
 
google-scholar.pngpn6krf5.jpg
 


 

Our journal is in TR-Dizin, DRJI (Directory of Research Journals Indexing, General Impact Factor, Google Scholar, Researchgate, CrossRef (DOI), ROAD, ASOS Index, Turk Medline Index, Eurasian Scientific Journal Index (ESJI), and Turkiye Citation Index.

EBSCO, DOAJ, OAJI and ProQuest Index are in process of evaluation. 

 

Journal articles are evaluated as "Double-Blind Peer Review"