Reproducible Machine Learning Methods for Lung Cancer Detection Using Computed Tomography Images: Algorithm Development and Validation

被引:37
作者
Yu, Kun-Hsing [1 ,2 ,3 ]
Lee, Tsung-Lu Michael [4 ]
Yen, Ming-Hsuan [5 ,6 ,7 ]
Kou, S. C. [2 ]
Rosen, Bruce [8 ,9 ]
Chiang, Jung-Hsien [7 ]
Kohane, Isaac S. [1 ,9 ]
机构
[1] Harvard Med Sch, Dept Biomed Informat, Boston, MA 02115 USA
[2] Harvard Univ, Dept Stat, Cambridge, MA 02138 USA
[3] Brigham & Womens Hosp, Dept Pathol, 75 Francis St, Boston, MA 02115 USA
[4] Kun Shan Univ, Dept Informat Engn, Tainan, Taiwan
[5] Natl Cheng Kung Univ, Grad Program Multimedia Syst & Intelligent Comp, Tainan, Taiwan
[6] Acad Sinica, Tainan, Taiwan
[7] Natl Cheng Kung Univ, Dept Comp Sci & Informat Engn, 1 Univ Rd, Tainan, Taiwan
[8] Massachusetts Gen Hosp, Dept Radiol, Athinoula A Martinos Ctr Biomed Imaging, Boston, MA USA
[9] Harvard Massachusetts Inst Technol, Div Hlth Sci & Technol, Boston, MA USA
基金
美国国家卫生研究院; 美国国家科学基金会;
关键词
computed tomography; spiral; lung cancer; machine learning; early detection of cancer; reproducibility of results; PULMONARY NODULES; AIDED DIAGNOSIS; CT; RADIOLOGISTS;
D O I
10.2196/16709
中图分类号
R19 [保健组织与事业(卫生事业管理)];
学科分类号
摘要
Background: Chest computed tomography (CT) is crucial for the detection of lung cancer, and many automated CT evaluation methods have been proposed. Due to the divergent software dependencies of the reported approaches, the developed methods are rarely compared or reproduced. Objective: The goal of the research was to generate reproducible machine learning modules for lung cancer detection and compare the approaches and performances of the award-winning algorithms developed in the Kaggle Data Science Bowl. Methods: We obtained the source codes of all award-winning solutions of the Kaggle Data Science Bowl Challenge, where participants developed automated CT evaluation methods to detect lung cancer (training set n=1397, public test set n=198, final test set n=506). The performance of the algorithms was evaluated by the log-loss function, and the Spearman correlation coefficient of the performance in the public and final test sets was computed. Results: Most solutions implemented distinct image preprocessing, segmentation, and classification modules. Variants of U-Net, VGGNet, and residual net were commonly used in nodule segmentation, and transfer learning was used in most of the classification algorithms. Substantial performance variations in the public and final test sets were observed (Spearman correlation coefficient = .39 among the top 10 teams). To ensure the reproducibility of results, we generated a Docker container for each of the top solutions. Conclusions: We compared the award-winning algorithms for lung cancer detection and generated reproducible Docker images for the top solutions. Although convolutional neural networks achieved decent accuracy, there is plenty of room for improvement regarding model generalizability.
引用
收藏
页数:11
相关论文
共 50 条
[31]   Detection of extremity chronic traumatic osteomyelitis by machine learning based on computed-tomography images A retrospective study [J].
Wu, Yifan ;
Lu, Xin ;
Hong, Jianqiao ;
Lin, Weijie ;
Chen, Shiming ;
Mou, Shenghong ;
Feng, Gang ;
Yan, Ruijian ;
Cheng, Zhiyuan .
MEDICINE, 2020, 99 (09)
[32]   Development of a deep learning model for early gastric cancer diagnosis using preoperative computed tomography images [J].
Gao, Zhihong ;
Yu, Zhuo ;
Zhang, Xiang ;
Chen, Chun ;
Pan, Zhifang ;
Chen, Xiaodong ;
Lin, Weihong ;
Chen, Jun ;
Zhuge, Qichuan ;
Shen, Xian .
FRONTIERS IN ONCOLOGY, 2023, 13
[33]   Radiomics and machine learning for osteoporosis detection using abdominal computed tomography: a retrospective multicenter study [J].
Liu, Zhai ;
Li, Yongjun ;
Zhang, Chenguang ;
Xu, Hui ;
Zhao, Junlu ;
Huang, Chencui ;
Chen, Xingzhi ;
Ren, Qingyun .
BMC MEDICAL IMAGING, 2025, 25 (01)
[34]   Machine Learning Detection and Characterization of Splenic Injuries on Abdominal Computed Tomography [J].
Hamghalam, Mohammad ;
Moreland, Robert ;
Gomez, David ;
Simpson, Amber ;
Lin, Hui Ming ;
Jandaghi, Ali Babaei ;
Tafur, Monica ;
Vlachou, Paraskevi A. ;
Wu, Matthew ;
Brassil, Michael ;
Crivellaro, Priscila ;
Mathur, Shobhit ;
Hosseinpour, Shahob ;
Colak, Errol .
CANADIAN ASSOCIATION OF RADIOLOGISTS JOURNAL-JOURNAL DE L ASSOCIATION CANADIENNE DES RADIOLOGISTES, 2024, 75 (03) :534-541
[35]   Development and validation of a nomogram model of lung metastasis in breast cancer based on machine learning algorithm and cytokines [J].
Li, Zhaoyi ;
Miao, Hao ;
Bao, Wei ;
Zhang, Lansheng .
BMC CANCER, 2025, 25 (01)
[36]   Transcriptome mapping of renal clear cell carcinoma revealed by machine learning algorithm based on enhanced computed tomography images [J].
Yang, Yu ;
Huang, Hang ;
Liang, Haote .
JOURNAL OF GENE MEDICINE, 2023, 25 (07)
[37]   Towards Improved Identification of Vertebral Fractures in Routine Computed Tomography (CT) Scans: Development and External Validation of a Machine Learning Algorithm [J].
Nicolaes, Joeri ;
Skjodt, Michael Kriegbaum ;
Raeymaeckers, Steven ;
Smith, Christopher Dyer ;
Abrahamsen, Bo ;
Fuerst, Thomas ;
Debois, Marc ;
Vandermeulen, Dirk ;
Libanati, Cesar .
JOURNAL OF BONE AND MINERAL RESEARCH, 2023, 38 (12) :1856-1866
[38]   HCI-Driven Machine Learning for Early Detection of Lung Cancer: An Ensemble Approach [J].
Sohaib, Muhammad .
UNIVERSAL ACCESS IN HUMAN-COMPUTER INTERACTION, PT I, UAHCI 2024, 2024, 14696 :311-325
[39]   Development and validation of machine learning algorithms for early detection of ankylosing spondylitis using magnetic resonance images [J].
Canayaz, Emre ;
Altikardes, Zehra Aysun ;
Unsal, Alparslan ;
Korkmaz, Hayriye ;
Gok, Mustafa .
TECHNOLOGY AND HEALTH CARE, 2025, 33 (03) :1182-1198
[40]   Estimation of Static Lung Volumes and Capacities From Spirometry Using Machine Learning: Algorithm Development and Validation [J].
Helgeson, Scott A. ;
Quicksall, Zachary S. ;
Johnson, Patrick W. ;
Lim, Kaiser G. ;
Carter, Rickey E. ;
Lee, Augustine S. .
JMIR AI, 2025, 4