Utilizing Attention-Enhanced Deep Neural Networks for Large-Scale Preliminary Diabetes Screening in Population Health Data

被引:0
作者
Hu, Hongwei [1 ,2 ]
Dong, Wenbo [3 ]
Yu, Jianming [1 ]
Guan, Shiyan [1 ]
Zhu, Xiaofei [1 ]
机构
[1] Elect Equipment Mfg Engn Technol Res & Dev Ctr Jia, Huaian 223003, Peoples R China
[2] Jiangsu Vocat Coll Elect & Informat, Huaian 223003, Peoples R China
[3] Univ Hong Kong, Dept Elect & Elect Engn, Hong Kong, Peoples R China
关键词
diabetes prediction; attention-enhanced deep neural network (AEDNN); attention-based feature weighting layer; Pima Indians diabetes dataset;
D O I
10.3390/electronics13214177
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Early screening for diabetes can promptly identify potential early stage patients, possibly delaying complications and reducing mortality rates. This paper presents a novel technique for early diabetes screening and prediction, called the Attention-Enhanced Deep Neural Network (AEDNN). The proposed AEDNN model incorporates an Attention-based Feature Weighting Layer combined with deep neural network layers to achieve precise diabetes prediction. In this study, we utilized the Diabetes-NHANES dataset and the Pima Indians Diabetes dataset. To handle significant missing values and outliers, group median imputation was applied. Oversampling techniques were used to balance the diabetes and non-diabetes groups. The data were processed through an Attention-based Feature Weighting Layer for feature extraction, producing a feature matrix. This matrix was subjected to Hadamard product operations with the raw data to obtain weighted data, which were subsequently input into deep neural network layers for training. The parameters were fine-tuned and the L2 regularization and dropout layers were added to enhance the generalization performance of the model. The model's reliability was thoroughly assessed through various metrics, including the accuracy, precision, recall, F1 score, mean squared error (MSE), and R2 score, as well as the ROC and AUC curves. The proposed model achieved a prediction accuracy of 98.4% in the Pima Indians Diabetes dataset. When the test dataset was expanded to the large-scale Diabetes-NHANES dataset, which contains 52,390 samples, the test precision of the model improved further to 99.82%, with an AUC of 0.9995. A comparative analysis was conducted using multiple models, including logistic regression with L1 regularization, support vector machine (SVM), random forest, K-nearest neighbors (KNNs), AdaBoost, XGBoost, and the latest semi-supervised XGBoost. The feature extraction method using attention mechanisms was compared with the classical feature selection methods, Lasso and Ridge. The experiments were performed on the same dataset, and the conclusion was that the Attention-based Ensemble Deep Neural Network (AEDNN) outperformed all the aforementioned methods. These results indicate that the model not only performs well on smaller datasets but also fully leverages its advantages on larger datasets, demonstrating strong generalization ability and robustness. The proposed model can effectively assist clinicians in the early screening of diabetes patients. This is particularly beneficial for the preliminary screening of high-risk individuals in large-scale, extensive healthcare datasets, followed by detailed examination and diagnosis. Compared to the existing methods, our AEDNN model showed an overall performance improvement of 1.75%.
引用
收藏
页数:31
相关论文
共 25 条
[1]  
Ahmed Benzir Md, 2024, Smart Health, V32, DOI 10.1016/j.smhl.2024.100457
[2]   High-Risk Prediction of Cardiovascular Diseases via Attention-Based Deep Neural Networks [J].
An, Ying ;
Huang, Nengjun ;
Chen, Xianlai ;
Wu, FangXiang ;
Wang, Jianxin .
IEEE-ACM TRANSACTIONS ON COMPUTATIONAL BIOLOGY AND BIOINFORMATICS, 2021, 18 (03) :1093-1105
[3]   A Novel Proposal for Deep Learning-Based Diabetes Prediction: Converting Clinical Data to Image Data [J].
Aslan, Muhammet Fatih ;
Sabanci, Kadir .
DIAGNOSTICS, 2023, 13 (04)
[4]   A position statement on screening and management of prediabetes in adults in primary care in Australia q [J].
Bell, Kirstine ;
Shaw, Jonathan E. ;
Maple-Brown, Louise ;
Ferris, Wendy ;
Gray, Susan ;
Murfet, Giuliana ;
Flavel, Rebecca ;
Maynard, Bernie ;
Ryrie, Hannah ;
Pritchard, Barry ;
Freeman, Rachel ;
Gordon, Brett A. .
DIABETES RESEARCH AND CLINICAL PRACTICE, 2020, 164
[5]   A deep neural network with modified random forest incremental interpretation approach for diagnosing diabetes in smart healthcare [J].
Chen, Tin-Chih Toly ;
Wu, Hsin-Chieh ;
Chiu, Min-Chi .
APPLIED SOFT COMPUTING, 2024, 152
[6]   Artificial intelligence of medical things for disease detection using ensemble deep learning and attention mechanism [J].
Djenouri, Youcef ;
Belhadi, Asma ;
Yazidi, Anis ;
Srivastava, Gautam ;
Lin, Jerry Chun-Wei .
EXPERT SYSTEMS, 2024, 41 (06)
[7]   Pediatric diabetes prediction using deep learning [J].
El-Bashbishy, Abeer El-Sayyid ;
El-Bakry, Hazem M. .
SCIENTIFIC REPORTS, 2024, 14 (01)
[8]   A Proposed Technique Using Machine Learning for the Prediction of Diabetes Disease through a Mobile App [J].
El-Sofany, Hosam ;
El-Seoud, Samir A. ;
Karam, Omar H. ;
Abd El-Latif, Yasser M. ;
Taj-Eddin, Islam A. T. F. .
INTERNATIONAL JOURNAL OF INTELLIGENT SYSTEMS, 2024, 2024
[9]  
Fernndez A., 2018, Pattern Recognit, V91, P313
[10]   Diabetes detection using deep learning techniques with oversampling and feature augmentation [J].
Garcia-Ordas, Maria Teresa ;
Benavides, Carmen ;
Benitez-Andrades, Jose Alberto ;
Alaiz-Moreton, Hector ;
Garcia-Rodriguez, Isaias .
COMPUTER METHODS AND PROGRAMS IN BIOMEDICINE, 2021, 202