Full Text

Turn on search term navigation

© 2022 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (https://creativecommons.org/licenses/by/4.0/). Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License.

Abstract

The prevalence of diabetes has been increasing in recent years, and previous research has found that machine-learning models are good diabetes prediction tools. The purpose of this study was to compare the efficacy of five different machine-learning models for diabetes prediction using lifestyle data from the National Health and Nutrition Examination Survey (NHANES) database. The 1999–2020 NHANES database yielded data on 17,833 individuals data based on demographic characteristics and lifestyle-related variables. To screen training data for machine models, the Akaike Information Criterion (AIC) forward propagation algorithm was utilized. For predicting diabetes, five machine-learning models (CATBoost, XGBoost, Random Forest (RF), Logistic Regression (LR), and Support Vector Machine (SVM)) were developed. Model performance was evaluated using accuracy, sensitivity, specificity, precision, F1 score, and receiver operating characteristic (ROC) curve. Among the five machine-learning models, the dietary intake levels of energy, carbohydrate, and fat, contributed the most to the prediction of diabetes patients. In terms of model performance, CATBoost ranks higher than RF, LG, XGBoost, and SVM. The best-performing machine-learning model among the five is CATBoost, which achieves an accuracy of 82.1% and an AUC of 0.83. Machine-learning models based on NHANES data can assist medical institutions in identifying diabetes patients.

Details

Title
Machine Learning Models for Data-Driven Prediction of Diabetes by Lifestyle Type
Author
Qin, Yifan 1 ; Wu, Jinlong 2   VIAFID ORCID Logo  ; Xiao, Wen 1 ; Wang, Kun 3 ; Huang, Anbing 1 ; Bowen, Liu 1 ; Yu, Jingxuan 1 ; Li, Chuhao 1 ; Yu, Fengyu 1 ; Ren, Zhanbing 1   VIAFID ORCID Logo 

 College of Physical Education, Shenzhen University, Shenzhen 518000, China 
 College of Physical Education, Southwest University, Chongqing 400715, China 
 Physical Education College, Yanching Institute of Technology, Langfang 065201, China 
First page
15027
Publication year
2022
Publication date
2022
Publisher
MDPI AG
ISSN
1661-7827
e-ISSN
1660-4601
Source type
Scholarly Journal
Language of publication
English
ProQuest document ID
2739429000
Copyright
© 2022 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (https://creativecommons.org/licenses/by/4.0/). Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License.