Predicting Diabetes Mellitus With Machine Learning Techniques

被引:365
|
作者
Zou, Quan [1 ,2 ]
Qu, Kaiyang [1 ]
Luo, Yamei [3 ]
Yin, Dehui [3 ]
Ju, Ying [4 ]
Tang, Hua [5 ]
机构
[1] Tianjin Univ, Sch Comp Sci & Technol, Tianjin, Peoples R China
[2] Univ Elect Sci & Technol China, Inst Fundamental & Frontier Sci, Chengdu, Sichuan, Peoples R China
[3] Southwest Med Univ, Sch Med Informat & Engn, Luzhou, Peoples R China
[4] Xiamen Univ, Sch Informat Sci & Technol, Xiamen, Peoples R China
[5] Southwest Med Univ, Sch Basic Med, Dept Pathophysiol, Luzhou, Peoples R China
关键词
diabetes mellitus; random forest; decision tree; neural network; machine learning; feature ranking; RANDOM FOREST; FEATURE-SELECTION; DIAGNOSIS; CLASSIFICATION; EXTRACTION; TOOL;
D O I
10.3389/fgene.2018.00515
中图分类号
Q3 [遗传学];
学科分类号
071007 ; 090102 ;
摘要
Diabetes mellitus is a chronic disease characterized by hyperglycemia. It may cause many complications. According to the growing morbidity in recent years, in 2040, the world's diabetic patients will reach 642 million, which means that one of the ten adults in the future is suffering from diabetes. There is no doubt that this alarming figure needs great attention. With the rapid development of machine learning, machine learning has been applied to many aspects of medical health. In this study, we used decision tree, random forest and neural network to predict diabetes mellitus. The dataset is the hospital physical examination data in Luzhou, China. It contains 14 attributes. In this study, five-fold cross validation was used to examine the models. In order to verity the universal applicability of the methods, we chose some methods that have the better performance to conduct independent test experiments. We randomly selected 68994 healthy people and diabetic patients' data, respectively as training set. Due to the data unbalance, we randomly extracted 5 times data. And the result is the average of these five experiments. In this study, we used principal component analysis (PCA) and minimum redundancy maximum relevance (mRMR) to reduce the dimensionality. The results showed that prediction with random forest could reach the highest accuracy (ACC = 0.8084) when all the attributes were used.
引用
收藏
页数:10
相关论文
共 50 条
  • [31] Glycemic and lipid variability for predicting complications and mortality in diabetes mellitus using machine learning
    Lee, Sharen
    Zhou, Jiandong
    Wong, Wing Tak
    Liu, Tong
    Wu, William K. K.
    Wong, Ian Chi Kei
    Zhang, Qingpeng
    Tse, Gary
    BMC ENDOCRINE DISORDERS, 2021, 21 (01)
  • [32] Non-invasive Diabetes Mellitus Detection System using Machine Learning Techniques
    Prabha, Anju
    Yadav, Jyoti
    Rani, Asha
    Singh, Vijander
    2021 11TH INTERNATIONAL CONFERENCE ON CLOUD COMPUTING, DATA SCIENCE & ENGINEERING (CONFLUENCE 2021), 2021, : 948 - 953
  • [33] Metabolic Syndrome and Development of Diabetes Mellitus: Predictive Modeling Based on Machine Learning Techniques
    Perveen, Sajida
    Shahbaz, Muhammad
    Keshavjee, Karim
    Guergachi, Aziz
    IEEE ACCESS, 2019, 7 : 1365 - 1375
  • [34] Integrated Embedded system for detecting diabetes mellitus using various machine learning techniques
    Konda R.
    Ramineni A.
    Jayashree J.
    Singavajhala N.
    Vanka S.A.
    EAI Endorsed Transactions on Pervasive Health and Technology, 2024, 10
  • [35] Predicting IRI Using Machine Learning Techniques
    Sharma, Ankit
    Sachdeva, S. N.
    Aggarwal, Praveen
    INTERNATIONAL JOURNAL OF PAVEMENT RESEARCH AND TECHNOLOGY, 2023, 16 (01) : 128 - 137
  • [36] Predicting IRI Using Machine Learning Techniques
    Ankit Sharma
    S. N. Sachdeva
    Praveen Aggarwal
    International Journal of Pavement Research and Technology, 2023, 16 : 128 - 137
  • [37] Machine Learning Techniques for Predicting Heart Diseases
    Taha, Mohammed A.
    Alsaidi, Saif Ali Abd Alradha
    Hussein, Reem Ali
    2022 INTERNATIONAL SYMPOSIUM ON INNOVATIVE INFORMATICS OF BISKRA, ISNIB, 2022, : 123 - 128
  • [38] Utility of Big Data in Predicting Short-Term Blood Glucose Levels in Type 1 Diabetes Mellitus Through Machine Learning Techniques
    Rodriguez-Rodriguez, Ignacio
    Chatzigiannakis, Ioannis
    Rodriguez, Jose-Victor
    Maranghi, Marianna
    Gentili, Michele
    Zamora-Izquierdo, Miguel-Angel
    SENSORS, 2019, 19 (20)
  • [39] A STUDY UTILIZING ADVANCED MACHINE LEARNING TECHNIQUES TO ANALYZE GESTATIONAL DIABETES MELLITUS AND ITS IMPLEMENTATIONS
    Shanmugam, Velu Chinnasamy
    Vijayalakshmi, C.
    Mynarani, M.
    Veerakumari, K. Pradeepa
    JP JOURNAL OF BIOSTATISTICS, 2024, 24 (02) : 227 - 237
  • [40] Prediction of gestational diabetes mellitus in the first 19 weeks of pregnancy using machine learning techniques
    Xiong, Yan
    Lin, Lu
    Chen, Yu
    Salerno, Stephen
    Li, Yi
    Zeng, Xiaoxi
    Li, Huafeng
    JOURNAL OF MATERNAL-FETAL & NEONATAL MEDICINE, 2022, 35 (13): : 2457 - 2463