BMC Bioinformatics | |
A diabetes prediction model based on Boruta feature selection and ensemble learning | |
Research | |
Yinbo Xin1  Suli Li1  Hongfang Zhou2  | |
[1] School of Computer Science and Engineering, Xi’an University of Technology, 710048, Xi’an, China;School of Computer Science and Engineering, Xi’an University of Technology, 710048, Xi’an, China;Shaanxi Key Laboratory of Network Computing and Security Technology, 710048, Xi’an, China; | |
关键词: Diabetes detection; Machine learning; Boruta feature selection; K-Means++; Ensemble learning; | |
DOI : 10.1186/s12859-023-05300-5 | |
received in 2022-07-15, accepted in 2023-04-21, 发布年份 2023 | |
来源: Springer | |
【 摘 要 】
Background and objectiveAs a common chronic disease, diabetes is called the “second killer” among modern diseases. Currently, there is no medical cure for diabetes. We can only rely on medication for auxiliary treatment. However, many diabetic patients still die each year. In addition, a considerable number of people do not pay attention to their physical health or opt out of treatment due to lack of money, which eventually leads to various complications. Therefore, diagnosing diabetes at an early stage and intervening early is necessary; thus, developing an early detection method for diabetes is essential.MethodsIn this study, a diabetes prediction model based on Boruta feature selection and ensemble learning is proposed. The model contains the use of Boruta feature selection, the extraction of salient features from datasets, the use of the K-Means++ algorithm for unsupervised clustering of data and stacking of an ensemble learning method for classification. It has been validated on a diabetes dataset.ResultsThe experiments were performed on the PIMA Indian diabetes dataset. The model was evaluated by accuracy, precision and F1 index. The obtained results show that the accuracy rate of the model reaches 98% and achieves good results.ConclusionCompared with other diabetes prediction models, this model achieved better results, and the obtained results indicate that this model is superior to other models in diabetes prediction and has better performance.
【 授权许可】
CC BY
© The Author(s) 2023
【 预 览 】
Files | Size | Format | View |
---|---|---|---|
RO202309074533793ZK.pdf | 2169KB | download | |
MediaObjects/12888_2023_4875_MOESM3_ESM.docx | 13KB | Other | download |
Fig. 1 | 1254KB | Image | download |
Fig. 13 | 101KB | Image | download |
Fig. 14 | 95KB | Image | download |
Fig. 1 | 419KB | Image | download |
Fig. 16 | 93KB | Image | download |
Fig. 2 | 4157KB | Image | download |
40517_2023_252_Article_IEq29.gif | 1KB | Image | download |
40517_2023_252_Article_IEq31.gif | 1KB | Image | download |
Fig. 2 | 83KB | Image | download |
40517_2023_252_Article_IEq45.gif | 1KB | Image | download |
40517_2023_252_Article_IEq47.gif | 1KB | Image | download |
40517_2023_252_Article_IEq55.gif | 1KB | Image | download |
Fig. 1 | 60KB | Image | download |
MediaObjects/13690_2023_1109_MOESM1_ESM.docx | 14KB | Other | download |
Fig. 2 | 1016KB | Image | download |
Fig. 1 | 203KB | Image | download |
Fig. 1 | 66KB | Image | download |
40517_2023_252_Article_IEq62.gif | 1KB | Image | download |
Fig. 2 | 235KB | Image | download |
MediaObjects/12888_2023_4862_MOESM1_ESM.docx | 10KB | Other | download |
MediaObjects/13690_2023_1108_MOESM1_ESM.docx | 64KB | Other | download |
Fig. 3 | 222KB | Image | download |
MediaObjects/13690_2023_1108_MOESM2_ESM.docx | 32KB | Other | download |
Fig. 1 | 121KB | Image | download |
Fig. 4 | 189KB | Image | download |
Fig. 2 | 663KB | Image | download |
40517_2023_252_Article_IEq69.gif | 1KB | Image | download |
MediaObjects/12888_2023_4843_MOESM1_ESM.docx | 25KB | Other | download |
MediaObjects/12888_2023_4843_MOESM2_ESM.docx | 26KB | Other | download |
MediaObjects/12888_2023_4843_MOESM3_ESM.docx | 165KB | Other | download |
40517_2023_252_Article_IEq73.gif | 1KB | Image | download |
40517_2023_252_Article_IEq74.gif | 1KB | Image | download |
Fig. 3 | 212KB | Image | download |
40517_2023_252_Article_IEq76.gif | 1KB | Image | download |
40517_2023_252_Article_IEq77.gif | 1KB | Image | download |
40517_2023_252_Article_IEq78.gif | 1KB | Image | download |
40517_2023_252_Article_IEq79.gif | 1KB | Image | download |
【 图 表 】
40517_2023_252_Article_IEq79.gif
40517_2023_252_Article_IEq78.gif
40517_2023_252_Article_IEq77.gif
40517_2023_252_Article_IEq76.gif
Fig. 3
40517_2023_252_Article_IEq74.gif
40517_2023_252_Article_IEq73.gif
40517_2023_252_Article_IEq69.gif
Fig. 2
Fig. 4
Fig. 1
Fig. 3
Fig. 2
40517_2023_252_Article_IEq62.gif
Fig. 1
Fig. 1
Fig. 2
Fig. 1
40517_2023_252_Article_IEq55.gif
40517_2023_252_Article_IEq47.gif
40517_2023_252_Article_IEq45.gif
Fig. 2
40517_2023_252_Article_IEq31.gif
40517_2023_252_Article_IEq29.gif
Fig. 2
Fig. 16
Fig. 1
Fig. 14
Fig. 13
Fig. 1
【 参考文献 】
- [1]
- [2]
- [3]
- [4]
- [5]
- [6]
- [7]
- [8]
- [9]
- [10]
- [11]
- [12]
- [13]
- [14]
- [15]
- [16]
- [17]
- [18]
- [19]
- [20]
- [21]
- [22]
- [23]
- [24]
- [25]
- [26]
- [27]
- [28]
- [29]
- [30]
- [31]
- [32]
- [33]
- [34]
- [35]
- [36]
- [37]
- [38]
- [39]
- [40]
- [41]