Imbalanced Classification in Diabetics Using Ensembled Machine Learning  被引量:1

在线阅读下载全文

作  者:M.Sandeep Kumar Mohammad Zubair Khan Sukumar Rajendran Ayman Noor A.Stephen Dass J.Prabhu 

机构地区:[1]School of Information Technology and Engineering,Vellore Institute of Technology,Vellore,Tamil Nadu,632014,India [2]Department of Computer Science and Information,Taibah University,Medina,Saudi Arabia [3]College of Computer Science and Engineering,Taibah University,Medina,Saudi Arabia

出  处:《Computers, Materials & Continua》2022年第9期4397-4409,共13页计算机、材料和连续体(英文)

摘  要:Diabetics is one of the world’s most common diseases which are caused by continued high levels of blood sugar.The risk of diabetics can be lowered if the diabetic is found at the early stage.In recent days,several machine learning models were developed to predict the diabetic presence at an early stage.In this paper,we propose an embedded-based machine learning model that combines the split-vote method and instance duplication to leverage an imbalanced dataset called PIMA Indian to increase the prediction of diabetics.The proposed method uses both the concept of over-sampling and under-sampling along with model weighting to increase the performance of classification.Different measures such as Accuracy,Precision,Recall,and F1-Score are used to evaluate the model.The results we obtained using K-Nearest Neighbor(kNN),Naïve Bayes(NB),Support Vector Machines(SVM),Random Forest(RF),Logistic Regression(LR),and Decision Trees(DT)were 89.32%,91.44%,95.78%,89.3%,81.76%,and 80.38%respectively.The SVM model is more efficient than other models which are 21.38%more than exiting machine learning-based works.

关 键 词:Diabetics classification imbalanced data split-vote instance duplication 

分 类 号:R587.1[医药卫生—内分泌] TP181[医药卫生—内科学]

 

参考文献:

正在载入数据...

 

二级参考文献:

正在载入数据...

 

耦合文献:

正在载入数据...

 

引证文献:

正在载入数据...

 

二级引证文献:

正在载入数据...

 

同被引文献:

正在载入数据...

 

相关期刊文献:

正在载入数据...

相关的主题
相关的作者对象
相关的机构对象