Analyzing the Efficacy of K-Means Clustering and Logistic Regression For Diabetes Prediction

Sohan Lal Gupta,Vikram Khandelwal,Vinod Katria,Dr Arpita Sharma,Anjali Pandey

doi:10.70135/seejph.vi.2454

Abstract

Diabetes causes a large number of deaths each year and a large number of people living with the disease do not realize their health condition early enough. In this study, we propose a data mining based model for early diagnosis and prediction of diabetes using the UCI database. Although K-means is simple and can be used for a wide variety of data types, it is quite sensitive to initial positions of cluster centers which determine the final cluster result, which either provides a sufficient and efficiently clustered dataset for the logistic regression model, or gives a lesser amount of data as a result of incorrect clustering of the original dataset, thereby limiting the performance of the logistic regression model. Our findings offer insights into the comparative strengths and weaknesses of each method, shedding light on their potential applications in diabetes diagnosis and risk assessment. A further experiment with a new dataset showed the applicability of our model for the predication of diabetes.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Analyzing the Efficacy of K-Means Clustering and Logistic Regression For Diabetes Prediction

Abstract

Talk to us

Similar Papers

More From: South Eastern European Journal of Public Health

Lead the way for us

Journal: South Eastern European Journal of Public Health	Publication Date: Nov 28, 2024
License type: CC BY-ND 4.0

Similar Papers

Improved logistic regression model for diabetes prediction by integrating PCA and K-means techniques
Changsheng Zhu ... Christian Uwa Idemudia
Informatics in Medicine Unlocked | VOL. 17
Changsheng Zhu, et. al.Changsheng Zhu ... Christian Uwa Idemudia
01 Jan 2019
Informatics in Medicine Unlocked | VOL. 17

Toward a Modern Era in Clinical Prediction: The TRIPOD Statement for Reporting Prediction Models
Navdeep Tangri ... David M Kent
American Journal of Kidney Diseases | VOL. 65
Navdeep Tangri, et. al.Navdeep Tangri ... David M Kent
15 Jan 2015
American Journal of Kidney Diseases | VOL. 65

U-shaped association between serum IGF2BP3 and T2DM: A cross-sectional study in Chinese population.
Xiaoying Wu ... Juying Tang
Journal of Diabetes | VOL. 15
Xiaoying Wu, et. al.Xiaoying Wu ... Juying Tang
09 Mar 2023
Journal of Diabetes | VOL. 15

Artificial Neural Network Models for Prediction of Acute Coronary Syndromes Using Clinical Data From the Time of Presentation
Robert F Harrison ... R Lee Kennedy
Annals of Emergency Medicine | VOL. 46
Robert F Harrison, et. al.Robert F Harrison ... R Lee Kennedy
27 Apr 2005
Annals of Emergency Medicine | VOL. 46

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Analyzing the Efficacy of K-Means Clustering and Logistic Regression For Diabetes Prediction

Abstract

Talk to us

Similar Papers

More From: South Eastern European Journal of Public Health