Analyzing the Efficacy of K-Means Clustering and Logistic Regression For Diabetes Prediction

Sohan Lal Gupta; Vikram Khandelwal; Vinod Katria; Dr. Arpita Sharma; Anjali Pandey

doi:10.70135/seejph.vi.2454

Analyzing the Efficacy of K-Means Clustering and Logistic Regression For Diabetes Prediction

Authors

Sohan Lal Gupta
Vikram Khandelwal
Vinod Katria
Dr. Arpita Sharma
Anjali Pandey

DOI:

https://doi.org/10.70135/seejph.vi.2454

Keywords:

Artificial Neural Network, Machine Learning, Support Vector Machine, knowledge discovery in databases, Online Analytical Processing..

Abstract

Diabetes causes a large number of deaths each year and a large number of people living with the disease do not realize their health condition early enough. In this study, we propose a data mining based model for early diagnosis and prediction of diabetes using the UCI database. Although K-means is simple and can be used for a wide variety of data types, it is quite sensitive to initial positions of cluster centers which determine the final cluster result, which either provides a sufficient and efficiently clustered dataset for the logistic regression model, or gives a lesser amount of data as a result of incorrect clustering of the original dataset, thereby limiting the performance of the logistic regression model. Our findings offer insights into the comparative strengths and weaknesses of each method, shedding light on their potential applications in diabetes diagnosis and risk assessment. A further experiment with a new dataset showed the applicability of our model for the predication of diabetes.

Downloads

Published

2024-11-28

How to Cite

Gupta, S. L., Khandelwal, V., Katria, V., Sharma, D. A., & Pandey, A. (2024). Analyzing the Efficacy of K-Means Clustering and Logistic Regression For Diabetes Prediction. South Eastern European Journal of Public Health, 1255–1262. https://doi.org/10.70135/seejph.vi.2454

Download Citation

Issue

Volume XXV 2024

Section

Articles

License

This work is licensed under a Creative Commons Attribution-NoDerivatives 4.0 International License.

Analyzing the Efficacy of K-Means Clustering and Logistic Regression For Diabetes Prediction

Authors

DOI:

Keywords:

Abstract

Downloads

Published

How to Cite

Issue

Section

License

Announcements

Call for Papers

indexing

Make a Submission

sidebar1

Benefits of Publishing Open Access

sidebar2

Public Health in Europe

sidebar3

Downloads