PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 24, 20241 citations

Analysis of various data imputation techniques for diabetes classification on PIMA dataset

View Full Paper
VJVishesh JainSSSanyam ShuklaNKNilay Khare

Key Points

Key points are not available for this paper at this time.

Abstract

Methodologies for addressing missing data in classification tasks must be rigorously evaluated in light of the rapidly expanding field of healthcare informatics. Using the PIMA Indian Diabetes dataset, this research provides a thorough analysis of data imputation methods related to diabetes classification. We evaluate four popular imputation techniques: Multivariate Imputation by Chained Equations (MICE), k-Nearest Neighbours (KNN), Mean, and Median. These techniques are applied to a variety of machine learning classifiers including Decision Trees, Random Forest, Support Vector Classifier (SVC), and Gaussian Naive Bayes Classifier. Our objective is to provide an understanding of how these techniques influence the predictive accuracy of classifiers in the context of diabetes diagnosis.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Jain et al. (2024) studied this question.

synapsesocial.com/papers/68e77c7cb6db6435876f066ehttps://doi.org/10.1109/sceecs61402.2024.10482050
Ask AI
Helpful
Bookmark
Share
View Full Paper