Analyze Important Features of PIMA Indian Database For Diabetes Prediction Using KNN

Aziz Perdana; Arief Hermawan; Donny Avianto

doi:10.32736/sisfokom.v12i1.1598

Jurnal Sisfokom (Mar 2023)

Analyze Important Features of PIMA Indian Database For Diabetes Prediction Using KNN

Aziz Perdana,
Arief Hermawan,
Donny Avianto

Affiliations

Aziz Perdana: Universitas Teknologi Yogyakarta
Arief Hermawan: Universitas Teknologi Yogyakarta
Donny Avianto: Universitas Teknologi Yogyakarta

DOI: https://doi.org/10.32736/sisfokom.v12i1.1598
Journal volume & issue: Vol. 12, no. 1
pp. 70 – 75

Abstract

Read online

Diabetes is a chronic, non-communicable disease, and a long-term health condition that affects how the body uses glucose, the type of sugar that gives energy. In Indonesia, diabetes ranks as the sixth highest cause of death, following conditions related to childbirth. In 2021, Indonesia has a total of 19.5 million diabetes patients, making it the fifth-highest in the world. Some machine learning research has used data from the PIDD (PIMA Indian Diabetes Dataset) to predict diabetes. In this research, in addition to prediction accuracy, data complexity is also important. This research analyzes important features in the PIMA Indian database using the KNN (k-nearest neighbor) method for classification. The results show that using KNN with k=22 value results in the highest accuracy of 83.12%. The analysis also found that the important features required by the KNN method to achieve high accuracy from the PIMA Indian database, in order of importance, are glucose, age, insulin, blood pressure, Body Mass Index, pregnancy, skin thickness, and diabetes pedigree function. However, when used in the KNN classification method, the diabetes pedigree function feature was found to be unnecessary, not relevant, and can be reduced.

Published in Jurnal Sisfokom

ISSN: 2301-7988 (Print); 2581-0588 (Online)
Publisher: LPPM ISB Atma Luhur
Country of publisher: Indonesia
LCC subjects: Technology: Technology (General): Industrial engineering. Management engineering: Information technology
Website: http://jurnal.atmaluhur.ac.id/index.php/sisfokom/

About the journal

Abstract

Keywords