An Effective Heart Disease Prediction Model based on Machine Learning Techniques

Rony Chowdhury Ripan; Iqbal H. Sarker; Md. Hasan Furhad; Md Musfique Anwar; Mohammed Moshiul Hoque

doi:10.20944/preprints202011.0744.v1

Submitted:

27 November 2020

Posted:

30 November 2020

You are already at the latest version

Abstract

This paper presents an effective heart disease prediction model through detecting the anomalies, also known as outliers, in healthcare data using the unsupervised K-means clustering algorithm. Most existing approaches for detecting anomalies are based on constructing profiles of normal instances. However, such techniques require an adequate number of normal profiles to justify those models. Our proposed model first evaluates an \textit{optimal} value of K using Silhouette method. Next, it intends to locate anomalies that are far from a certain threshold distance with respect to their clusters. Finally, the five most popular classification techniques such as K-Nearest Neighbor (KNN), Random Forest (RF), Support Vector Machines (SVM), Naive Bayes (NB), and Logistic Regression (LR) are applied to build the resultant prediction model. The effectiveness of the proposed methodology is justified using a benchmark dataset of heart disease.

Keywords:

Anomaly detection

;

Healthcare

;

K-means clustering

;

Heart disease prediction

;

Data analytics

;

Machine learning

Subject:

Computer Science and Mathematics - Algebra and Number Theory

Copyright: This open access article is published under a Creative Commons CC BY 4.0 license, which permit the free download, distribution, and reuse, provided that the author and preprint are cited in any reuse.

An Effective Heart Disease Prediction Model based on Machine Learning Techniques

Abstract

Keywords:

Subject:

MDPI Initiatives

Important Links

Subscribe