Data mining classification techniques - comparison for better accuracy in prediction of cardiovascular disease
Richa Sharma,
Shailendra Narayan Singh and
Sujata Khatri
International Journal of Data Analysis Techniques and Strategies, 2019, vol. 11, issue 4, 356-373
Abstract:
Cardiovascular disease is a broad term which includes stroke or any disorder in the cardiovascular system that has the heart at its centre. This disease is a critical cause of mortality every year across the globe. Data mining utilises a variety of techniques and algorithms that could help to draw some interesting conclusions about cardiovascular disease. Data mining in healthcare can assist in predicting disease. This study aims to gain knowledge from a heart disease dataset and analyse several data mining classification techniques seeking improved accuracy and a lesser error rate in the results. The data set for the experiment is chosen from the UCI machine learning repository database. The dataset is analysed using two different data mining tools, i.e., WEKA and Tanagra. The analysis was done using the 10 fold cross validation technique. The results show that the Naive Bayes algorithm and the C-PLS algorithm outperform others with an accuracy of 83.71% and 84.44% respectively.
Keywords: data mining; classification techniques; machine learning tools; cardiovascular disease; KNN; Naïve Bayes; C-PLS; decision tree. (search for similar items in EconPapers)
Date: 2019
References: Add references at CitEc
Citations:
Downloads: (external link)
http://www.inderscience.com/link.php?id=103756 (text/html)
Access to full text is restricted to subscribers.
Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.
Export reference: BibTeX
RIS (EndNote, ProCite, RefMan)
HTML/Text
Persistent link: https://EconPapers.repec.org/RePEc:ids:injdan:v:11:y:2019:i:4:p:356-373
Access Statistics for this article
More articles in International Journal of Data Analysis Techniques and Strategies from Inderscience Enterprises Ltd
Bibliographic data for series maintained by Sarah Parker ().