EconPapers    
Economics at your fingertips  
 

Hierarchical Cluster Analysis – Various Approaches to Data Preparation

Z. Pacáková and J. Poláčková

AGRIS on-line Papers in Economics and Informatics, 2013, vol. 05, issue 3, 11

Abstract: The article deals with two various approaches to data preparation to avoid multicollinearity. The aim of the article is to find similarities among the e-communication level of EU states using hierarchical cluster analysis. The original set of fourteen indicators was first reduced on the basis of correlation analysis while in case of high correlation indicator of higher variability was included in further analysis. Secondly the data were transformed using principal component analysis while the principal components are poorly correlated. For further analysis five principal components explaining about 92% of variance were selected. Hierarchical cluster analysis was performed both based on the reduced data set and the principal component scores. Both times three clusters were assumed following Pseudo t-Squared and Pseudo F Statistic, but the final clusters were not identical. An important characteristic to compare the two results found was to look at the proportion of variance accounted for by the clusters which was about ten percent higher for the principal component scores (57.8% compared to 47%). Therefore it can be stated, that in case of using principal component scores as an input variables for cluster analysis with explained proportion high enough (about 92% for in our analysis), the loss of information is lower compared to data reduction on the basis of correlation analysis.

Keywords: Research; and; Development/Tech; Change/Emerging; Technologies (search for similar items in EconPapers)
Date: 2013
References: View references in EconPapers View complete reference list from CitEc
Citations:

Downloads: (external link)
https://ageconsearch.umn.edu/record/157585/files/a ... kova_polackova-2.pdf (application/pdf)

Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.

Export reference: BibTeX RIS (EndNote, ProCite, RefMan) HTML/Text

Persistent link: https://EconPapers.repec.org/RePEc:ags:aolpei:157585

DOI: 10.22004/ag.econ.157585

Access Statistics for this article

More articles in AGRIS on-line Papers in Economics and Informatics from Czech University of Life Sciences Prague, Faculty of Economics and Management Contact information at EDIRC.
Bibliographic data for series maintained by AgEcon Search ().

 
Page updated 2025-03-19
Handle: RePEc:ags:aolpei:157585