K-means cluster analysis, Advanced Statistics

Assignment Help:

K-means cluster analysis is the method of cluster analysis in which from an initial partition of observations into K clusters, each observation in turn is analysed and reassigned, if suitable, to a different cluster in an attempt to optimize some predefined numerical criterion that measures in some sense the 'quality' of cluster solution. Several such clustering criteria have been suggested, but the most usually used arise from considering the features of the within groups, between groups and whole matrices of sums of squares and the cross products (W, B, T) which can be described for every partition of the observations into the particular number of groups. The two most ordinary of the clustering criteria developing from these matrices are given as follows

minimization of trace W

minimization of determinant W

The first of these has tendency to produce the 'spherical' clusters, the second to produce clusters that all have same shape, though this will not necessarily be spherical in shape. 

 


Related Discussions:- K-means cluster analysis

Define radical statistics group, Radical statistics group : The national ne...

Radical statistics group : The national network of the social scientists in United Kingdom committed to the critique of statistics as taken in use in the policy making procedure. T

Describe lorenz curve., Lorenz curve : Essentially the graphical representa...

Lorenz curve : Essentially the graphical representation of cumulative distribution of the variable, most often used for the income. If the risks of disease are not monotonically in

Chernoff''s faces, Chernoff's faces : A method or technique for representin...

Chernoff's faces : A method or technique for representing the multivariate data graphically. Each observation is represented by the computer-created face, the features of which are

Basic reproduction number, Basic reproduction number : A term used in the t...

Basic reproduction number : A term used in the theory of infectious diseases for the number of secondary cases which one case would generate in a completely susceptible population.

Frailty, A term usually used for unobserved individual heterogeneity. Such ...

A term usually used for unobserved individual heterogeneity. Such variation is of main concern in the medical statistics particularly in the analysis of the survival times where ha

Ordinal variable, Ordinal variable is a measurement which allows a sample ...

Ordinal variable is a measurement which allows a sample of the individuals to be ranked with respect to some characteristic but where differences at different points of the scale

Density estimation, Procedures for estimating the probability distributions...

Procedures for estimating the probability distributions without supposing any particular functional form. Constructing the histogram is perhaps the easiest example of such type of

Missing data - reasons for screening data, Missing Data - Reasons for scree...

Missing Data - Reasons for screening data In case of any missing data, the researcher needs to conduct tests to ascertain that the pattern of these missing cases is random.

Describe item-total correlation, Item-total correlation is an  extensively...

Item-total correlation is an  extensively used method for checking the homogeneity of the scale made up of number of items. It is simply the Pearson's product moment correlation c

Protopathic bias, Protopathic bias is the type of bias (also called as rev...

Protopathic bias is the type of bias (also called as reverse-causality) that is a consequence of differential misclassification of the exposure related to timing of occurrence. It

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd