K-means cluster analysis, Advanced Statistics

Assignment Help:

K-means cluster analysis is the method of cluster analysis in which from an initial partition of observations into K clusters, each observation in turn is analysed and reassigned, if suitable, to a different cluster in an attempt to optimize some predefined numerical criterion that measures in some sense the 'quality' of cluster solution. Several such clustering criteria have been suggested, but the most usually used arise from considering the features of the within groups, between groups and whole matrices of sums of squares and the cross products (W, B, T) which can be described for every partition of the observations into the particular number of groups. The two most ordinary of the clustering criteria developing from these matrices are given as follows

minimization of trace W

minimization of determinant W

The first of these has tendency to produce the 'spherical' clusters, the second to produce clusters that all have same shape, though this will not necessarily be spherical in shape. 

 


Related Discussions:- K-means cluster analysis

Pattern recognition, Pattern recognition is a term for a technology that r...

Pattern recognition is a term for a technology that recognizes and analyses patterns automatically by machine and which has been used successfully in many areas of application inc

Bimodal distribution, Bimodal distribution : The probability distribution, ...

Bimodal distribution : The probability distribution, or we can simply say the frequency distribution, with two modes. Figure 15 shows the example of each of them

Empirical likelihood, An approach of using the likelihood as the basis of e...

An approach of using the likelihood as the basis of estimation without the requirement to specify a parametric family for data. Empirical likelihood can be viewed as the example of

Cohort component method, Cohort component method : A broadly used method or...

Cohort component method : A broadly used method or technique of forecasting the age- and sex-speci?c population to the upcoming years, in which the initial population is strati?ed

Reliability theory, Reliability theory is the theory which attempts to det...

Reliability theory is the theory which attempts to determine the reliability of the complex system from knowledge of the reliabilities of the components. Interest might centre on

Maximum likelihood estimation, Maximum likelihood estimation is an estimat...

Maximum likelihood estimation is an estimation procedure involving maximization of the likelihood or the log-likelihood with respect to the parameters. Such type of estimators is

Determine the probablity, Dr. Stallter has been teaching basic statistics f...

Dr. Stallter has been teaching basic statistics for many years. She knows that 80% of the students will complete the assigned problems. She has also determined that among those who

Hot deck, Hot deck is a method broadly used in surveys for imputing the mi...

Hot deck is a method broadly used in surveys for imputing the missing values. In its easiest form the method includes sampling with replacement m values from the sample respondent

Correlation matrix, Correlation matrix : A square, symmetric matrix with th...

Correlation matrix : A square, symmetric matrix with the rows and columns corresponding to the variables, in which the non diagonal elements are correlations between the pairs of t

Homework help, Q1: The growth in bad debt expense for Aptara Pvt. Ltd. Comp...

Q1: The growth in bad debt expense for Aptara Pvt. Ltd. Company over the last 20 years is as follows. 1997 0.11 1998 0.09 1999 0.08 2000 0.08 2001 0.1 2002 0.11 2003 0.12 2004 0.1

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd