K-means cluster analysis, Advanced Statistics

Assignment Help:

K-means cluster analysis is the method of cluster analysis in which from an initial partition of observations into K clusters, each observation in turn is analysed and reassigned, if suitable, to a different cluster in an attempt to optimize some predefined numerical criterion that measures in some sense the 'quality' of cluster solution. Several such clustering criteria have been suggested, but the most usually used arise from considering the features of the within groups, between groups and whole matrices of sums of squares and the cross products (W, B, T) which can be described for every partition of the observations into the particular number of groups. The two most ordinary of the clustering criteria developing from these matrices are given as follows

minimization of trace W

minimization of determinant W

The first of these has tendency to produce the 'spherical' clusters, the second to produce clusters that all have same shape, though this will not necessarily be spherical in shape. 

 


Related Discussions:- K-means cluster analysis

Composite sampling, A procedure whereby the collection of multiple sample u...

A procedure whereby the collection of multiple sample units are combined in their entirety or in part, to form the new sample. One or more succeeding measurements are taken on the

Friedman''s two-way analysis of variance, The distribution free or techniqu...

The distribution free or technique which is the analogue of the analysis of variance for the design with two factors. It can be applied to data sets which do not meet the assumptio

Mba, Mention the characteristics of Statistics. Explain any two application...

Mention the characteristics of Statistics. Explain any two applications of Statistics.

Codominance, Codominance : The relationship between genotype at the locus a...

Codominance : The relationship between genotype at the locus and a phenotype to which it in?uences. If an individuals with heterozygote (such as, AB) genotype is phenotypically dif

Explain longitudinal data, Longitudinal data : The data arising when each o...

Longitudinal data : The data arising when each of the number of subjects or patients give rise to the vector of measurements representing same variable observed at the number of di

Catastrophe theory, Catastrophe theory : A theory of how little is the cont...

Catastrophe theory : A theory of how little is the continuous changes in the independent variables which can have unexpected, discontinuous effects on the dependent variables. Exam

Excel, Software which started out as the spreadsheet targeting at manipulat...

Software which started out as the spreadsheet targeting at manipulating the tables of number for financial analysis, which has now developed into a more flexible package for workin

Intention-to-treat analysis, Intention-to-treat analysis is the process in...

Intention-to-treat analysis is the process in which all the patients randomly allocated to a treatment in the clinical trial are analyzed together as representing that particular

Likelihood, Likelihood is the probability of a set of observations provide...

Likelihood is the probability of a set of observations provided the value of some parameter or the set of parameters. For instance, the likelihood of the random sample of n observ

Cross over design, The type of longitudinal study in which the subjects rec...

The type of longitudinal study in which the subjects receive different treatments on the various occasions. Random allocation is required to determine the order in which the treatm

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd